Michael Wagner 0019

dblp:21/4084-19 · DBLP profile ↗
← Back
5ranked-venue papers
2as first author
3since 2021 · last 2022
—ORCID · conflict

Domains — the database's venue-derived domains; a paper can count in several

Artificial intelligence and machine learning · 5 · 2 first-author · 3 since 2021Graphics, computer vision, multimedia, augmented reality and games · 4 · 2 first-author · 2 since 2021

Expertise — from the expertise taxonomy: the topics of the expert's papers under the CCF categories. A weight counts papers with recency: 1 for a paper about the topic, 0.3 when the topic is its context, halved every five years.

Artificial intelligence
1 paper
Information extraction and text analysis · 77% Language models and text generation · 23%

Topics — the 1 heaviest of 2, each with the papers that count most for it

TopicWeightPapersLastEvidence papers
Natural language and speech › Language models and text generation › pre-trained language model
pretrained language model analysis
0.212022
Characterizing Idioms: Conventionality and Contingency · ACL (1) 2022

Methods — techniques the papers use, named apart from their topics

XLNet · 0.6BERT · 0.6
YearPublicationVenuePosition
2022 Characterizing Idioms: Conventionality and Contingency
abstract
Idioms are unlike most phrases in two important ways.First, words in an idiom have non-canonical meanings.Second, the noncanonical meanings of words in an idiom are contingent on the presence of other words in the idiom.Linguistic theories differ on whether these properties depend on one another, as well as whether special theoretical machinery is needed to accommodate idioms.We define two measures that correspond to the properties above, and we implement them using BERT (Devlin et al., 2019) and XLNet (Yang et al., 2019).We show that English idioms fall at the expected intersection of the two dimensions, but that the dimensions themselves are not correlated.Our results suggest that special machinery to handle idioms may not be warranted.
Michaela Socolof, Jackie Chi Kit Cheung, Michael Wagner 0019, Timothy J. O'Donnell
ACL (1)3
2021 ProsoBeast Prosody Annotation Tool
abstract
The labelling of speech corpora is a laborious and time-consuming process. The ProsoBeast Annotation Tool seeks to ease and accelerate this process by providing an interactive 2D representation of the prosodic landscape of the data, in which contours are distributed based on their similarity. This interactive map allows the user to inspect and label the utterances. The tool integrates several state-of-the-art methods for dimensionality reduction and feature embedding, including variational autoencoders. The user can use these to find a good representation for their data. In addition, as most of these methods are stochastic, each can be used to generate an unlimited number of different prosodic maps. The web app then allows the user to seamlessly switch between these alternative representations in the annotation process. Experiments with a sample prosodically rich dataset have shown that the tool manages to find good representations of varied data and is helpful both for annotation and label correction. The tool is released as free software for use by the community.
Branislav Gerazov, Michael Wagner 0019
Interspeech2
2021 Parsing Speech for Grouping and Prominence, and the Typology of Rhythm
abstract
Humans appear to be wired to perceive acoustic events rhythmically. English speakers, for example, tend to perceive alternating short and long sounds as a series of binary groups with a final beat (iambs), and alternating soft and loud sounds as a series of trochees. This generalization, often called the ‘Iambic-trochaic Law’ (ITL), although viewed as an auditory universal by some, has been argued to be shaped by language experience. Earlier work on the ITL had a crucial limitation, in that it did not tease apart the percepts of grouping and prominence, which the notions of iamb and trochee inherently confound. We explore how intensity and duration relate to percepts of prominence and grouping in six languages (English, French, German, Japanese, Mandarin, and Spanish). The results show that the ITL is not universal, and that cue interpretation is shaped by language experience. However, there are also invariances: Duration appears relatively robust across languages as a cue to prominence (longer syllables are perceived as stressed), and intensity for grouping (louder syllables are perceived as initial). The results show the beginnings of a rhythmic typology based on how the dimensions of grouping and prominence are cued.
Michael Wagner 0019, Alvaro Iturralde Zurita
Interspeech1
2017 Montreal Forced Aligner: Trainable Text-Speech Alignment Using Kaldi
Michael McAuliffe, Michaela Socolof, Sarah Mihuc, Michael Wagner 0019, Morgan Sonderegger
INTERSPEECH4
2017 Three Dimensions of Sentence Prosody and Their (Non-)Interactions
Michael Wagner 0019, Michael McAuliffe
INTERSPEECH1