EDBT 2026 Demo / reviewers in the wild / expert
Benjamin Elie
dblp:154/1400
· DBLP profile ↗
11ranked-venue papers
10as first author
8since 2021 · last 2025
0000-0001-5889-763XORCID · corroborated
Domains — the database's venue-derived domains; a paper can count in several
Graphics, computer vision, multimedia, augmented reality and games · 11 · 10 first-author · 8 since 2021Artificial intelligence and machine learning · 7 · 6 first-author · 6 since 2021
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2025 | Self-supervised Optimality-Guided Learning of Speech ArticulationabstractThis paper introduces a novel approach for modeling speech articulatory planning based on Optimal Control Theory. The presented approach uses an internal feed-forward controller model that learns to predict optimal articulatory commands minimizing a context-dependent objective function. This objective function combines conflicting tasks of minimizing articulatory effort and maximizing the recognition probability of a target vowel based on acoustic characteristics. We present a self-supervised optimality-guided architecture for training the feedforward internal model that directly uses the objective function as a training loss. Simulations involving isolated vowels of American-English show that online training of the internal model enables feedforward estimation of near-optimal articulatory parameters. Juraj Simko, Benjamin Elie, Alice Turk |
INTERSPEECH | 2 |
| 2024 | Articulatory Configurations across Genders and Periods in French Radio and TV archivesabstractInternational audience Benjamin Elie, David Doukhan, Rémi Uro, Lucas Ondel Yang, Albert Rilliard, Simon Devauchelle |
INTERSPEECH | 1 |
| 2024 | A data-driven model of acoustic speech intelligibility for optimization-based models of speech productionabstractThis paper presents a data-driven model of intelligibility which is intended to be used in an optimization-based model of speech production.The BiLSTM-based model is trained as a phoneme classifier and takes a sequence of real articulatory trajectories as input and returns the probability of phonemes over time.The optimization minimizes a cost function which is the weighted sum of the conflicting demands of being intelligible and least articulatory effort.The data-driven intelligibility model presented in this paper is used to compute the intelligibility score.Simulations support Lindblom's hypo-and hyper-articulation theory of speech, as the degree of hyper-articulation of speech can be modified and tuned along a continuum by balancing the importance given to both requirements of intelligibility and least articulatory effort. Benjamin Elie, Juraj Simko, Alice Turk |
INTERSPEECH | 1 |
| 2024 | Optimization-based planning of speech articulation using general Tau TheoryabstractThis paper presents a model of speech articulation planning and generation based on General Tau Theory and Optimal Control Theory. Because General Tau Theory assumes that articulatory targets are always reached, the model accounts for speech variation via context-dependent articulatory targets. Targets are chosen via the optimization of a composite objective function. This function models three different task requirements: maximal intelligibility, minimal articulatory effort and minimal utterance duration. The paper shows that systematic phonetic variability can be reproduced by adjusting the weights assigned to each task requirement. Weights can be adjusted globally to simulate different speech styles, and can be adjusted locally to simulate different levels of prosodic prominence. The solution of the optimization procedure contains Tau equation parameter values for each articulatory movement, namely position of the articulator at the movement offset, movement duration, and a parameter which relates to the shape of the movement’s velocity profile. The paper presents simulations which illustrate the ability of the model to predict or reproduce several well-known characteristics of speech. These phenomena include close-to-symmetric velocity profiles for articulatory movement, variation related to speech rate, centralization of unstressed vowels, lengthening of stressed vowels, lenition of unstressed lingual stop consonants, and coarticulation of stop consonants. Benjamin Elie, Juraj Simko, Alice Turk |
Speech Commun. | 1 |
| 2023 | Optimal control of speech with context-dependent articulatory targetsabstractThis paper presents a computational implementation of phonetic planning which consists of choosing the position of articulatory targets which satisfy conflicting linguistic and extra-linguistic requirements. We present a minimal model that considers intelligibility and least effort as task requirements. To achieve the context-dependent variability of targets, our model approximates intelligibility as a function of target phoneme recognition probability given a vector of articulatory parameters. Preliminary experiments show that our minimal computational model of phonetic planning is able to predict two types of hypoarticulation by adjusting the weight assigned to effort: vowel centralization and stop consonant lenition. Benjamin Elie, Juraj Simko, Alice Turk |
INTERSPEECH | 1 |
| 2023 | Estimating virtual targets for lingual stop consonants using general Tau theoryabstractThis paper investigates the existence and position of virtual targets during the production of stop consonants. Using the equations from general Tau theory to model the time-course of tongue constriction formation movements, targets were estimated by fitting these equations on observed tongue constriction variables extracted from real EMA data from 2 native speakers of English. Results suggest that targets are virtual for 50 to 60% of movements. For these movements, virtual targets of the tongue tip constriction are predicted to occur around 0.1 cm beyond the palate, and virtual targets for the tongue dorsum constriction are predicted to occur between 0.05 and 0.2 cm beyond the palate. Our results suggest that the time-course of movement is planned so that the onset of closure occurs with relatively high velocity: closure onset is generally located very close in time to the time of peak velocity. Benjamin Elie, Alice Turk |
INTERSPEECH | 1 |
| 2023 | Modeling trajectories of human speech articulators using general Tau theoryabstractThis paper presents an application of general Tau theory to the modeling and analysis of articulatory trajectories in speech. We evaluated the model using electromagnetic articulometry data from 12 native speakers of English reading a common text, where trajectories of the following sensors were fitted: lower and upper lips, jaw, and three tongue sensors. Additionally, we analyzed trajectories of the lip aperture signal. Our experiments show that the general Tau theory model gives a better fit than existing (i) methods based on critically damped oscillators, and (ii) a method based on sequential target approximation. These findings support the hypothesis of Tau-guided movements of articulators during speech production. In the second part of the paper, our Tau theory analysis shows that articulatory movements follow similar velocity profile distributions across speakers. In particular, the value of the shape parameter κ of the Tau theory equation is identically distributed across speakers, following a unimodal distribution. The statistical mode of the distribution corresponds to the value of κ that generates a symmetric velocity profile. The analysis of the statistical distribution of κ values also reveals that its variance decreases when greater articulatory effort is required, such that produced articulatory effort remains close to that predicted by the theoretical minimal cost function based on forces acting on the moving articulator. This provides new evidence that articulatory effort is optimized during speech production. Benjamin Elie, David N. Lee, Alice Turk |
Speech Commun. | 1 |
| 2021 | Modeling the Effect of Military Oxygen Masks on Speech CharacteristicsabstractInternational audience Benjamin Elie, Jodie Gauvain, Jean-Luc Gauvain, Lori Lamel |
Interspeech | 1 |
| 2017 | Glottal Opening and Strategies of Production of FricativesabstractInternational audience Benjamin Elie, Yves Laprie |
INTERSPEECH | 1 |
| 2016 | A glottal chink model for the synthesis of voiced fricativesabstractThis paper presents a simulation framework that enables a glottal chink model to be integrated into a time-domain continuous speech synthesizer along with self-oscillating vocal folds. The glottis is then made up of two main separated components: a self-oscillating part and a constantly open chink. This feature allows the simulation of voiced fricatives, thanks to a self-oscillating model of the vocal folds to generate the voiced source, and the glottal opening that is necessary to generate the frication noise. Numerical simulations show the accuracy of the model to simulate voiced fricative, and also phonetic assimilation, such as sonorization and devoicing. The simulation framework is also used to show that the phonatory/articulatory space for generating voiced fricatives is different according to the desired sound: for instance, the minimal glottal opening for generating frication noise is shorter for /z/ than for /3/. Benjamin Elie, Yves Laprie |
ICASSP | 1 |
| 2016 | Extension of the single-matrix formulation of the vocal tract: Consideration of bilateral channels and connection of self-oscillating models of the vocal folds with a glottal chink
Benjamin Elie, Yves Laprie |
Speech Commun. | 1 |