Nicolas Audibert

dblp:09/6477 · DBLP profile ↗
← Back
24ranked-venue papers
7as first author
8since 2021 · last 2025
0000-0003-3648-9322ORCID · corroborated

Domains — the database's venue-derived domains; a paper can count in several

Artificial intelligence and machine learning · 20 · 5 first-author · 7 since 2021Graphics, computer vision, multimedia, augmented reality and games · 16 · 5 first-author · 6 since 2021Human-computer interaction and ubiquitous computing · 3 · 2 first-author · 1 since 2021
YearPublicationVenuePosition
2025 Corpus-Based Insights into Mandarin Neutral Tone: Effects of Tonal Context and Structural Patterns in Spontaneous Speech
abstract
Abstract book: https://www.isca-archive.org/interspeech_2025/booklet.pdf /// Full conference proceedings: https://www.isca-archive.org/interspeech_2025/
Nicolas Audibert, Yaru Wu, Martine Adda-Decker
INTERSPEECH2
2024 Do Speaker-dependent Vowel Characteristics depend on Speech Style?
abstract
International audience
Nicolas Audibert, Cécile Fougeron, Christine Meunier
INTERSPEECH1
2023 Performative Vocal Synthesis for Foreign Language Intonation Practice
abstract
Typical foreign language (L2) pronunciation training focuses mainly on individual sounds. Intonation, the patterns of pitch change across words or phrases is often neglected, despite its key role in word-level intelligibility and in the expression of attitudes and affect. This paper examines hand-controlled real-time vocal synthesis, known as Performative Vocal Synthesis (PVS), as an interaction technique for practicing L2 intonation in computer aided pronunciation training (CAPT).
Xiao Xiao 0001, Barbara Kühnert, Nicolas Audibert, Grégoire Locqueville, Claire Pillot-Loiseau, Haohan Zhang 0004, Christophe d'Alessandro
CHI3
2023 Evaluation of delexicalization methods for research on emotional speech
abstract
International audience
Nicolas Audibert, Francesca Carbone, Maud Champagne-Lavau, Aurélien Said Housseini, Caterina Petrone
INTERSPEECH1
2022 Intra-speaker phonetic variation in read speech: comparison with inter-speaker variability in a controlled population
abstract
International audience
Nicolas Audibert, Cécile Fougeron
INTERSPEECH1
2022 Comparison of 5 methods for the evaluation of intelligibility in mild to moderate French dysarthric speech
abstract
Altered quality of the phonetic-acoustic information in the speech signal in the case of motor speech disorders may reduce its intelligibility. Monitoring intelligibility is part of the standard clinical assessment of patients. It is also a valuable tool to index the evolution of the speech disorder. However, measuring intelligibility raises methodological debates concerning: the type of linguistic material on which the assessment is based (non-words, words, continuous speech), the evaluation protocol and type of scores (scale-based rating, transcription or recognition tests), and the advantages and disadvantages of listener vs. automatic-based approaches (subjective vs. objective, expertise level, types of models used). In this paper, the intelligibility of the speech of 32 French patients presenting mild to moderate dysarthria and 17 elderly speakers is assessed with five different methods: impressionistic clinician judgment on continuous speech, number of words recognized in an interactive face-to-face setting and in an on-line testing of the same material by 75 judges, automatic feature-based and automatic speech recognition-based methods (both on short sentences). The implications of the different methods for clinical practice are discussed.
Cécile Fougeron, Nicolas Audibert, Ina Kodrasi, Parvaneh Janbakhshi, Michaela Pernon, Nathalie Lévêque, Stephanie Borel, Marina Laganaro, Hervé Bourlard, Frédéric Assal
INTERSPEECH2
2022 PATATRA and PATAFreq: two French databases for the documentation of within-speaker variability in speech
abstract
Our knowledge on speech is historically built on data comparing different speakers or data averaged across speakers. Consequently, little is known on the variability in the speech of a single individual. Experimental studies have shown that speakers adapt to the linguistic and the speaking contexts, and modify their speech according to their emotional or biological condition, etc. However, it is unclear how much speakers vary from one repetition to the next, and how comparable are recordings that are collected days, months or years apart. In this paper, we introduce two French databases which contain recordings of 9 to 11 speakers recorded over 9 to 18 sessions, allowing comparisons of speech tasks with a different delay between the repetitions: 3 repetitions within the same session, 6 to 10 repetitions on different days during a two months period, 5 to 9 repetitions on different years. Speakers are recorded on a large set of speech tasks including read and spontaneous speech as well as speech-like performance tasks. In this paper, we provide detailed descriptions of the two databases and available annotations. We conclude by an illustration on how these data can inform on within-speaker variability of speech.
Cécile Fougeron, Nicolas Audibert, Cédric Gendrot, Estelle Chardenon, Louise Wohmann
LREC2
2021 Prosodic Disambiguation Using Chironomic Stylization of Intonation with Native and Non-Native Speakers
abstract
International audience
Nicolas Audibert, Grégoire Locqueville, Christophe d'Alessandro, Barbara Kühnert, Claire Pillot-Loiseau
Interspeech2
2020 Towards Interactive Annotation for Hesitation in Conversational Speech
abstract
Manual annotation of speech corpora is expensive in both human resources and time. Furthermore, recognizing affects in spontaneous, non acted speech presents a challenge for humans and machines. The aim of the present study is to automatize the labeling of hesitant speech as a marker of expressed uncertainty. That is why, the NCCFr-corpus was manually annotated for ‘degree of hesitation’ on a continuous scale between -3 and 3 and the affective dimensions ‘activation, valence and control’. In total, 5834 chunks of the NCCFr-corpus were manually annotated. Acoustic analyses were carried out based on these annotations. Furthermore, regression models were trained in order to allow automatic prediction of hesitation for speech chunks that do not have a manual annotation. Preliminary results show that the number of filled pauses as well as vowel duration increase with the degree of hesitation, and that automatic prediction of the hesitation degree reaches encouraging RMSE results of 1.6.
Jane Wottawa, Marie Tahon, Apolline Marin, Nicolas Audibert
LREC4
2019 " Gra[f] e!" Word-Final Devoicing of Obstruents in Standard French: An Acoustic Study Based on Large Corpora
abstract
International audience
Adèle Jatteau, Ioana Vasilescu, Lori Lamel, Martine Adda-Decker, Nicolas Audibert
INTERSPEECH5
2017 Relationships Between Speech Timing and Perceived Hostility in a French Corpus of Political Debates
abstract
International audience
Charlotte Kouklia, Nicolas Audibert
INTERSPEECH2
2016 The Effects of Prosody on French V-to-V Coarticulation: A Corpus-Based Study
abstract
International audience
Giuseppina Turco, Cécile Fougeron, Nicolas Audibert
INTERSPEECH3
2014 An educational platform to capture, visualize and analyze rare singing
Patrick Chawah, Samer Al Kork, Thibaut Fux, Martine Adda-Decker, Angélique Amelot, Nicolas Audibert, Bruce Denby, Gérard Dreyfus, Aurore Jaumard-Hakoun, Claire Pillot-Loiseau, Pierre Roussel-Ragot, Maureen Stone 0001, Kele Xu, Lise Crevier-Buchman
INTERSPEECH6
2013 Is protrusion of French rounded vowels affected by prosodic positions?
abstract
International audience
Laurianne Georgeton, Nicolas Audibert
INTERSPEECH2
2013 Perceptual, acoustic and electroglottographic correlates of 3 aggressive attitudes in French: a pilot study
Charlotte Kouklia, Nicolas Audibert
INTERSPEECH2
2011 Speaker verification by inexperienced and experienced listeners vs. speaker verification system
abstract
This paper describes the participation of the LIA in the Human Assisted Speaker Recognition (HASR) task of the NIST-SRE 2010 evaluation campaign and its extension to a larger number of listeners.The human performance in such unfavorable conditions is analyzed in relation to the decision of a speaker recognition automatic system. Results of the perception test showed an important inter-trial variability (from 3% to 90% of correct answers for non-target trials) whereas there was no significant difference between the experienced and inexperienced listeners. Some complementarity between speaker verification system and human decisions was also found.
Juliette Kahn, Nicolas Audibert, Solange Rossato, Jean-François Bonastre
ICASSP2
2011 Comparison of Nasalance Measurements from Accelerometers and Microphones and Preliminary Development of Novel Features
abstract
International audience
Nicolas Audibert, Angélique Amelot
INTERSPEECH1
2008 Multimodal Spontaneous Expressive Speech Corpus for Hungarian
Márk Fék, Nicolas Audibert, János Szabó, Albert Rilliard, Géza Németh, Véronique Aubergé
LREC2
2007 Gradient or Contours Cues? A Gating Experiment for the Timing of the Emotional Information
Nicolas Audibert, Véronique Aubergé
ACII1
2005 The Relative Weights of the Different Prosodic Dimensions in Expressive Speech: A Resynthesis Study
Nicolas Audibert, Véronique Aubergé, Albert Rilliard
ACII1
2005 The prosodic dimensions of emotion in speech: the relative weights of parameters
abstract
International audience
Nicolas Audibert, Véronique Aubergé, Albert Rilliard
INTERSPEECH1
2004 E-Wiz: a Trapper Protocol for Hunting the Expressive Speech Corpora in Lab
Véronique Aubergé, Nicolas Audibert, Albert Rilliard
LREC2
2004 Evaluating an Authentic Audio-Visual Expressive Speech Corpus
Albert Rilliard, Véronique Aubergé, Nicolas Audibert
LREC3
2003 Why and how to control the authentic emotional speech corpora
Véronique Aubergé, Nicolas Audibert, Albert Rilliard
INTERSPEECH2