VLDB 2026 Research / reviewers in the wild / expert
Namhyun Cho
dblp:140/8714
· DBLP profile ↗
7ranked-venue papers
2as first author
7since 2021 · last 2025
—ORCID · none
Domains — the database's venue-derived domains; a paper can count in several
Graphics, computer vision, multimedia, augmented reality and games · 7 · 2 first-author · 7 since 2021Artificial intelligence and machine learning · 6 · 2 first-author · 6 since 2021
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2025 | Unleashing the Inner Monster: Demonstrating High-Fidelity Human to Non-Human Voice Conversion
Namhyun Cho, Sunmin Kim, Minsu Kang, Seolhee Lee, Choonghyeon Lee, Yangsun Lee 0003 |
INTERSPEECH | 1 |
| 2025 | When Humans Growl and Birds Speak: High-Fidelity Voice Conversion from Human to Animal and Designed Sounds
Minsu Kang, Seolhee Lee, Choonghyeon Lee, Namhyun Cho |
INTERSPEECH | 4 |
| 2024 | Iphonmatchnet: Zero-Shot User-Defined Keyword Spotting Using Implicit Acoustic Echo CancellationabstractIn response to the increasing interest in human–machine communication across various domains, this paper introduces a novel approach called iPhonMatchNet, which addresses the challenge of barge-in scenarios, wherein user speech overlaps with device playback audio, thereby creating a self-referencing problem. The proposed model leverages implicit acoustic echo cancellation (iAEC) techniques to increase the efficiency of user-defined keyword spotting models, achieving a remarkable 95% reduction in mean absolute error with a minimal increase in model size (0.13%) compared to the baseline model, PhonMatchNet. We also present an efficient model structure and demonstrate its capability to learn iAEC functionality without requiring a clean signal. The findings of our study indicate that the proposed model achieves competitive performance in real-world deployment conditions of smart devices. Yong-Hyeok Lee, Namhyun Cho |
ICASSP | 2 |
| 2024 | Centroid Estimation with Transformer-Based Speaker Embedder for Robust Target Speaker Extraction
Woon-Haeng Heo, Joongyu Maeng, Yoseb Kang, Namhyun Cho |
INTERSPEECH | 4 |
| 2023 | Fast Enrollable Streaming Keyword Spotting System: Training and Inference using a Web Browser
Namhyun Cho, Sunmin Kim, Yoseb Kang, Heeman Kim |
INTERSPEECH | 1 |
| 2023 | Focus-attention-enhanced Crossmodal Transformer with Metric Learning for Multimodal Speech Emotion Recognition
Keulbit Kim, Namhyun Cho |
INTERSPEECH | 2 |
| 2023 | PhonMatchNet: Phoneme-Guided Zero-Shot Keyword Spotting for User-Defined KeywordsabstractThis study presents a novel zero-shot user-defined keyword spotting model that utilizes the audio-phoneme relationship of the keyword to improve performance. Unlike the previous approach that estimates at utterance level, we use both utterance and phoneme level information. Our proposed method comprises a two-stream speech encoder architecture, self-attention-based pattern extractor, and phoneme-level detection loss for high performance in various pronunciation environments. Based on experimental results, our proposed model outperforms the baseline model and achieves competitive performance compared with full-shot keyword spotting models. Our proposed model significantly improves the EER and AUC across all datasets, including familiar words, proper nouns, and indistinguishable pronunciations, with an average relative improvement of 67% and 80%, respectively. The implementation code of our proposed model is available at https://github.com/ncsoft/PhonMatchNet. Yong-Hyeok Lee, Namhyun Cho |
INTERSPEECH | 2 |