EDBT 2026 Demo / reviewers in the wild / expert
Hitoshi Suda
dblp:237/0155
· DBLP profile ↗
4ranked-venue papers
4as first author
3since 2021 · last 2025
0000-0003-2648-363XORCID · corroborated
Domains — the database's venue-derived domains; a paper can count in several
Artificial intelligence and machine learning · 4 · 4 first-author · 3 since 2021Graphics, computer vision, multimedia, augmented reality and games · 3 · 3 first-author · 2 since 2021
Expertise — from the expertise taxonomy: the topics of the expert's papers under the CCF categories. A weight counts papers with recency: 1 for a paper about the topic, 0.3 when the topic is its context, halved every five years.
| Computer graphics and multimedia
1 paper |
Audio and music processing · 100% |
Topics — the 2 heaviest of 2, each with the papers that count most for it
| Topic | Weight | Papers | Last | Evidence papers |
|---|---|---|---|---|
Audio and music processing › music information retrieval
singing information processing |
0.6 | 1 | 2022 | Singer Diarization for Polyphonic Music With Unison Singing · IEEE ACM Trans. Audio Speech Lang. Process. 2022 |
Audio and music processing › source separation › music source separation
singing voice separation |
0.2 | 1 | 2022 | Singer Diarization for Polyphonic Music With Unison Singing · IEEE ACM Trans. Audio Speech Lang. Process. 2022 |
Methods — techniques the papers use, named apart from their topics
cosacorr score · 0.6arcface · 0.6
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2025 | Voice Conversion for Likability Control via Automated Rating of Speech Synthesis Corpora
Hitoshi Suda, Shinnosuke Takamichi, Satoru Fukayama |
INTERSPEECH | 1 |
| 2024 | Who Finds This Voice Attractive? A Large-Scale Experiment Using In-the-Wild Data
Hitoshi Suda, Aya Watanabe, Shinnosuke Takamichi |
INTERSPEECH | 1 |
| 2022 | Singer Diarization for Polyphonic Music With Unison SingingabstractThis paper introduces a new framework for singer diarization, which is a technique to reveal who sings when in songs with multiple singers. Although various techniques have been developed to analyze and extract features of singing voices in musical audio signals, most of them assume that a song is sung by a single singer, and singer diarization for multiple singers has not been well studied in the field of singing information processing. To deal with multiple speakers in speech analysis, speaker diarization has been explored to handle overlapped speech voices, but cannot handle singing voices well because of acoustic differences between singing and speech voices. This paper therefore proposes a new diarization framework specialized in singing voices. To achieve high accuracy in overlap detection, this paper proposes a novel acoustic feature named Cosacorr score, which is helpful in estimating whether a song is sung by more than one singer. After extracting singing voices from polyphonic music by using a singing voice separation technique, the framework adopts an existing ArcFace technique to extract discriminative singer representations from short segments of the separated singing voices. The framework is evaluated by using a new private dataset of unison singing voices, which is constructed using commercially available compact discs (CDs). The experimental results show that the proposed framework outperformed the baseline method for speaker diarization in terms of diarization error rate (DER). Hitoshi Suda, Daisuke Saito, Satoru Fukayama, Tomoyasu Nakano, Masataka Goto |
IEEE ACM Trans. Audio Speech Lang. Process. | 1 |
| 2020 | Nonparallel Training of Exemplar-Based Voice Conversion System Using INCA-Based Alignment Technique
Hitoshi Suda, Gaku Kotani, Daisuke Saito |
INTERSPEECH | 1 |