VLDB 2026 Research / reviewers in the wild / expert
Mingdu Huangfu
dblp:365/9571
· DBLP profile ↗
1ranked-venue papers
0as first author
1since 2021 · last 2024
—ORCID · none
Domains — the database's venue-derived domains; a paper can count in several
Artificial intelligence and machine learning · 1 · 1 since 2021Graphics, computer vision, multimedia, augmented reality and games · 1 · 1 since 2021
Expertise — from the expertise taxonomy: the topics of the expert's papers under the CCF categories. A weight counts papers with recency: 1 for a paper about the topic, 0.3 when the topic is its context, halved every five years.
| Artificial intelligence
1 paper |
Speech recognition and synthesis · 50% Language models and text generation · 50% |
Topics — the 2 heaviest of 2, each with the papers that count most for it
| Topic | Weight | Papers | Last | Evidence papers |
|---|---|---|---|---|
Natural language and speech › Language models and text generation › text generation › large language model generation
prompt-based generation |
0.8 | 1 | 2024 | SIG: Speaker Identification in Literature via Prompt-Based Generation · AAAI 2024 |
Natural language and speech › Speech recognition and synthesis › speaker recognition
speaker identification |
0.8 | 1 | 2024 | SIG: Speaker Identification in Literature via Prompt-Based Generation · AAAI 2024 |
Methods — techniques the papers use, named apart from their topics
prompt-based generation · 0.8large language model · 0.8
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2024 | SIG: Speaker Identification in Literature via Prompt-Based GenerationabstractIdentifying speakers of quotations in narratives is an important task in literary analysis, with challenging scenarios including the out-of-domain inference for unseen speakers, and non-explicit cases where there are no speaker mentions in surrounding context. In this work, we propose a simple and effective approach SIG, a generation-based method that verbalizes the task and quotation input based on designed prompt templates, which also enables easy integration of other auxiliary tasks that further bolster the speaker identification performance. The prediction can either come from direct generation by the model, or be determined by the highest generation probability of each speaker candidate. Based on our approach design, SIG supports out-of-domain evaluation, and achieves open-world classification paradigm that is able to accept any forms of candidate input. We perform both cross-domain evaluation and in-domain evaluation on PDNC, the largest dataset of this task, where empirical results suggest that SIG outperforms previous baselines of complicated designs, as well as the zero-shot ChatGPT, especially excelling at those hard non-explicit scenarios by up to 17% improvement. Additional experiments on another dataset WP further corroborate the efficacy of SIG. Zhenlin Su, Liyan Xu, Mingdu Huangfu |
AAAI | 5 |