VLDB 2026 Research / reviewers in the wild / expert
David Alfonso-Hermelo
dblp:241/1750
· DBLP profile ↗
5ranked-venue papers
0as first author
4since 2021 · last 2024
0009-0009-4591-3077ORCID · corroborated
Domains — the database's venue-derived domains; a paper can count in several
Artificial intelligence and machine learning · 4 · 3 since 2021Databases, data management, data science and information retrieval · 1 · 1 since 2021
Expertise — from the expertise taxonomy: the topics of the expert's papers under the CCF categories. A weight counts papers with recency: 1 for a paper about the topic, 0.3 when the topic is its context, halved every five years.
| Databases, data mining, and information retrieval
3 papers |
Information retrieval · 100% | |
| Artificial intelligence
2 papers |
Knowledge representation and reasoning · 50% Information extraction and text analysis · 50% |
Topics — the 7 heaviest of 7, each with the papers that count most for it
| Topic | Weight | Papers | Last | Evidence papers |
|---|---|---|---|---|
Information retrieval
cross-language information retrieval |
1.0 | 2 | 2024 | CIRAL: A Test Collection for CLIR Evaluations in African Languages · SIGIR 2024 EUROPA: A Legal Multilingual Keyphrase Generation Dataset · ACL (1) 2024 |
Natural language and speech › Information extraction and text analysis
keyphrase generation |
0.8 | 1 | 2024 | EUROPA: A Legal Multilingual Keyphrase Generation Dataset · ACL (1) 2024 |
Knowledge, reasoning and agents › Knowledge representation and reasoning › knowledge graph
knowledge graph querying |
0.8 | 1 | 2024 | EWEK-QA : Enhanced Web and Efficient Knowledge Graph Retrieval for Citation-based Question Answering Systems · ACL (1) 2024 |
Information retrieval
evaluation |
0.8 | 1 | 2024 | CIRAL: A Test Collection for CLIR Evaluations in African Languages · SIGIR 2024 |
Information retrieval › evaluation
test collection |
0.8 | 1 | 2024 | CIRAL: A Test Collection for CLIR Evaluations in African Languages · SIGIR 2024 |
Information retrieval › evaluation
relevance judgment |
0.2 | 1 | 2024 | CIRAL: A Test Collection for CLIR Evaluations in African Languages · SIGIR 2024 |
Information retrieval › web search
web information retrieval |
0.2 | 1 | 2024 | EWEK-QA : Enhanced Web and Efficient Knowledge Graph Retrieval for Citation-based Question Answering Systems · ACL (1) 2024 |
Methods — techniques the papers use, named apart from their topics
dataset construction · 1.5
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2024 | EWEK-QA : Enhanced Web and Efficient Knowledge Graph Retrieval for Citation-based Question Answering SystemsabstractMohammad Dehghan, Mohammad Alomrani, Sunyam Bagga, David Alfonso-Hermelo, Khalil Bibi, Abbas Ghaddar, Yingxue Zhang, Xiaoguang Li, Jianye Hao, Qun Liu, Jimmy Lin, Boxing Chen, Prasanna Parthasarathi, Mahdi Biparva, Mehdi Rezagholizadeh. Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers). 2024. Mohammad Dehghan, Mohammad Ali Alomrani, Sunyam Bagga, David Alfonso-Hermelo, Khalil Bibi, Abbas Ghaddar, Yingxue Zhang 0001, Jianye Hao, Qun Liu 0001, Jimmy Lin, Boxing Chen, Prasanna Parthasarathi, Mahdi Biparva, Mehdi Rezagholizadeh |
ACL (1) | 4 |
| 2024 | EUROPA: A Legal Multilingual Keyphrase Generation DatasetabstractOlivier Salaün, Frédéric Piedboeuf, Guillaume Le Berre, David Alfonso-Hermelo, Philippe Langlais. Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers). 2024. Olivier Salaün, Frédéric Piedboeuf, Guillaume Le Berre, David Alfonso-Hermelo, Philippe Langlais |
ACL (1) | 4 |
| 2024 | CIRAL: A Test Collection for CLIR Evaluations in African LanguagesabstractCross-lingual information retrieval (CLIR) continues to be an actively studied topic in information retrieval (IR), and there have been consistent efforts in curating test collections to support its research. However, there is a lack of high-quality human-annotated CLIR resources for African languages: the few existing collections are mostly curated synthetically or from sources with limited corpora for these languages. We present CIRAL, a test collection for cross-lingual retrieval with English queries and passages in four African languages: Hausa, Somali, Swahili, and Yoruba. CIRAL's corpora are obtained from Indigenous African websites and consist of a total of over 2.5 million passages. We gathered over 1,600 queries and 30k high-quality binary relevance judgments annotated by native speakers of the languages. Additional pools were also obtained at CIRAL's shared task, which was hosted at the Forum for Information Retrieval Evaluation 2023 to encourage community participation in CLIR for African languages. We describe the design and curation process of our test collection and provide reproducible baselines that demonstrate CIRAL's utility in evaluating the effectiveness of systems. CIRAL is available at https://github.com/ciralproject/ciral. Mofe Adeyemi, Akintunde Oladipo, Xinyu Zhang 0018, David Alfonso-Hermelo, Mehdi Rezagholizadeh, Boxing Chen, Abdul-Hakeem Omotayo, Idris Abdulmumin, Naome A. Etori, Toyib Babatunde Musa, Samuel Fanijo, Oluwabusayo Olufunke Awoyomi, Saheed Abdullahi Salahudeen, Labaran Adamu Mohammed, Daud Abolade, Falalu Ibrahim Lawan, Maryam Sabo Abubakar, Ruqayya Nasir Iro, Amina Abubakar Imam, Shafie Abdi Mohamed, Hanad Mohamud Mohamed, Tunde Ajayi, Jimmy Lin |
SIGIR | 4 |
| 2023 | MIRACL: A Multilingual Retrieval Dataset Covering 18 Diverse LanguagesabstractAbstract MIRACL is a multilingual dataset for ad hoc retrieval across 18 languages that collectively encompass over three billion native speakers around the world. This resource is designed to support monolingual retrieval tasks, where the queries and the corpora are in the same language. In total, we have gathered over 726k high-quality relevance judgments for 78k queries over Wikipedia in these languages, where all annotations have been performed by native speakers hired by our team. MIRACL covers languages that are both typologically close as well as distant from 10 language families and 13 sub-families, associated with varying amounts of publicly available resources. Extensive automatic heuristic verification and manual assessments were performed during the annotation process to control data quality. In total, MIRACL represents an investment of around five person-years of human annotator effort. Our goal is to spur research on improving retrieval across a continuum of languages, thus enhancing information access capabilities for diverse populations around the world, particularly those that have traditionally been underserved. MIRACL is available at http://miracl.ai/. Xinyu Zhang 0018, Nandan Thakur, Odunayo Ogundepo, Ehsan Kamalloo, David Alfonso-Hermelo, Qun Liu 0001, Mehdi Rezagholizadeh, Jimmy Lin |
Trans. Assoc. Comput. Linguistics | 5 |
| 2020 | Human or Neural Translation?abstractShivendra Bhardwaj, David Alfonso Hermelo, Phillippe Langlais, Gabriel Bernier-Colborne, Cyril Goutte, Michel Simard. Proceedings of the 28th International Conference on Computational Linguistics. 2020. Shivendra Bhardwaj, David Alfonso-Hermelo, Philippe Langlais, Gabriel Bernier-Colborne, Cyril Goutte, Michel Simard |
COLING | 2 |