VLDB 2026 Research / reviewers in the wild / expert
Patrizia Paggio
dblp:31/6683
· DBLP profile ↗
20ranked-venue papers
7as first author
4since 2021 · last 2026
0000-0002-2484-2275ORCID · corroborated
Domains — the database's venue-derived domains; a paper can count in several
Artificial intelligence and machine learning · 20 · 7 first-author · 4 since 2021Databases, data management, data science and information retrieval · 3
Expertise — from the expertise taxonomy: the topics of the expert's papers under the CCF categories. A weight counts papers with recency: 1 for a paper about the topic, 0.3 when the topic is its context, halved every five years.
| Artificial intelligence
1 paper |
Representation and self-supervised learning · 100% | |
| Interdisciplinary, comprehensive, and emerging computing
1 paper |
Computational social science and digital humanities · 100% |
Topics — the 1 heaviest of 2, each with the papers that count most for it
| Topic | Weight | Papers | Last | Evidence papers |
|---|---|---|---|---|
Computational social science and digital humanities
historical linguistics |
0.2 | 1 | 2022 | Letters From the Past: Modeling Historical Sound Change Through Diachronic Character Embeddings · ACL (1) 2022 |
Methods — techniques the papers use, named apart from their topics
PPMI embedding · 1.1
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2026 | The MultiplEYE Text Corpus: Towards a Diverse and Ever-Expanding Multilingual Text Corpus
Ramune Kaspere, Anna Bondar, Sergiu Nisioi, Maja Stegenwallner-Schütz, Hanne B. Søndergaard Knudsen, Ana Matic Skoric, Eva Pavlinusic Vilus, Dorota Klimek-Jankowska, Chiara Tschirner, Not Battesta Soliva, Deborah N. Jakobi, Cui Ding, Dima Abu Romi, Cengiz Acartürk, Matilda Agdler, Anton Marius Alexandru, Mohd Faizan Ansari, Annalisa Arcidiacono, Elizabete Ausma Velta Barisa, Ana Bautista, Lisa Beinborn, Yevgeni Berzak, Nedeljka Bjelanovic, Anna Isabelle Bothmann, Jan Brasser, Caterina Cacioli, Anila Çepani, Ilze Ceple, Adelina Çerpja, Dalí Chirino, Jan Chromý, Alessandro Corona Mendozza, Iria de-Dios-Flores, Nazik Dinçtopal Deniz, Ana Dosen, Kristian Elersic, Inmaculada Fajardo, Zigmunds Freibergs, Angelina Ganebnaya, Jessica Gomes, Annjo Klungervik Greenall, Alba Haveriku, Anamaria Hodivoianu, Yu-Yin Hsu, Amanda Isaksen, Andreia Janeiro, Kristine M. Jensen de López, Aleksandar Jevremovic, Vojislav Jovanovic, Hanna Kedzierska, Nik Kharlamov, Sara Kosutar, Nelda Kote, Vanja Kovic, Izabela Krejtz, Thyra Krosness, Oleksandra Kuvshynova, Eilam Lavy, Ella Lion, Marta Lockiewicz, Kaidi Lõo, Paula Luegi, Mircea Mihai Marin, Clara Martin, Svitlana Matvieieva, Diane C. Mézière, Xavier Mínguez-López, Valeriia Modina, Jurgita Motiejuniene, Marie-Luise Müller, Tolgonai Nasipbek kyzy, Jamal Abdul Nasir, Johanne Sofie Krog Nedergård, Aysegül Özkan, Patrizia Paggio, Marijan Palmovic, Maria Christina Panagiotopoulou, Alberto Parola, Helena Pérez, Klaudia Petersen, Anja Podlesek, Eva Pospísilová, Marta Praulina, Mikulás Preininger, Loredana Punga, Diego Rossini, Spela Rot, Habib Sani Yahaya, Irina A. Sekerina, Anne Gabija Skadina, Jordi Solé i Casals, Lonneke van der Plas, Saara M. Varjopuro, Spyridoula Varlokosta, João Veríssimo, Oskari Juhapekka Virtanen, Nemanja Vracar, Mila Dimitrova-Vulchanova, Ahmad Mustapha Wali, Peizheng Wu, Nilgün Yücel, Stefan Frank, Nora Hollenstein, Lena A. Jäger, Somayeh Bakhtiari |
LREC | 77 |
| 2026 | Multimodal Entrainment and Feedback in Online Group Meetings
Patrizia Paggio, Manex Agirrezabal, Giulia Di Cristina, Bart Jongejan, Costanza Navarretta |
LREC | 1 |
| 2024 | Multimodal Behaviour in an Online Environment: The GEHM Zoom Corpus CollectionabstractThis paper introduces a novel multimodal corpus consisting of 12 video recordings of Zoom meetings held in English by an international group of researchers from September 2021 to March 2023. The meetings have an average duration of about 40 minutes each, for a total of 8 hours. The number of participants varies from 5 to 9 per meeting. The participants’ speech was transcribed automatically using WhisperX, while visual coordinates of several keypoints of the participants’ head, their shoulders and wrists, were extracted using OpenPose. The audio-visual recordings will be distributed together with the orthographic transcription as well as the visual coordinates. In the paper we describe the way the corpus was collected, transcribed and enriched with the visual coordinates, we give descriptive statistics concerning both the speech transcription and the visual keypoint values and we present and discuss visualisations of these values. Finally, we carry out a short preliminary analysis of the role of feedback in the meetings, and show how visualising the coordinates extracted via OpenPose can be used to see how gestural behaviour supports the use of feedback words during the interaction. Patrizia Paggio, Manex Agirrezabal, Costanza Navarretta, Leo Vitasovic |
LREC/COLING | 1 |
| 2022 | Letters From the Past: Modeling Historical Sound Change Through Diachronic Character EmbeddingsabstractWhile a great deal of work has been done on NLP approaches to lexical semantic change detection, other aspects of language change have received less attention from the NLP community. In this paper, we address the detection of sound change through historical spelling. We propose that a sound change can be captured by comparing the relative distance through time between the distributions of the characters involved before and after the change has taken place. We model these distributions using PPMI character embeddings. We verify this hypothesis in synthetic data and then test the method’s ability to trace the well-known historical change of lenition of plosives in Danish historical sources. We show that the models are able to identify several of the changes under consideration and to uncover meaningful contexts in which they appeared. The methodology has the potential to contribute to the study of open questions such as the relative chronology of sound shifts and their geographical distribution. Sidsel Boldsen, Patrizia Paggio |
ACL (1) | 2 |
| 2020 | Dialogue Act Annotation in a Multimodal Corpus of First Encounter DialoguesabstractThis paper deals with the annotation of dialogue acts in a multimodal corpus of first encounter dialogues, i.e. face-to- face dialogues in which two people who meet for the first time talk with no particular purpose other than just talking. More specifically, we describe the method used to annotate dialogue acts in the corpus, including the evaluation of the annotations. Then, we present descriptive statistics of the annotation, particularly focusing on which dialogue acts often follow each other across speakers and which dialogue acts overlap with gestural behaviour. Finally, we discuss how feedback is expressed in the corpus by means of feedback dialogue acts with or without co-occurring gestural behaviour, i.e. multimodal vs. unimodal feedback. Costanza Navarretta, Patrizia Paggio |
LREC | 2 |
| 2018 | Classifying the Informative Behaviour of Emoji in Microblogs
Giulia Donato, Patrizia Paggio |
LREC | 2 |
| 2018 | Face2Text: Collecting an Annotated Image Description Corpus for the Generation of Rich Face Descriptions
Albert Gatt, Marc Tanti, Adrian Muscat, Patrizia Paggio, Reuben A. Farrugia, Claudia Borg, Kenneth P. Camilleri, Mike Rosner, Lonneke van der Plas |
LREC | 4 |
| 2014 | Learning when to point: A data-driven approach
Albert Gatt, Patrizia Paggio |
COLING | 2 |
| 2012 | Feedback in Nordic First-Encounters: a Comparative Study
Costanza Navarretta, Elisabeth Ahlsén, Jens Allwood, Kristiina Jokinen, Patrizia Paggio |
LREC | 5 |
| 2012 | Multimodal Behaviour and Feedback in Different Types of Interaction
Costanza Navarretta, Patrizia Paggio |
LREC | 2 |
| 2010 | The NOMCO Multimodal Nordic Resource - Goals and Characteristics
Patrizia Paggio, Jens Allwood, Elisabeth Ahlsén, Kristiina Jokinen, Costanza Navarretta |
LREC | 1 |
| 2006 | Information Structure and Pauses in a Corpus of Spoken Danish
Patrizia Paggio |
EACL | 1 |
| 2006 | Annotating Information Structure in a Corpus of Spoken Danish
Patrizia Paggio |
LREC | 1 |
| 2004 | Ontology-Based Question Answering in a Federation of University Sites: The MOSES Case Study
Paolo Atzeni, Roberto Basili 0001, Dorte Haltrup Hansen, Paolo Missier, Patrizia Paggio, Maria Teresa Pazienza, Fabio Massimo Zanzotto |
NLDB | 5 |
| 2004 | Content-based text querying with ontological descriptors
Troels Andreasen, Per Anker Jensen, Jørgen Fischer Nilsson, Patrizia Paggio, Bolette S. Pedersen, Hanne Erdman Thomsen |
Data Knowl. Eng. | 4 |
| 2002 | Semantic Lexical Resources Applied to Content-based Querying - the OntoQuery Project
Bolette S. Pedersen, Patrizia Paggio |
LREC | 2 |
| 2002 | Ontological Extraction of Content for Text Querying
Troels Andreasen, Per Anker Jensen, Jørgen Fischer Nilsson, Patrizia Paggio, Bolette S. Pedersen, Hanne Erdman Thomsen |
NLDB | 4 |
| 1998 | Evaluation in the SCARRIE project
Patrizia Paggio, Bradley Music |
LREC | 1 |
| 1998 | Validating the TEMAA LE evaluation methodology: a case study on Danish spelling checkers
Patrizia Paggio, Nancy L. Underwood |
Nat. Lang. Eng. | 1 |
| 1991 | A Preference Mechanism Based On Multiple Criteria Resolution
Ioannis Dologlou, Giovanni Malnati, Patrizia Paggio |
EACL | 3 |