Patrizia Paggio

dblp:31/6683 · DBLP profile ↗
← Back
20ranked-venue papers
7as first author
4since 2021 · last 2026
0000-0002-2484-2275ORCID · corroborated

Domains — the database's venue-derived domains; a paper can count in several

Artificial intelligence and machine learning · 20 · 7 first-author · 4 since 2021Databases, data management, data science and information retrieval · 3

Expertise — from the expertise taxonomy: the topics of the expert's papers under the CCF categories. A weight counts papers with recency: 1 for a paper about the topic, 0.3 when the topic is its context, halved every five years.

Artificial intelligence
1 paper
Representation and self-supervised learning · 100%
Interdisciplinary, comprehensive, and emerging computing
1 paper
Computational social science and digital humanities · 100%

Topics — the 1 heaviest of 2, each with the papers that count most for it

TopicWeightPapersLastEvidence papers
Computational social science and digital humanities
historical linguistics
0.212022
Letters From the Past: Modeling Historical Sound Change Through Diachronic Character Embeddings · ACL (1) 2022

Methods — techniques the papers use, named apart from their topics

PPMI embedding · 1.1
YearPublicationVenuePosition
2026 The MultiplEYE Text Corpus: Towards a Diverse and Ever-Expanding Multilingual Text Corpus
Ramune Kaspere, Anna Bondar, Sergiu Nisioi, Maja Stegenwallner-Schütz, Hanne B. Søndergaard Knudsen, Ana Matic Skoric, Eva Pavlinusic Vilus, Dorota Klimek-Jankowska, Chiara Tschirner, Not Battesta Soliva, Deborah N. Jakobi, Cui Ding, Dima Abu Romi, Cengiz Acartürk, Matilda Agdler, Anton Marius Alexandru, Mohd Faizan Ansari, Annalisa Arcidiacono, Elizabete Ausma Velta Barisa, Ana Bautista, Lisa Beinborn, Yevgeni Berzak, Nedeljka Bjelanovic, Anna Isabelle Bothmann, Jan Brasser, Caterina Cacioli, Anila Çepani, Ilze Ceple, Adelina Çerpja, Dalí Chirino, Jan Chromý, Alessandro Corona Mendozza, Iria de-Dios-Flores, Nazik Dinçtopal Deniz, Ana Dosen, Kristian Elersic, Inmaculada Fajardo, Zigmunds Freibergs, Angelina Ganebnaya, Jessica Gomes, Annjo Klungervik Greenall, Alba Haveriku, Anamaria Hodivoianu, Yu-Yin Hsu, Amanda Isaksen, Andreia Janeiro, Kristine M. Jensen de López, Aleksandar Jevremovic, Vojislav Jovanovic, Hanna Kedzierska, Nik Kharlamov, Sara Kosutar, Nelda Kote, Vanja Kovic, Izabela Krejtz, Thyra Krosness, Oleksandra Kuvshynova, Eilam Lavy, Ella Lion, Marta Lockiewicz, Kaidi Lõo, Paula Luegi, Mircea Mihai Marin, Clara Martin, Svitlana Matvieieva, Diane C. Mézière, Xavier Mínguez-López, Valeriia Modina, Jurgita Motiejuniene, Marie-Luise Müller, Tolgonai Nasipbek kyzy, Jamal Abdul Nasir, Johanne Sofie Krog Nedergård, Aysegül Özkan, Patrizia Paggio, Marijan Palmovic, Maria Christina Panagiotopoulou, Alberto Parola, Helena Pérez, Klaudia Petersen, Anja Podlesek, Eva Pospísilová, Marta Praulina, Mikulás Preininger, Loredana Punga, Diego Rossini, Spela Rot, Habib Sani Yahaya, Irina A. Sekerina, Anne Gabija Skadina, Jordi Solé i Casals, Lonneke van der Plas, Saara M. Varjopuro, Spyridoula Varlokosta, João Veríssimo, Oskari Juhapekka Virtanen, Nemanja Vracar, Mila Dimitrova-Vulchanova, Ahmad Mustapha Wali, Peizheng Wu, Nilgün Yücel, Stefan Frank, Nora Hollenstein, Lena A. Jäger, Somayeh Bakhtiari
LREC77
2026 Multimodal Entrainment and Feedback in Online Group Meetings
Patrizia Paggio, Manex Agirrezabal, Giulia Di Cristina, Bart Jongejan, Costanza Navarretta
LREC1
2024 Multimodal Behaviour in an Online Environment: The GEHM Zoom Corpus Collection
abstract
This paper introduces a novel multimodal corpus consisting of 12 video recordings of Zoom meetings held in English by an international group of researchers from September 2021 to March 2023. The meetings have an average duration of about 40 minutes each, for a total of 8 hours. The number of participants varies from 5 to 9 per meeting. The participants’ speech was transcribed automatically using WhisperX, while visual coordinates of several keypoints of the participants’ head, their shoulders and wrists, were extracted using OpenPose. The audio-visual recordings will be distributed together with the orthographic transcription as well as the visual coordinates. In the paper we describe the way the corpus was collected, transcribed and enriched with the visual coordinates, we give descriptive statistics concerning both the speech transcription and the visual keypoint values and we present and discuss visualisations of these values. Finally, we carry out a short preliminary analysis of the role of feedback in the meetings, and show how visualising the coordinates extracted via OpenPose can be used to see how gestural behaviour supports the use of feedback words during the interaction.
Patrizia Paggio, Manex Agirrezabal, Costanza Navarretta, Leo Vitasovic
LREC/COLING1
2022 Letters From the Past: Modeling Historical Sound Change Through Diachronic Character Embeddings
abstract
While a great deal of work has been done on NLP approaches to lexical semantic change detection, other aspects of language change have received less attention from the NLP community. In this paper, we address the detection of sound change through historical spelling. We propose that a sound change can be captured by comparing the relative distance through time between the distributions of the characters involved before and after the change has taken place. We model these distributions using PPMI character embeddings. We verify this hypothesis in synthetic data and then test the method’s ability to trace the well-known historical change of lenition of plosives in Danish historical sources. We show that the models are able to identify several of the changes under consideration and to uncover meaningful contexts in which they appeared. The methodology has the potential to contribute to the study of open questions such as the relative chronology of sound shifts and their geographical distribution.
Sidsel Boldsen, Patrizia Paggio
ACL (1)2
2020 Dialogue Act Annotation in a Multimodal Corpus of First Encounter Dialogues
abstract
This paper deals with the annotation of dialogue acts in a multimodal corpus of first encounter dialogues, i.e. face-to- face dialogues in which two people who meet for the first time talk with no particular purpose other than just talking. More specifically, we describe the method used to annotate dialogue acts in the corpus, including the evaluation of the annotations. Then, we present descriptive statistics of the annotation, particularly focusing on which dialogue acts often follow each other across speakers and which dialogue acts overlap with gestural behaviour. Finally, we discuss how feedback is expressed in the corpus by means of feedback dialogue acts with or without co-occurring gestural behaviour, i.e. multimodal vs. unimodal feedback.
Costanza Navarretta, Patrizia Paggio
LREC2
2018 Classifying the Informative Behaviour of Emoji in Microblogs
Giulia Donato, Patrizia Paggio
LREC2
2018 Face2Text: Collecting an Annotated Image Description Corpus for the Generation of Rich Face Descriptions
Albert Gatt, Marc Tanti, Adrian Muscat, Patrizia Paggio, Reuben A. Farrugia, Claudia Borg, Kenneth P. Camilleri, Mike Rosner, Lonneke van der Plas
LREC4
2014 Learning when to point: A data-driven approach
Albert Gatt, Patrizia Paggio
COLING2
2012 Feedback in Nordic First-Encounters: a Comparative Study
Costanza Navarretta, Elisabeth Ahlsén, Jens Allwood, Kristiina Jokinen, Patrizia Paggio
LREC5
2012 Multimodal Behaviour and Feedback in Different Types of Interaction
Costanza Navarretta, Patrizia Paggio
LREC2
2010 The NOMCO Multimodal Nordic Resource - Goals and Characteristics
Patrizia Paggio, Jens Allwood, Elisabeth Ahlsén, Kristiina Jokinen, Costanza Navarretta
LREC1
2006 Information Structure and Pauses in a Corpus of Spoken Danish
Patrizia Paggio
EACL1
2006 Annotating Information Structure in a Corpus of Spoken Danish
Patrizia Paggio
LREC1
2004 Ontology-Based Question Answering in a Federation of University Sites: The MOSES Case Study
Paolo Atzeni, Roberto Basili 0001, Dorte Haltrup Hansen, Paolo Missier, Patrizia Paggio, Maria Teresa Pazienza, Fabio Massimo Zanzotto
NLDB5
2004 Content-based text querying with ontological descriptors
Troels Andreasen, Per Anker Jensen, Jørgen Fischer Nilsson, Patrizia Paggio, Bolette S. Pedersen, Hanne Erdman Thomsen
Data Knowl. Eng.4
2002 Semantic Lexical Resources Applied to Content-based Querying - the OntoQuery Project
Bolette S. Pedersen, Patrizia Paggio
LREC2
2002 Ontological Extraction of Content for Text Querying
Troels Andreasen, Per Anker Jensen, Jørgen Fischer Nilsson, Patrizia Paggio, Bolette S. Pedersen, Hanne Erdman Thomsen
NLDB4
1998 Evaluation in the SCARRIE project
Patrizia Paggio, Bradley Music
LREC1
1998 Validating the TEMAA LE evaluation methodology: a case study on Danish spelling checkers
Patrizia Paggio, Nancy L. Underwood
Nat. Lang. Eng.1
1991 A Preference Mechanism Based On Multiple Criteria Resolution
Ioannis Dologlou, Giovanni Malnati, Patrizia Paggio
EACL3