VLDB 2026 Research / reviewers in the wild / expert
Thierry Declerck
dblp:05/6145
· DBLP profile ↗
52ranked-venue papers
23as first author
8since 2021 · last 2023
0000-0002-9450-6648ORCID · verified
Domains — the database's venue-derived domains; a paper can count in several
Artificial intelligence and machine learning · 47 · 21 first-author · 8 since 2021Databases, data management, data science and information retrieval · 11 · 7 first-author · 5 since 2021Applied, interdisciplinary, general and emerging computing · 4 · 1 first-author · 2 since 2021Graphics, computer vision, multimedia, augmented reality and games · 3 · 1 first-author
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2023 | What's the Meaning of Superhuman Performance in Today's NLU?abstractSimone Tedeschi, Johan Bos, Thierry Declerck, Jan Hajič, Daniel Hershcovich, Eduard Hovy, Alexander Koller, Simon Krek, Steven Schockaert, Rico Sennrich, Ekaterina Shutova, Roberto Navigli. Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers). 2023. Simone Tedeschi, Johan Bos, Thierry Declerck, Jan Hajic 0001, Daniel Hershcovich, Eduard H. Hovy, Alexander Koller, Simon Krek, Steven Schockaert, Rico Sennrich, Ekaterina Shutova, Roberto Navigli |
ACL (1) | 3 |
| 2023 | A uniform RDF-based Representation of the Interlinking of Wordnets and Sign Language Data
Thierry Declerck, Sam Bigeard, Dorians Callus, Benjamin Matthews, Sussi Olsen, Loran Ripard Xuereb |
LDK | 1 |
| 2023 | Leveraging DBnary Data to Enrich Information of Multiword Terms in Wiktionary
Gilles Sérasset, Thierry Declerck, Lenka Bajcetic |
LDK | 2 |
| 2023 | Towards an RDF Representation of the Infrastructure consisting in using Wordnets as a conceptual Interlingua between multilingual Sign Language DatasetsabstractWe present ongoing work dealing with a Linked Data compliant representation of infrastructures using wordnets for connecting multilingual Sign Language data sets.We build for this on already existing RDF and OntoLex representations of Open Multilingual Wordnet (OMW) data sets and work done by the European EAS-IER research project on the use of the CSV files of OMW for linking glosses and basic semantic information associated with Sign Language data sets in two languages: German and Greek.In this context, we started the transformation into RDF of a Danish data set, which links Danish Sign Language data and the wordnet for Danish, DanNet.The final objective of our work is to include Sign Language data sets (and their conceptual cross-linking via wordnets) in the Linguistic Linked Open Data cloud. Thierry Declerck, Thomas Troelsgård, Sussi Olsen |
GWC | 1 |
| 2023 | Are there just WordNets or also SignNets?abstractFor Sign Languages (SLs), can we create a SignNet, like a WordNet for spoken languages: a network of semantic relations between constitutive elements of SLs?We first discuss approaches that link SL data to wordnets, or integrate such elements with some adaptations into the structure of WordNet.Then, we present requirements for a SignNet, which is built on SL data and then linked to WordNet. Ineke Schuurman, Thierry Declerck, Caro Brosens, Margot Janssens, Vincent Vandeghinste, Bram Vanroy |
GWC | 2 |
| 2022 | Using Wiktionary to Create Specialized Lexical Resources and DatasetsabstractThis paper describes an approach aiming at utilizing Wiktionary data for creating specialized lexical datasets which can be used for enriching other lexical (semantic) resources or for generating datasets that can be used for evaluating or improving NLP tasks, like Word Sense Disambiguation, Word-in-Context challenges, or Sense Linking across lexicons and dictionaries. We have focused on Wiktionary data about pronunciation information in English, and grammatical number and grammatical gender in German. Lenka Bajcetic, Thierry Declerck |
LREC | 2 |
| 2022 | Towards a new Ontology for Sign LanguagesabstractWe present the current status of a new ontology for representing constitutive elements of Sign Languages (SL). This development emerged from investigations on how to represent multimodal lexical data in the OntoLex-Lemon framework, with the goal to publish such data in the Linguistic Linked Open Data (LLOD) cloud. While studying the literature and various sites dealing with sign languages, we saw the need to harmonise all the data categories (or features) defined and used in those sources, and to organise them in an ontology to which lexical descriptions in OntoLex-Lemon could be linked. We make the code of the first version of this ontology available, so that it can be further developed collaboratively by both the Linked Data and the SL communities Thierry Declerck |
LREC | 1 |
| 2021 | Towards the Addition of Pronunciation Information to Lexical Semantic ResourcesabstractThis paper describes ongoing work aiming at adding pronunciation information to lexical semantic resources, with a focus on open wordnets. Our goal is not only to add a new modality to those semantic networks, but also to mark heteronyms listed in them with the pronunciation information associated with their different meanings. This work could contribute in the longer term to the disambiguation of multi-modal resources, which are combining text and speech. Thierry Declerck, Lenka Bajcetic |
GWC | 1 |
| 2020 | A Multilingual Evaluation Dataset for Monolingual Word Sense AlignmentabstractAligning senses across resources and languages is a challenging task with beneficial applications in the field of natural language processing and electronic lexicography. In this paper, we describe our efforts in manually aligning monolingual dictionaries. The alignment is carried out at sense-level for various resources in 15 languages. Moreover, senses are annotated with possible semantic relationships such as broadness, narrowness, relatedness, and equivalence. In comparison to previous datasets for this task, this dataset covers a wide range of languages and resources and focuses on the more challenging task of linking general-purpose language. We believe that our data will pave the way for further advances in alignment and evaluation of word senses by creating new solutions, particularly those notoriously requiring data such as neural networks. Our resources are publicly available at https://github.com/elexis-eu/MWSA. Sina Ahmadi, John P. McCrae, Sanni Nimb, Anas Fahad Khan, Monica Monachini, Bolette S. Pedersen, Thierry Declerck, Tanja Wissik, Andrea Bellandi, Irene Pisani, Thomas Troelsgård, Sussi Olsen, Simon Krek, Veronika Lipp, Tamás Váradi, László Simon, András Gyorffy, Carole Tiberius, Tanneke Schoonheim, Yifat Ben Moshe, Maya Rudich, Raya Abu Ahmad, Dorielle Lonke, Kira Kovalenko, Margit Langemets, Jelena Kallas, Oksana Dereza, Theodorus Fransen, David Cillessen, David Lindemann, Mikel Alonso, Ana Salgado, José-Luis Sancho-Gómez, Rafael-J. Ureña-Ruiz, Jordi Porta-Zamorano, Kiril Ivanov Simov, Petya Osenova, Zara Kancheva, Ivaylo Radev, Ranka Stankovic, Andrej Perdih, Dejan Gabrovsek |
LREC | 7 |
| 2020 | Recent Developments for the Linguistic Linked Open Data InfrastructureabstractIn this paper we describe the contributions made by the European H2020 project “Prêt-à-LLOD” (‘Ready-to-use Multilingual Linked Language Data for Knowledge Services across Sectors’) to the further development of the Linguistic Linked Open Data (LLOD) infrastructure. Prêt-à-LLOD aims to develop a new methodology for building data value chains applicable to a wide range of sectors and applications and based around language resources and language technologies that can be integrated by means of semantic technologies. We describe the methods implemented for increasing the number of language data sets in the LLOD. We also present the approach for ensuring interoperability and for porting LLOD data sets and services to other infrastructures, as well as the contribution of the projects to existing standards. Thierry Declerck, John P. McCrae, Matthias Hartung, Jorge Gracia, Christian Chiarcos, Elena Montiel-Ponsoda, Philipp Cimiano, Artem Revenko, Roser Saurí, Deirdre Lee, Stefania Racioppa, Jamal Abdul Nasir, Matthias Orlikowski, Marta Lanau-Coronas, Christian Fäth, Mariano Rico, Mohammad Fazleh Elahi, Maria Khvalchik, Meritxell González, Katharine Cooney |
LREC | 1 |
| 2020 | Language Data Sharing in European Public Services - Overcoming Obstacles and Creating Sustainable Data Sharing InfrastructuresabstractData is key in training modern language technologies. In this paper, we summarise the findings of the first pan-European study on obstacles to sharing language data across 29 EU Member States and CEF-affiliated countries carried out under the ELRC White Paper action on Sustainable Language Data Sharing to Support Language Equality in Multilingual Europe. Why Language Data Matters. We present the methodology of the study, the obstacles identified and report on recommendations on how to overcome those. The obstacles are classified into (1) lack of appreciation of the value of language data, (2) structural challenges, (3) disposition towards CAT tools and lack of digital skills, (4) inadequate language data management practices, (5) limited access to outsourced translations, and (6) legal concerns. Recommendations are grouped into addressing the European/national policy level, and the organisational/institutional level. Lilli Smal, Andrea Lösch, Josef van Genabith, Maria Giagkou, Thierry Declerck, Stephan Busemann |
LREC | 5 |
| 2019 | Towards the Detection and Formal Representation of Semantic Shifts in Inflectional Morphology
Dagmar Gromann, Thierry Declerck |
LDK | 2 |
| 2019 | OntoLex as a possible Bridge between WordNets and full lexical DescriptionsabstractIn this paper we describe our current work on representing a recently created German lexical semantics resource in OntoLex-Lemon and in conformance with Word-Net specifications.Besides presenting the representation effort, we show the utilization of OntoLex-Lemon to bridge from WordNet-like resources to full lexical descriptions and extend the coverage of WordNets to other types of lexical data, such as decomposition results, exemplified for German data, and inflectional phenomena, here outlined for English data. Thierry Declerck, Melanie Siegel |
GWC | 1 |
| 2018 | An Integrated Formal Representation for Terminological and Lexical Data included in Classification Schemes
Thierry Declerck, Kseniya Egorova, Eileen Schnur |
LREC | 1 |
| 2018 | Comparing Pretrained Multilingual Word Embeddings on an Ontology Alignment Task
Dagmar Gromann, Thierry Declerck |
LREC | 2 |
| 2018 | European Language Resource Coordination: Collecting Language Resources for Public Sector Multilingual Information Management
Andrea Lösch, Valérie Mapelli, Stelios Piperidis, Andrejs Vasiljevs, Lilli Smal, Thierry Declerck, Eileen Schnur, Khalid Choukri, Josef van Genabith |
LREC | 6 |
| 2016 | Monolingual Social Media Datasets for Detecting Contradiction and Entailment
Piroska Lendvai, Isabelle Augenstein, Kalina Bontcheva, Thierry Declerck |
LREC | 4 |
| 2016 | The Open Linguistics Working Group: Developing the Linguistic Linked Open Data Cloud
John P. McCrae, Christian Chiarcos, Francis Bond, Philipp Cimiano, Thierry Declerck, Gerard de Melo, Jorge Gracia, Sebastian Hellmann 0001, Bettina Klimek, Steven Moran, Petya Osenova, Antonio Pareja-Lora, Jonathan Pool |
LREC | 5 |
| 2016 | Towards a WordNet based Classification of Actors in FolktalesabstractIn the context of a student software project we are investigating the use of Word-Net for improving the automatic detection and classification of actors (or characters) mentioned in folktales.Our starting point is the book "Classification of International Folktales", out of which we extract text segments that name the different actors involved in tales, taking advantage of patterns used by its author, Hans-Jörg Uther.We apply on those text segments functions that are implemented in the NLTK interface to WordNet in order to obtain lexical semantic information to enrich the original naming of characters proposed in the "Classification of International Folktales" and to support their translation in other languages. Thierry Declerck, Tyler Klement, Antonia Kostova |
GWC | 1 |
| 2014 | Harmonization of German Lexical Resources for Opinion Mining
Thierry Declerck, Hans-Ulrich Krieger |
LREC | 1 |
| 2014 | A SKOS-based Schema for TEI encoded Dictionaries at ICLTT
Thierry Declerck, Karlheinz Mörth, Eveline Wandl-Vogt |
LREC | 1 |
| 2014 | TMO ― The Federated Ontology of the TrendMiner Project
Hans-Ulrich Krieger, Thierry Declerck |
LREC | 2 |
| 2013 | The DDI corpus: An annotated corpus with pharmacological substances and drug-drug interactions
María Herrero-Zazo, Isabel Segura-Bedmar, Paloma Martínez, Thierry Declerck |
J. Biomed. Informatics | 4 |
| 2013 | Trends in semantic and digital media technologiesabstractThe objective of this special issue is to report on recent trends in digital and media technologies responding to the challenges of managing and accessing multimedia (images, audio, video, 3D/4D material, etc.).In a highly selective review procedure we accepted contributions describing recent work that aims at narrowing the large disparity between the low-level multimedia descriptors and the richness of subjectivity of semantics in user queries and human interpretations of audiovisual data.The articles in this special issue can be grouped into following categories: Multimedia Analysis (Section 1), Multimedia Ontologies and Data Integration (Section 2), and Social Media and Retrieval (Section 3). Marcin Grzegorzek, Michael Granitzer, Stefan M. Rüger, Michael Sintek, Thierry Declerck, Massimo Romanelli |
Multim. Tools Appl. | 5 |
| 2012 | MFO-The Federated Financial Ontology for the MONNET Project
Hans-Ulrich Krieger, Thierry Declerck, Ashok Kumar Nedunchezhian |
KEOD | 2 |
| 2012 | Accessing and standardizing Wiktionary lexical entries for the translation of labels in Cultural Heritage taxonomies
Thierry Declerck, Karlheinz Mörth, Piroska Lendvai |
LREC | 1 |
| 2012 | The META-SHARE Metadata Schema for the Description of Language Resources
Maria Gavrilidou, Penny Labropoulou, Elina Desipri, Stelios Piperidis, Harris Papageorgiou, Monica Monachini, Francesca Frontini, Thierry Declerck, Gil Francopoulo, Victoria Arranz, Valérie Mapelli |
LREC | 8 |
| 2011 | A Text Technology Infrastructure for Annotating Corpora in the eHumanities
Thierry Declerck, Ulrike Czeitschner, Karlheinz Mörth, Claudia Resch, Gerhard Budin |
TPDL | 1 |
| 2011 | Linguistic and Semantic Representation of the Thompson's Motif-Index of Folk-Literature
Thierry Declerck, Piroska Lendvai |
TPDL | 1 |
| 2010 | Towards a Standardized Linguistic Annotation of the Textual Content of Labels in Knowledge Representation Systems
Thierry Declerck, Piroska Lendvai |
LREC | 1 |
| 2010 | Extraction, Merging, and Monitoring of Company Data from Heterogeneous Sources
Christian Federmann, Thierry Declerck |
LREC | 2 |
| 2010 | LAF/GrAF-grounded Representation of Dependency Structures
Yoshihiko Hayashi, Thierry Declerck, Chiharu Narawa |
LREC | 2 |
| 2010 | Integration of Linguistic Markup into Semantic Models of Folk Narratives: The Fairy Tale Use Case
Piroska Lendvai, Thierry Declerck, Sándor Darányi, Pablo Gervás, Raquel Hervás, Scott A. Malec, Federico Peinado |
LREC | 2 |
| 2008 | Ontology-Driven Human Language Technology for Semantic-Based Business IntelligenceabstractIn this poster submission, we describe the actual state of development of textual analysis and ontology-based information extraction in real world applications, as they are defined in the context of the European R&D project “MUSING” dealing with Business Intelligence. We present in some details the actual state of ontology development, including a time and domain ontologies, which are guiding information extraction onto an ontology population task. Thierry Declerck, Hans-Ulrich Krieger, Horacio Saggion, Marcus Spies |
ECAI | 1 |
| 2008 | A Hybrid Reasoning Architecture for Business Intelligence ApplicationsabstractWe describe an implemented hybrid reasoning architecture that is used in an EU-funded project called MUSING (www.musing.eu) which is dedicated towards the investigation of semantic-based business intelligence solutions. The reasoning platform builds on publicly available software, such as Pellet, OWLIM, Jena, and Sesame. The project uses and extends existing OWL ontologies (e.g., PROTON) and assumes rule-based reasoning to take place on top of OWL. We describe the pros and cons of each subsystem w.r.t. the needs we have encountered during our investigation. We explain the specific reasoning architecture that is based on a sequence and a fixpoint computation of three reasoners which we might sloppily write as Pellet + (OWLIM + Jena)^*. Pellet is used for checking the initial consistency of the ontology, whereas OWLIM and Jena are employed to execute rules outside the expressiveness of OWL. However, OWLIM is way much faster than Jena, but neither has means to do numerical comparison nor arithmetic. We explain our choice why SWRL is not enough and why we believe that binary OWL properties lead to an unwanted proliferation of objects, making representation and reasoning extremely complex. Hans-Ulrich Krieger, Bernd Kiefer, Thierry Declerck |
HIS | 3 |
| 2008 | Foundation of a Component-based Flexible Registry for Language Resources and Technology
Daan Broeder, Thierry Declerck, Erhard W. Hinrichs, Stelios Piperidis, Laurent Romary, Nicoletta Calzolari, Peter Wittenburg |
LREC | 2 |
| 2008 | A Framework for Standardized Syntactic Annotation
Thierry Declerck |
LREC | 1 |
| 2006 | SynAF: Towards a Standard for Syntactic Annotation
Thierry Declerck |
LREC | 1 |
| 2006 | Multilingual Lexical Semantic Resources for Ontology Translation
Thierry Declerck, Asunción Gómez-Pérez, Ovidiu Vela, Zeno Gantner, David Manzano-Macho |
LREC | 1 |
| 2006 | Generic NLP Tools for Supporting Shallow Ontology Building
Thierry Declerck, Mihaela Vela |
LREC | 1 |
| 2004 | A Large Metadata Domain of Language Resources
Daan Broeder, Thierry Declerck, Laurent Romary, Markus Uneson, Sven Strömqvist, Peter Wittenburg |
LREC | 2 |
| 2004 | Towards Ontology Engineering Based on Linguistic Analysis
Paul Buitelaar, Daniel Olejnik, Mihaela Hutanu, Alexander Schutz, Thierry Declerck, Michael Sintek |
LREC | 5 |
| 2004 | Towards a Language Infrastructure for the Semantic Web
Thierry Declerck, Paul Buitelaar, Nicoletta Calzolari, Alessandro Lenci |
LREC | 1 |
| 2004 | Towards an International Standard on Feature Structure Representation
Kiyong Lee, Lou Burnard, Laurent Romary, Éric Villemonte de la Clergerie, Thierry Declerck, Syd Bauman, Harry Bunt, Lionel Clément, Tomaz Erjavec, Azim Roussanaly, Claude Roux |
LREC | 5 |
| 2003 | Event-Coreference across Multiple, Multi-lingual Sources in the Mumis Project
Horacio Saggion, Hamish Cunningham, Jan Kuper, Thierry Declerck, Peter Wittenburg |
EACL | 4 |
| 2003 | Intelligent Multimedia Indexing and Retrieval through Multi-source Information Extraction and Merging
Jan Kuper, Horacio Saggion, Hamish Cunningham, Thierry Declerck, Franciska de Jong, Dennis Reidsma, Yorick Wilks, Peter Wittenburg |
IJCAI | 4 |
| 2002 | LREP: A Language Repository Exchange Protocol
Daan Broeder, Peter Wittenburg, Thierry Declerck, Laurent Romary |
LREC | 3 |
| 2002 | COLLATE: Competence Center in Speech and Language Technology
Joanne Capstick, Hans Uszkoreit, Wolfgang Wahlster, Thierry Declerck, Gregor Erbach, Anthony Jameson, Brigitte Jörg, Reinhard Karger, Tillmann Wegst |
LREC | 4 |
| 2000 | The New Edition of the Natural Language Software Registry (an Initiative of ACL hosted at DFKI)
Thierry Declerck, Alexander Werner Jachmann, Hans Uszkoreit |
LREC | 1 |
| 1998 | Evaluation of the NPL components of an information extraction system for german
Thierry Declerck, Judith Klein, Günter Neumann |
LREC | 1 |
| 1996 | Dealing with Cross-Sentential Anaphora Resolution in ALEP
Thierry Declerck |
COLING | 1 |
| 1996 | Lean Formalisms, Linguistic Theory and Applications. Grammar Development in ALEP
Paul Schmidt, Axel Theofilidis, Sibylle Rieder, Thierry Declerck |
COLING | 4 |