Sussi Olsen

dblp:56/8155 · DBLP profile ↗
← Back
7ranked-venue papers in the field
0as first author
4since 2021 · last 2023
—ORCID · none

Domains — venue-derived; a paper can count in several

Other / Interdisciplinary · 6Knowledge Engineering, Semantic Web & Information Systems · 1
YearPublicationVenuePosition
2023 A uniform RDF-based Representation of the Interlinking of Wordnets and Sign Language Data
Thierry Declerck, Sam Bigeard, Dorians Callus, Benjamin Matthews, Sussi Olsen, Loran Ripard Xuereb
LDK5
2023 Towards an RDF Representation of the Infrastructure consisting in using Wordnets as a conceptual Interlingua between multilingual Sign Language Datasets
abstract
We present ongoing work dealing with a Linked Data compliant representation of infrastructures using wordnets for connecting multilingual Sign Language data sets.We build for this on already existing RDF and OntoLex representations of Open Multilingual Wordnet (OMW) data sets and work done by the European EAS-IER research project on the use of the CSV files of OMW for linking glosses and basic semantic information associated with Sign Language data sets in two languages: German and Greek.In this context, we started the transformation into RDF of a Danish data set, which links Danish Sign Language data and the wordnet for Danish, DanNet.The final objective of our work is to include Sign Language data sets (and their conceptual cross-linking via wordnets) in the Linguistic Linked Open Data cloud.
Thierry Declerck, Thomas Troelsgård, Sussi Olsen
GWC3
2023 Reusing the Danish WordNet for a New Central Word Register for Danish - a Project Report
abstract
In this paper we report on a new Danish lexical initiative, the Central Word Register for Danish, (COR), which aims at providing an open-source, well curated and large-coverage lexicon for AI purposes.The semantic part of the lexicon (COR-S) relies to a large extent on the lexical-semantic information provided in the Danish wordnet, DanNet.However, we have taken the opportunity to evaluate and curate the wordnet information while compiling the new resource.Some information types have been simplified and more systematically curated.This is the case for the hyponymy relations, the ontological typing, and the sense inventory, i.e. the treatment of polysemy, including systematic polysemy.
Bolette S. Pedersen, Sanni Nimb, Nathalie Carmen Hau Sørensen, Sussi Olsen, Ida Flørke, Thomas Troelsgård
GWC4
2021 DanNet2: Extending the coverage of adjectives in DanNet based on thesaurus data (project presentation)
abstract
The paper describes work in progress in the DanNet2 project financed by the Carlsberg Foundation.The project aim is to extend the original Danish wordnet, DanNet, in several ways.Main focus is on extension of the coverage and description of the adjectives, a part of speech that was rather sparsely described in the original wordnet.We describe the methodology and initial work of semiautomatically transferring adjectives from the Danish Thesaurus to the wordnet with the aim of easily enlarging the coverage from 3,000 to approx.13,000 adjectival synsets.Transfer is performed by manually encoding all missing adjectival subsection headwords from the thesaurus and thereafter employing a semiautomatic procedure where adjectives from the same subsection are transferred to the wordnet as either 1) near synonyms to the section's headword, 2) hyponyms to the section's headword, or 3) as members of the same synset as the headword.We also discuss how to deal with the problem of multiple representations of the same sense in the thesaurus, and present other types of information from the thesaurus that we plan to integrate, such as thematic and sentiment information.
Sanni Nimb, Bolette S. Pedersen, Sussi Olsen
GWC3
2019 Merging DanNet with Princeton Wordnet
abstract
In this paper we describe the merge of the Danish wordnet, DanNet, with Princeton Wordnet applying a two-step approach.We first link from the English Princeton core to Danish (5,000 base concepts) and then proceed to linking the rest of the Danish vocabulary to English, thus going from Danish to English.Since the Danish wordnet is built bottom-up from Danish lexica and corpora, all taxonomies are monolingually based and thus not necessarily directly compatible with the coverage and structure of the Princeton WordNet.This fact proves to pose some challenges to the linking procedure since a considerable number of the links cannot be realised via the preferred crosslanguage synonym link which implies a more or less precise correlation between the two concepts.Instead, a subpart of the links are realised through near synonym or hyponymy links to compensate for the fact that no precise translation can be found in the target resource.The tool WordnetLoom is currently used for manual linking but procedures for a more automatic procedure in future is discussed.We conclude that the two resources actually differ from each other quite more than expected, both vocabulary-and structure-wise.
Bolette S. Pedersen, Sanni Nimb, Ida Rørmann Olsen, Sussi Olsen
GWC4
2018 Towards a principled approach to sense clustering - a case study of wordnet and dictionary senses in Danish
abstract
Our aim is to develop principled methods for sense clustering which can make existing lexical resources practically useful in NLPnot too fine-grained to be operational and yet finegrained enough to be worth the trouble.Where traditional dictionaries have a highly structured sense inventory typically describing the vocabulary by means of main-and subsenses, wordnets are generally fine-grained and unstructured.We present a series of clustering and annotation experiments with 10 of the most polysemous nouns in Danish.We combine the structured information of a traditional Danish dictionary with the ontological types found in the Danish wordnet, DanNet.This constellation enables us to automatically cluster senses in a principled way and improve inter-annotator agreement and wsd performance.
Bolette S. Pedersen, Manex Agirrezabal, Sanni Nimb, Ida Rørmann Olsen, Sussi Olsen
GWC5
2016 An empirically grounded expansion of the supersense inventory
abstract
In this article we present an expansion of the supersense inventory.All new supersenses are extensions of members of the current inventory, which we postulate by identifying semantically coherent groups of synsets.We cover the expansion of the already-established supernsense inventory for nouns and verbs, the addition of coarse supersenses for adjectives in absence of a canonical supersense inventory, and supersenses for verbal satellites.We evaluate the viability of the new senses examining the annotation agreement, frequency and co-ocurrence patterns.
Héctor Martínez Alonso, Anders Johannsen, Sanni Nimb, Sussi Olsen, Bolette S. Pedersen
GWC4