| Challenge: | CLARIN Concept Registry supports semantic interoperability, but does not extend beyond it. |
| Approach: | They propose to ground CMDI-based metadata using the CLARIN concept registry . they propose to use schema.org to map CMDi-based profiles to schema.com terms . |
| Outcome: | The proposed tool can map CMDI-based profiles to schema.org terms . the tool can be used to map CCR-based metadata to schema ontologies . |
Similar Papers
Interoperability in an Infrastructure Enabling Multidisciplinary Research: The case of CLARIN (2020.lrec-1)
Copied to clipboard
| Challenge: | CLARIN supports the use and study of language data in general and aims to increase the potential for comparative research of cultural and societal phenomena across languages and disciplines. |
| Approach: | They describe the interoperability requirements that arise through the existing ambitions and emerging frameworks. |
| Outcome: | The proposed frameworks will address interoperability requirements at several levels, including organisation and ecosystem, design of workflow services, data curation, performance measurement and collaboration. |
Metadata Collection Records for Language Resources (L18-1)
Copied to clipboard
| Challenge: | a pilot project aimed at bringing metadata records to the CLARIN context has been conducted . a virtual language observatory (VLO) was developed to provide an entry point to the language resources available in the infrastructure. |
| Approach: | They propose to implement a CMDI profile for Dutch language resources . they propose an interface for creating, editing, listing, copying and exporting metadata records . |
| Outcome: | The proposed interface is validated in a pilot with 45 Dutch language resources . the proposed interface provides a user interface for creating, editing, listing, copying and exporting descriptions of metadata collection records. |
Annotation Interoperability for the Post-ISOCat Era (2020.lrec-1)
Copied to clipboard
| Challenge: | Using ISOCat successor solutions, annotation standards have been developed since 2010 . |
| Approach: | They describe ISOCat successor solutions and annotation standardization efforts since 2010 . they describe low-cost harmonization of post-ISOCat vocabularies by means of linked ontologies . |
| Outcome: | The proposed ontologies are linked with the Ontologie of Linguistic Annotation and ISOCat, the GOLD ontology, the Typological Database Systems ontological and a large number of annotation schemes. |
CLARIN’s Key Resource Families (L18-1)
Copied to clipboard
| Challenge: | CLARIN is a European Research Infrastructure that supports the accessibility of language resources and technologies to researchers from the Digital Humanities and Social Sciences. |
| Approach: | They propose to present key resource families in a uniform way for researchers to use in their research using the CLARIN infrastructure. |
| Outcome: | The key resource families are newspaper, parliamentary, CMC (computer-mediated communication), and parallel corpora. |
AMR Beyond the Sentence: the Multi-sentence AMR corpus (C18-1)
Copied to clipboard
| Challenge: | Abstract Meaning Representation (AMR) is limited to capturing the semantics of individual sentences. |
| Approach: | They propose a corpus that annotates coreference and similar phenomena on top of existing AMRs. |
| Outcome: | The proposed corpus is compared with existing corpora on sentence-level semantics . it shows that it can be used for information extraction and question answering . |
entity-linkings: A Unified Library for Entity Linking (2026.eacl-demo)
Copied to clipboard
| Challenge: | Entity linking (EL) is the task of mapping named entities in text to canonical entries in a knowledge base. |
| Approach: | They propose a unified library for using and developing entity linking systems . a strong emphasis is placed on usability, making it highly extensible . |
| Outcome: | a new library aims to disambiguate named entities in text by mapping them to canonical entries in a knowledge base. |
A Survey of AMR Applications (2024.emnlp-main)
Copied to clipboard
| Challenge: | Abstract Meaning Representation (AMR) is a semantic representation that takes the form of a rooted, directed graph. |
| Approach: | They analyze more than 100 papers which use Abstract Meaning Representation (AMR) they highlight the range of applications for which AMR has been harnessed and techniques for incorporating it . they also highlight broader AMR engineering patterns and outline areas of future work that seem ripe for AMR incorporation. |
| Outcome: | The results highlight the range of applications for which AMR has been harnessed and the techniques for incorporating it into those applications. |
Towards Entity Spaces (2020.lrec-1)
Copied to clipboard
| Challenge: | Entities are a central element of knowledge bases and are used in many knowledge-centric tasks including text analysis. |
| Approach: | They propose to use entity spaces to represent a set of associated entities with near-identity to provide a handle to an amorphous grouping of entities. |
| Outcome: | The proposed representations improve recall of entity linking in English by using disambiguation pages. |
Quantification Annotation in ISO 24617-12, Second Draft (2022.lrec-1)
Copied to clipboard
Harry Bunt, Maxime Amblard, Johan Bos, Karën Fort, Bruno Guillaume, Philippe de Groote, Chuyuan Li, Pierre Ludmann, Michel Musiol, Siyana Pavlova, Guy Perrier, Sylvain Pogodalla
| Challenge: | a project aimed at establishing an interoperable annotation schema for quantification phenomena was relaunched in early 2022 due to the Covid-19 pandemic . |
| Approach: | This paper describes the continuation of a project that aims at establishing an interoperable annotation schema for quantification phenomena as part of the ISO suite of semantic annotation standards. |
| Outcome: | The proposed schema is part of the ISO suite of semantic annotation standards known as the Semantic Annotation Framework (SemAF). |
FAIRification of LeiLanD (2024.lrec-main)
Copied to clipboard
| Challenge: | LeiLanD is a searchable catalogue of language datasets collected at LUCL and other institutes of Leiden University. |
| Approach: | They propose to use a standardised metadata format called CMDI to improve the findability of Leiden language datasets. |
| Outcome: | The proposed catalogue has enhanced the findability and accessibility of incredibly diverse datasets. |