Increasing CMDI’s Semantic Interoperability with schema.org (2022.lrec-1)

Copied to clipboard

Challenge: CLARIN Concept Registry supports semantic interoperability, but does not extend beyond it.
Approach: They propose to ground CMDI-based metadata using the CLARIN concept registry . they propose to use schema.org to map CMDi-based profiles to schema.com terms .
Outcome: The proposed tool can map CMDI-based profiles to schema.org terms . the tool can be used to map CCR-based metadata to schema ontologies .

Similar Papers

Interoperability in an Infrastructure Enabling Multidisciplinary Research: The case of CLARIN (2020.lrec-1)

Copied to clipboard

Challenge: CLARIN supports the use and study of language data in general and aims to increase the potential for comparative research of cultural and societal phenomena across languages and disciplines.
Approach: They describe the interoperability requirements that arise through the existing ambitions and emerging frameworks.
Outcome: The proposed frameworks will address interoperability requirements at several levels, including organisation and ecosystem, design of workflow services, data curation, performance measurement and collaboration.
Metadata Collection Records for Language Resources (L18-1)

Copied to clipboard

Challenge: a pilot project aimed at bringing metadata records to the CLARIN context has been conducted . a virtual language observatory (VLO) was developed to provide an entry point to the language resources available in the infrastructure.
Approach: They propose to implement a CMDI profile for Dutch language resources . they propose an interface for creating, editing, listing, copying and exporting metadata records .
Outcome: The proposed interface is validated in a pilot with 45 Dutch language resources . the proposed interface provides a user interface for creating, editing, listing, copying and exporting descriptions of metadata collection records.
Annotation Interoperability for the Post-ISOCat Era (2020.lrec-1)

Copied to clipboard

Challenge: Using ISOCat successor solutions, annotation standards have been developed since 2010 .
Approach: They describe ISOCat successor solutions and annotation standardization efforts since 2010 . they describe low-cost harmonization of post-ISOCat vocabularies by means of linked ontologies .
Outcome: The proposed ontologies are linked with the Ontologie of Linguistic Annotation and ISOCat, the GOLD ontology, the Typological Database Systems ontological and a large number of annotation schemes.
CLARIN’s Key Resource Families (L18-1)

Copied to clipboard

Challenge: CLARIN is a European Research Infrastructure that supports the accessibility of language resources and technologies to researchers from the Digital Humanities and Social Sciences.
Approach: They propose to present key resource families in a uniform way for researchers to use in their research using the CLARIN infrastructure.
Outcome: The key resource families are newspaper, parliamentary, CMC (computer-mediated communication), and parallel corpora.
AMR Beyond the Sentence: the Multi-sentence AMR corpus (C18-1)

Copied to clipboard

Challenge: Abstract Meaning Representation (AMR) is limited to capturing the semantics of individual sentences.
Approach: They propose a corpus that annotates coreference and similar phenomena on top of existing AMRs.
Outcome: The proposed corpus is compared with existing corpora on sentence-level semantics . it shows that it can be used for information extraction and question answering .
entity-linkings: A Unified Library for Entity Linking (2026.eacl-demo)

Copied to clipboard

Challenge: Entity linking (EL) is the task of mapping named entities in text to canonical entries in a knowledge base.
Approach: They propose a unified library for using and developing entity linking systems . a strong emphasis is placed on usability, making it highly extensible .
Outcome: a new library aims to disambiguate named entities in text by mapping them to canonical entries in a knowledge base.
A Survey of AMR Applications (2024.emnlp-main)

Copied to clipboard

Challenge: Abstract Meaning Representation (AMR) is a semantic representation that takes the form of a rooted, directed graph.
Approach: They analyze more than 100 papers which use Abstract Meaning Representation (AMR) they highlight the range of applications for which AMR has been harnessed and techniques for incorporating it . they also highlight broader AMR engineering patterns and outline areas of future work that seem ripe for AMR incorporation.
Outcome: The results highlight the range of applications for which AMR has been harnessed and the techniques for incorporating it into those applications.
Towards Entity Spaces (2020.lrec-1)

Copied to clipboard

Challenge: Entities are a central element of knowledge bases and are used in many knowledge-centric tasks including text analysis.
Approach: They propose to use entity spaces to represent a set of associated entities with near-identity to provide a handle to an amorphous grouping of entities.
Outcome: The proposed representations improve recall of entity linking in English by using disambiguation pages.
Quantification Annotation in ISO 24617-12, Second Draft (2022.lrec-1)

Copied to clipboard

Challenge: a project aimed at establishing an interoperable annotation schema for quantification phenomena was relaunched in early 2022 due to the Covid-19 pandemic .
Approach: This paper describes the continuation of a project that aims at establishing an interoperable annotation schema for quantification phenomena as part of the ISO suite of semantic annotation standards.
Outcome: The proposed schema is part of the ISO suite of semantic annotation standards known as the Semantic Annotation Framework (SemAF).
FAIRification of LeiLanD (2024.lrec-main)

Copied to clipboard

Challenge: LeiLanD is a searchable catalogue of language datasets collected at LUCL and other institutes of Leiden University.
Approach: They propose to use a standardised metadata format called CMDI to improve the findability of Leiden language datasets.
Outcome: The proposed catalogue has enhanced the findability and accessibility of incredibly diverse datasets.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations