Papers with LLOD

9 papers
An Integrated Formal Representation for Terminological and Lexical Data included in Classification Schemes (L18-1)

Copied to clipboard

Challenge: e-lexicography is a field of study dealing with the automated creation of specialized multilingual dictionaries from structured data.
Approach: They propose to use a SKOS-XL vocabulary for modelling the multilingual terminological part of comparable taxonomies and OntoLex-Lemon for modelling multilingual lexical entries.
Outcome: The proposed model can be explicitly cross-linked in the context of the Linguistic Linked Open Data (LLOD).
ISO-based Annotated Multilingual Parallel Corpus for Discourse Markers (2022.lrec-1)

Copied to clipboard

Challenge: Discourse markers carry information about the discourse structure and organization, and also signal local dependencies or epistemic stance of speaker.
Approach: They propose an ISO-based annotated multilingual parallel corpus for discourse markers . they propose an annotation scheme for discourse relations with a plug-in to ISO 24617-2 .
Outcome: The proposed language resource is based on an ISO-based annotated multilingual parallel corpus of discourse markers.
Croatian Idioms Integration: Enhancing the LIdioms Multilingual Linked Idioms Dataset (2024.lrec-main)

Copied to clipboard

Challenge: Existing datasets that include idioms from English, German, Italian, Portuguese and Russian do not include a comprehensive representation of idiomatic expressions in Croatian.
Approach: They propose to extend existing RDF-based multilingual representation of idioms to include 1,042 Croatian idiomes in an Ontolex Lemon format.
Outcome: The proposed resource includes 1,042 Croatian idioms in an Ontolex Lemon format to foster translation initiatives and facilitate intercultural exchange.
Towards a Linked Open Data Edition of Sumerian Corpora (L18-1)

Copied to clipboard

Challenge: Linguistic Linked Open Data (LLOD) is a flourishing line of research in the language resource community . existing LLOD standards and vocabularies are not widely used in this community despite its popularity .
Approach: They propose to use Linguistic Linked Open Data to link a Sumerian corpus with lexical resources . they use a linguistically annotated archive to create a corpus of cuneiform texts .
Outcome: The proposed LLOD framework is used in assyriology, with philological resources underrepresented . the proposed framework is based on a linguistically annotated corpus of Sumerian texts .
Towards a new Ontology for Sign Languages (2022.lrec-1)

Copied to clipboard

Challenge: Linked Data (LD) compliant datasets for sign languages are not available in the LLOD cloud.
Approach: They propose to create an ontology for representing constitutive elements of Sign Languages (SL) they propose to publish such data in the Linguistic Linked Open Data cloud.
Outcome: The proposed ontology can be used to represent sign languages in the Linguistic Linked Open Data cloud.
The Index Thomisticus Treebank as Linked Data in the LiLa Knowledge Base (2022.lrec-1)

Copied to clipboard

Challenge: a series of Latin treebanks with word-by-word account of syntax and morphology of Latin texts have been published only in recent years.
Approach: They propose to publish Latin treebanks that contain morphology and syntax annotations . they propose to use principles of the Linguistic Linked Open Data community .
Outcome: The proposed approach enables interoperability between corpora and lexical resources for Latin . language learning and corpus-based research are the most obvious applications .
From Linguistic Linked Data to Big Data (2024.lrec-main)

Copied to clipboard

Challenge: Language data on the LOD cloud has grown in number, size, and variety . Linked (Open) Data (LLOD) is a standardized way of representing and sharing linguistic datasets .
Approach: They propose to combine LLOD and Big Data to improve interoperability of linguistic datasets . they propose to use a machine-readable format to represent and share linguistic data .
Outcome: This paper examines the use cases of Linked (Open) Data and Big Data in language data.
Recent Developments for the Linguistic Linked Open Data Infrastructure (2020.lrec-1)

Copied to clipboard

Challenge: Language data is rarely 'ready-to-use' and language technology specialists spend over 80% of their time cleaning, organizing and collecting language datasets.
Approach: They propose a methodology for building data value chains based around language resources and language technologies that can be integrated by means of semantic technologies.
Outcome: The proposed methodology is based on language resources and language technologies that can be integrated by means of semantic technologies.
Interoperability of Language-related Information: Mapping the BLL Thesaurus to Lexvo and Glottolog (L18-1)

Copied to clipboard

Challenge: The Bibliography of Linguistic Literature (BLL Thesaurus) has been used since 2013 in the context of the Lin gu is tik portal, a hub for linguistically relevant information.
Approach: They propose to use Lexvo and Glottolog to facilitate interoperability between the BLL Thesaurus and terminological repositories in the Linguistic Linked Open Data cloud.
Outcome: The proposed model is based on Lexvo and Glottolog and is able to connect to the Linguistic Linked Open Data cloud.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations