Papers with LLOD
An Integrated Formal Representation for Terminological and Lexical Data included in Classification Schemes (L18-1)
Copied to clipboard
| Challenge: | e-lexicography is a field of study dealing with the automated creation of specialized multilingual dictionaries from structured data. |
| Approach: | They propose to use a SKOS-XL vocabulary for modelling the multilingual terminological part of comparable taxonomies and OntoLex-Lemon for modelling multilingual lexical entries. |
| Outcome: | The proposed model can be explicitly cross-linked in the context of the Linguistic Linked Open Data (LLOD). |
ISO-based Annotated Multilingual Parallel Corpus for Discourse Markers (2022.lrec-1)
Copied to clipboard
Purificação Silvano, Mariana Damova, Giedrė Valūnaitė Oleškevičienė, Chaya Liebeskind, Christian Chiarcos, Dimitar Trajanov, Ciprian-Octavian Truică, Elena-Simona Apostol, Anna Baczkowska
| Challenge: | Discourse markers carry information about the discourse structure and organization, and also signal local dependencies or epistemic stance of speaker. |
| Approach: | They propose an ISO-based annotated multilingual parallel corpus for discourse markers . they propose an annotation scheme for discourse relations with a plug-in to ISO 24617-2 . |
| Outcome: | The proposed language resource is based on an ISO-based annotated multilingual parallel corpus of discourse markers. |
Croatian Idioms Integration: Enhancing the LIdioms Multilingual Linked Idioms Dataset (2024.lrec-main)
Copied to clipboard
| Challenge: | Existing datasets that include idioms from English, German, Italian, Portuguese and Russian do not include a comprehensive representation of idiomatic expressions in Croatian. |
| Approach: | They propose to extend existing RDF-based multilingual representation of idioms to include 1,042 Croatian idiomes in an Ontolex Lemon format. |
| Outcome: | The proposed resource includes 1,042 Croatian idioms in an Ontolex Lemon format to foster translation initiatives and facilitate intercultural exchange. |
Towards a Linked Open Data Edition of Sumerian Corpora (L18-1)
Copied to clipboard
| Challenge: | Linguistic Linked Open Data (LLOD) is a flourishing line of research in the language resource community . existing LLOD standards and vocabularies are not widely used in this community despite its popularity . |
| Approach: | They propose to use Linguistic Linked Open Data to link a Sumerian corpus with lexical resources . they use a linguistically annotated archive to create a corpus of cuneiform texts . |
| Outcome: | The proposed LLOD framework is used in assyriology, with philological resources underrepresented . the proposed framework is based on a linguistically annotated corpus of Sumerian texts . |
Towards a new Ontology for Sign Languages (2022.lrec-1)
Copied to clipboard
| Challenge: | Linked Data (LD) compliant datasets for sign languages are not available in the LLOD cloud. |
| Approach: | They propose to create an ontology for representing constitutive elements of Sign Languages (SL) they propose to publish such data in the Linguistic Linked Open Data cloud. |
| Outcome: | The proposed ontology can be used to represent sign languages in the Linguistic Linked Open Data cloud. |
The Index Thomisticus Treebank as Linked Data in the LiLa Knowledge Base (2022.lrec-1)
Copied to clipboard
| Challenge: | a series of Latin treebanks with word-by-word account of syntax and morphology of Latin texts have been published only in recent years. |
| Approach: | They propose to publish Latin treebanks that contain morphology and syntax annotations . they propose to use principles of the Linguistic Linked Open Data community . |
| Outcome: | The proposed approach enables interoperability between corpora and lexical resources for Latin . language learning and corpus-based research are the most obvious applications . |
From Linguistic Linked Data to Big Data (2024.lrec-main)
Copied to clipboard
Dimitar Trajanov, Elena Apostol, Radovan Garabik, Katerina Gkirtzou, Dagmar Gromann, Chaya Liebeskind, Cosimo Palma, Michael Rosner, Alexia Sampri, Gilles Sérasset, Blerina Spahiu, Ciprian-Octavian Truică, Giedre Valunaite Oleskeviciene
| Challenge: | Language data on the LOD cloud has grown in number, size, and variety . Linked (Open) Data (LLOD) is a standardized way of representing and sharing linguistic datasets . |
| Approach: | They propose to combine LLOD and Big Data to improve interoperability of linguistic datasets . they propose to use a machine-readable format to represent and share linguistic data . |
| Outcome: | This paper examines the use cases of Linked (Open) Data and Big Data in language data. |
Recent Developments for the Linguistic Linked Open Data Infrastructure (2020.lrec-1)
Copied to clipboard
Thierry Declerck, John Philip McCrae, Matthias Hartung, Jorge Gracia, Christian Chiarcos, Elena Montiel-Ponsoda, Philipp Cimiano, Artem Revenko, Roser Saurí, Deirdre Lee, Stefania Racioppa, Jamal Abdul Nasir, Matthias Orlikowsk, Marta Lanau-Coronas, Christian Fäth, Mariano Rico, Mohammad Fazleh Elahi, Maria Khvalchik, Meritxell Gonzalez, Katharine Cooney
| Challenge: | Language data is rarely 'ready-to-use' and language technology specialists spend over 80% of their time cleaning, organizing and collecting language datasets. |
| Approach: | They propose a methodology for building data value chains based around language resources and language technologies that can be integrated by means of semantic technologies. |
| Outcome: | The proposed methodology is based on language resources and language technologies that can be integrated by means of semantic technologies. |
Interoperability of Language-related Information: Mapping the BLL Thesaurus to Lexvo and Glottolog (L18-1)
Copied to clipboard
| Challenge: | The Bibliography of Linguistic Literature (BLL Thesaurus) has been used since 2013 in the context of the Lin gu is tik portal, a hub for linguistically relevant information. |
| Approach: | They propose to use Lexvo and Glottolog to facilitate interoperability between the BLL Thesaurus and terminological repositories in the Linguistic Linked Open Data cloud. |
| Outcome: | The proposed model is based on Lexvo and Glottolog and is able to connect to the Linguistic Linked Open Data cloud. |