Challenge: European Language Resource Coordination (ELRC) initiated a number of actions to support the collection of Language Resources (LRs) within the public sector in EU member and CEF-affiliated countries.
Approach: They propose to initiate actions to support the collection of Language Resources (LRs) within the public sector in EU member and CEF-affiliated countries.
Outcome: The European Language Resource Coordination (ELRC) consortium initiated a number of actions to support the collection of Language Resources (LRs) within the public sector in EU member and CEF-affiliated countries.

Similar Papers

Managing Public Sector Data for Multilingual Applications Development (L18-1)

Copied to clipboard

Challenge: eTranslation is a digital service that enables multilingual communication across public administrations in 30 European countries.
Approach: They propose to develop a repository infrastructure specifically tailored to the needs of the eTranslation service of the European Commission.
Outcome: The ELRC-SHARE repository is designed and developed specifically for the eTranslation service of the European Commission.
Collecting Language Resources from Public Administrations in the Nordic and Baltic Countries (L18-1)

Copied to clipboard

Challenge: Several large-scale projects and initiatives have been undertaken in this century to collect language resources and create LR repositories and infrastructures on a pan-European scale.
Approach: They present the work of Tilde on collecting language resources from government institutions and other public administrations in the Nordic and Baltic countries.
Outcome: The results of the European Language Resources Coordination (ELRC) action in the Nordic and Baltic countries are presented.
Language Data Sharing in European Public Services – Overcoming Obstacles and Creating Sustainable Data Sharing Infrastructures (2020.lrec-1)

Copied to clipboard

Challenge: Data is key in training modern language technologies.
Approach: They summarise findings of first pan-European study on barriers to language data sharing . they identify structural challenges, disposition towards CAT tools and lack of digital skills . overcoming language barriers is one of the main challenges european citizens face .
Outcome: The paper summarises the findings of the first pan-European study on barriers to language data sharing . the findings highlight the barriers and recommend solutions to overcome them .
Language Resources to Support Language Diversity – the ELRA Achievements (2022.lrec-1)

Copied to clipboard

Challenge: ELRA and its operational agency ELDA have continued to increase their catalogue of Language Resources (LRs) over the past few years, ELLA and ELTA have contributed to improve the access to multilingual information in the context of the pandemic .
Approach: ELRA and its operational agency ELDA have increased their catalogue of Language Resources (LRs) over the past few years, they have established partnerships for the production of various types of LRs.
Outcome: ELRA and its operational agency ELDA have contributed to improve the access to multilingual information in the context of the pandemic, develop tools for the de-identification of texts in the legal and medical domains, and support the EU eTranslation Machine Translation system.
ELRC Action: Covering Confidentiality, Correctness and Cross-linguality (2022.lrec-1)

Copied to clipboard

Challenge: ELRC aims to reduce language barriers by assessing language technology (LT) specifications . automated anonymisation and multilingual fake news processing are two of the most extensive LT assessments .
Approach: They describe language technology (LT) assessments carried out by the European Commission . they zoom in on two of the most extensive assessments, namely automated anonymisation and multilingual fake news processing.
Outcome: The language technology (LT) assessments carried out by the European Commission are detailed in this paper . they include a consultation round with stakeholders from public organisations, academia and industry . the ELRC action aims to create proof-of-concept environments integrating relevant tools and services .
Discovering Parallel Language Resources for Training MT Engines (L18-1)

Copied to clipboard

Challenge: Web crawling is an efficient way for compiling the monolingual, parallel and/or domain-specific corpora needed for machine translation and other HLT applications.
Approach: They propose a system for compiling monolingual, parallel and/or domain-specific corpora . ILSP-FC is a web crawling system that generates bilingual lexica and terminology lists .
Outcome: The ILSP Focused Crawler is a system developed by researchers at the IL SP/Athena RIC for the acquisition of such resources.
European Language Grid: An Overview (2020.lrec-1)

Copied to clipboard

Challenge: European LT business is dominated by hundreds of SMEs and a few large players, with technologies that outperform the global players.
Approach: European Language Grid (ELG) project addresses this by establishing the ELG as the primary platform for LT in Europe.
Outcome: European Language Grid (ELG) will be primary platform for LT in Europe . it will provide access to hundreds of commercial and non-commercial LTs for all European languages, including running tools and services as well as data sets and resources.
Digital Language Infrastructures – Documenting Language Actors (2020.lrec-1)

Copied to clipboard

Challenge: Existing language infrastructures focus on large institutions, but smaller institutions could benefit from them.
Approach: They propose to reach out to smaller local language actors on a local scope . they highlight the need to connect these institutions to existing infrastructures .
Outcome: The proposed project aims to reach out to smaller local language actors on a local scope and discuss challenges related to this ambition.
Making Metadata Fit for Next Generation Language Technology Platforms: The Metadata Schema of the European Language Grid (2020.lrec-1)

Copied to clipboard

Challenge: Metadata are a key factor in the management, sharing and usage of digital assets . the European Language Grid project aims to be the primary hub and marketplace for industry-relevant Language Technology in Europe.
Approach: They propose a rich metadata schema catering for the description of Language Resources and Technologies.
Outcome: The proposed schema powers the European Language Grid platform that aims to be the primary hub and marketplace for industry-relevant Language Technology in Europe.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations