European Language Resource Coordination: Collecting Language Resources for Public Sector Multilingual Information Management (L18-1)
Copied to clipboard
Andrea Lösch, Valérie Mapelli, Stelios Piperidis, Andrejs Vasiļjevs, Lilli Smal, Thierry Declerck, Eileen Schnur, Khalid Choukri, Josef van Genabith
| Challenge: | European Language Resource Coordination (ELRC) initiated a number of actions to support the collection of Language Resources (LRs) within the public sector in EU member and CEF-affiliated countries. |
| Approach: | They propose to initiate actions to support the collection of Language Resources (LRs) within the public sector in EU member and CEF-affiliated countries. |
| Outcome: | The European Language Resource Coordination (ELRC) consortium initiated a number of actions to support the collection of Language Resources (LRs) within the public sector in EU member and CEF-affiliated countries. |
Similar Papers
Managing Public Sector Data for Multilingual Applications Development (L18-1)
Copied to clipboard
| Challenge: | eTranslation is a digital service that enables multilingual communication across public administrations in 30 European countries. |
| Approach: | They propose to develop a repository infrastructure specifically tailored to the needs of the eTranslation service of the European Commission. |
| Outcome: | The ELRC-SHARE repository is designed and developed specifically for the eTranslation service of the European Commission. |
Collecting Language Resources from Public Administrations in the Nordic and Baltic Countries (L18-1)
Copied to clipboard
| Challenge: | Several large-scale projects and initiatives have been undertaken in this century to collect language resources and create LR repositories and infrastructures on a pan-European scale. |
| Approach: | They present the work of Tilde on collecting language resources from government institutions and other public administrations in the Nordic and Baltic countries. |
| Outcome: | The results of the European Language Resources Coordination (ELRC) action in the Nordic and Baltic countries are presented. |
Language Data Sharing in European Public Services – Overcoming Obstacles and Creating Sustainable Data Sharing Infrastructures (2020.lrec-1)
Copied to clipboard
| Challenge: | Data is key in training modern language technologies. |
| Approach: | They summarise findings of first pan-European study on barriers to language data sharing . they identify structural challenges, disposition towards CAT tools and lack of digital skills . overcoming language barriers is one of the main challenges european citizens face . |
| Outcome: | The paper summarises the findings of the first pan-European study on barriers to language data sharing . the findings highlight the barriers and recommend solutions to overcome them . |
Language Resources to Support Language Diversity – the ELRA Achievements (2022.lrec-1)
Copied to clipboard
| Challenge: | ELRA and its operational agency ELDA have continued to increase their catalogue of Language Resources (LRs) over the past few years, ELLA and ELTA have contributed to improve the access to multilingual information in the context of the pandemic . |
| Approach: | ELRA and its operational agency ELDA have increased their catalogue of Language Resources (LRs) over the past few years, they have established partnerships for the production of various types of LRs. |
| Outcome: | ELRA and its operational agency ELDA have contributed to improve the access to multilingual information in the context of the pandemic, develop tools for the de-identification of texts in the legal and medical domains, and support the EU eTranslation Machine Translation system. |
ELRC Action: Covering Confidentiality, Correctness and Cross-linguality (2022.lrec-1)
Copied to clipboard
Tom Vanallemeersch, Arne Defauw, Sara Szoc, Alina Kramchaninova, Joachim Van den Bogaert, Andrea Lösch
| Challenge: | ELRC aims to reduce language barriers by assessing language technology (LT) specifications . automated anonymisation and multilingual fake news processing are two of the most extensive LT assessments . |
| Approach: | They describe language technology (LT) assessments carried out by the European Commission . they zoom in on two of the most extensive assessments, namely automated anonymisation and multilingual fake news processing. |
| Outcome: | The language technology (LT) assessments carried out by the European Commission are detailed in this paper . they include a consultation round with stakeholders from public organisations, academia and industry . the ELRC action aims to create proof-of-concept environments integrating relevant tools and services . |
European Language Grid: A Joint Platform for the European Language Technology Community (2021.eacl-demos)
Copied to clipboard
Georg Rehm, Stelios Piperidis, Kalina Bontcheva, Jan Hajic, Victoria Arranz, Andrejs Vasiļjevs, Gerhard Backfried, Jose Manuel Gomez-Perez, Ulrich Germann, Rémi Calizzano, Nils Feldhus, Stefanie Hegele, Florian Kintzel, Katrin Marheinecke, Julian Moreno-Schneider, Dimitris Galanis, Penny Labropoulou, Miltos Deligiannis, Katerina Gkirtzou, Athanasia Kolovou, Dimitris Gkoumas, Leon Voukoutis, Ian Roberts, Jana Hamrlova, Dusan Varis, Lukas Kacena, Khalid Choukri, Valérie Mapelli, Mickaël Rigault, Julija Melnika, Miro Janosik, Katja Prinz, Andres Garcia-Silva, Cristian Berrio, Ondrej Klejch, Steve Renals
| Challenge: | Europe is a multilingual society, in which dozens of languages are spoken. |
| Approach: | They describe the European Language Grid, which is targeted to evolve into the primary platform and marketplace for LT in Europe by providing one umbrella platform for the European LT landscape. |
| Outcome: | The European Language Grid (ELG) will provide access to 1300 services for all European languages as well as thousands of data sets. |
Discovering Parallel Language Resources for Training MT Engines (L18-1)
Copied to clipboard
| Challenge: | Web crawling is an efficient way for compiling the monolingual, parallel and/or domain-specific corpora needed for machine translation and other HLT applications. |
| Approach: | They propose a system for compiling monolingual, parallel and/or domain-specific corpora . ILSP-FC is a web crawling system that generates bilingual lexica and terminology lists . |
| Outcome: | The ILSP Focused Crawler is a system developed by researchers at the IL SP/Athena RIC for the acquisition of such resources. |
European Language Grid: An Overview (2020.lrec-1)
Copied to clipboard
Georg Rehm, Maria Berger, Ela Elsholz, Stefanie Hegele, Florian Kintzel, Katrin Marheinecke, Stelios Piperidis, Miltos Deligiannis, Dimitris Galanis, Katerina Gkirtzou, Penny Labropoulou, Kalina Bontcheva, David Jones, Ian Roberts, Jan Hajič, Jana Hamrlová, Lukáš Kačena, Khalid Choukri, Victoria Arranz, Andrejs Vasiļjevs, Orians Anvari, Andis Lagzdiņš, Jūlija Meļņika, Gerhard Backfried, Erinç Dikici, Miroslav Janosik, Katja Prinz, Christoph Prinz, Severin Stampler, Dorothea Thomas-Aniola, José Manuel Gómez-Pérez, Andres Garcia Silva, Christian Berrío, Ulrich Germann, Steve Renals, Ondrej Klejch
| Challenge: | European LT business is dominated by hundreds of SMEs and a few large players, with technologies that outperform the global players. |
| Approach: | European Language Grid (ELG) project addresses this by establishing the ELG as the primary platform for LT in Europe. |
| Outcome: | European Language Grid (ELG) will be primary platform for LT in Europe . it will provide access to hundreds of commercial and non-commercial LTs for all European languages, including running tools and services as well as data sets and resources. |
Digital Language Infrastructures – Documenting Language Actors (2020.lrec-1)
Copied to clipboard
| Challenge: | Existing language infrastructures focus on large institutions, but smaller institutions could benefit from them. |
| Approach: | They propose to reach out to smaller local language actors on a local scope . they highlight the need to connect these institutions to existing infrastructures . |
| Outcome: | The proposed project aims to reach out to smaller local language actors on a local scope and discuss challenges related to this ambition. |
Making Metadata Fit for Next Generation Language Technology Platforms: The Metadata Schema of the European Language Grid (2020.lrec-1)
Copied to clipboard
Penny Labropoulou, Katerina Gkirtzou, Maria Gavriilidou, Miltos Deligiannis, Dimitris Galanis, Stelios Piperidis, Georg Rehm, Maria Berger, Valérie Mapelli, Michael Rigault, Victoria Arranz, Khalid Choukri, Gerhard Backfried, José Manuel Gómez-Pérez, Andres Garcia-Silva
| Challenge: | Metadata are a key factor in the management, sharing and usage of digital assets . the European Language Grid project aims to be the primary hub and marketplace for industry-relevant Language Technology in Europe. |
| Approach: | They propose a rich metadata schema catering for the description of Language Resources and Technologies. |
| Outcome: | The proposed schema powers the European Language Grid platform that aims to be the primary hub and marketplace for industry-relevant Language Technology in Europe. |