Bridging the LAPPS Grid and CLARIN (L18-1)

Copied to clipboard

Challenge: The LAPPS-CLARIN project is creating a "trust network" between the Language Applications Grid and WebLicht workflow engine . the goal is to allow users on one side of the bridge to gain appropriately authenticated access to the other .
Approach: The LAPPS-CLARIN project is creating a "trust network" between the Language Applications Grid and WebLicht workflow engine hosted by the CLARIN-D Center in Tübingen.
Outcome: The LAPPS-CLARIN project is creating a "trust network" between the Language Applications (LAPPS) Grid and the WebLicht workflow engine hosted by the CLARIN-D Center in Tübingen.

Similar Papers

Interoperability in an Infrastructure Enabling Multidisciplinary Research: The case of CLARIN (2020.lrec-1)

Copied to clipboard

Challenge: CLARIN supports the use and study of language data in general and aims to increase the potential for comparative research of cultural and societal phenomena across languages and disciplines.
Approach: They describe the interoperability requirements that arise through the existing ambitions and emerging frameworks.
Outcome: The proposed frameworks will address interoperability requirements at several levels, including organisation and ecosystem, design of workflow services, data curation, performance measurement and collaboration.
Mining Biomedical Publications With The LAPPS Grid (L18-1)

Copied to clipboard

Challenge: Natural language processing (NLP) text mining can increase productivity and innovation in the sciences by orders of magnitude.
Approach: The Language Applications Grid is an infrastructure for rapid development of natural language processing applications (NLP) it provides an intuitive and easy-to-use platform for users to exploit NLP tools and resources . the Grid integrates the services and resources provided by PubAnnotation to greatly enhance the user's ability to annotate scientific publications .
Outcome: The Language Applications (LAPPS) Grid is an infrastructure for rapid development of natural language processing applications (NLP) it integrates services and resources provided by PubAnnotation to greatly enhance user's ability to annotate scientific publications and share the results.
CLARIN’s Key Resource Families (L18-1)

Copied to clipboard

Challenge: CLARIN is a European Research Infrastructure that supports the accessibility of language resources and technologies to researchers from the Digital Humanities and Social Sciences.
Approach: They propose to present key resource families in a uniform way for researchers to use in their research using the CLARIN infrastructure.
Outcome: The key resource families are newspaper, parliamentary, CMC (computer-mediated communication), and parallel corpora.
CLARIN: Towards FAIR and Responsible Data Science Using Language Resources (L18-1)

Copied to clipboard

Challenge: CLARIN is a European Research Infrastructure providing access to language resources and tools for researchers in the humanities and social sciences.
Approach: This paper outlines the CLARIN vision and strategy . it explains how the design and implementation of CLARINS are compliant with the FAIR principles .
Outcome: The paper outlines the CLARIN vision and strategy and explains how it is compliant with the FAIR principles: findability, accessibility, interoperability and reusability of data.
Introducing the CLARIN Knowledge Centre for Linguistic Diversity and Language Documentation (L18-1)

Copied to clipboard

Challenge: Knowledge Centres comprise physical institutions with particular expertise in certain areas and are committed to providing their expertise in the form of reliable knowledge-sharing services.
Approach: They propose to build a Knowledge Sharing Infrastructure (KSI) to ensure existing knowledge and expertise is easily available for the CLARIN community and for the humanities research communities for which CLARINS is being developed.
Outcome: The CLARIN Knowledge Centre for Linguistic Diversity and Language Documentation (CKLD) is a virtual distributed centre comprising institutions at the Universities of London, Cologne and Hamburg.
Handling Big Data and Sensitive Data Using EUDAT’s Generic Execution Framework and the WebLicht Workflow Engine. (L18-1)

Copied to clipboard

Challenge: a new workflow engine for web-based tools and workflow engines can be used to process big data and data with restrictive property rights.
Approach: They propose to bring WebLicht workflow engine with EUDAT-based Generic Execution Framework to address this issue.
Outcome: The proposed workflow engine can handle large data sets with restrictive property rights . the EUDAT project is developing the Generic Execution Framework (GEF)
Interchange Formats for Visualization: LIF and MMIF (2020.lrec-1)

Copied to clipboard

Challenge: In this paper, we discuss the enhanced data visualization capabilities enabled by interoperating computational linguistics and natural language processing (NLP) applications.
Approach: They propose to use interchange formats to enable enhanced data visualization . they propose to combine CL tools with openly available visualization tools .
Outcome: The proposed formats can be used to create visualizations and manipulate annotations in multiple ways.
Italian NLP for Everyone: Resources and Models from EVALITA to the European Language Grid (2022.lrec-1)

Copied to clipboard

Challenge: European Language Grid enables researchers and practitioners to easily distribute and use NLP resources and models.
Approach: They propose to integrate Italian NLP resources into the European Language Grid . they show how easy it is to use the integrated systems and demonstrate how seamless it is .
Outcome: The European Language Grid enables researchers and practitioners to easily distribute and use NLP resources and models.
ParCourE: A Parallel Corpus Explorer for a Massively Multilingual Corpus (2021.acl-demo)

Copied to clipboard

Challenge: 7000 languages worldwide are spoken, but most research is focused on English . multilinguality is essential for multilingual research, and is a key component of the process.
Approach: They propose a wordaligned parallel corpus that can be browsed using an online tool . they use the word alignment tools SimAlign and BabelNet to find the alignments .
Outcome: The proposed tool can be set up for any parallel corpus and explores its quality and properties.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations