Papers by Marc Verhagen

6 papers
Bridging the LAPPS Grid and CLARIN (L18-1)

Copied to clipboard

Challenge: The LAPPS-CLARIN project is creating a "trust network" between the Language Applications Grid and WebLicht workflow engine . the goal is to allow users on one side of the bridge to gain appropriately authenticated access to the other .
Approach: The LAPPS-CLARIN project is creating a "trust network" between the Language Applications Grid and WebLicht workflow engine hosted by the CLARIN-D Center in Tübingen.
Outcome: The LAPPS-CLARIN project is creating a "trust network" between the Language Applications (LAPPS) Grid and the WebLicht workflow engine hosted by the CLARIN-D Center in Tübingen.
Interchange Formats for Visualization: LIF and MMIF (2020.lrec-1)

Copied to clipboard

Challenge: In this paper, we discuss the enhanced data visualization capabilities enabled by interoperating computational linguistics and natural language processing (NLP) applications.
Approach: They propose to use interchange formats to enable enhanced data visualization . they propose to combine CL tools with openly available visualization tools .
Outcome: The proposed formats can be used to create visualizations and manipulate annotations in multiple ways.
Evaluating Retrieval for Multi-domain Scientific Publications (2022.lrec-1)

Copied to clipboard

Challenge: a new framework for retrieval and mining of scientific publications is being developed . the AskMe retrieval engine is a bridge between xDD's publication database and the LAPPS Grid suite of NLP tools.
Approach: They evaluate AskMe retrievalengine using BEIR benchmark datasets . they aim to determine when and why certain approaches perform well on in-domain and out-of-domain data.
Outcome: The AskMe retrieval engine performs well on both in-domain and out-of-domain data.
Exploration and Discovery of the COVID-19 Literature through Semantic Visualization (2021.naacl-srw)

Copied to clipboard

Challenge: Existing semantic visualization methods are limited in finding connections between corpora targeting a specific topic.
Approach: They propose to use semantic visualization to explore large datasets of complex networks by exploiting the semantics of the relations in them.
Outcome: The proposed method can enable exploration and discovery over large datasets of complex networks by exploiting the semantics of the relations in them.
The CLAMS Platform at Work: Processing Audiovisual Data from the American Archive of Public Broadcasting (2022.lrec-1)

Copied to clipboard

Challenge: The Computational Linguistics Applications for Multimedia Services (CLAMS) platform provides access to computational content analysis tools for multimedia material.
Approach: They describe the CLAMS platform as it is and its initial prototype implementation from 2019 . they use a common multi-modal representation language called MMIF to create a workflow .
Outcome: The CLAMS platform is a new version of an initial prototype from 2019 . it can be used to add metadata to mass-digitized multimedia collections . the proposed version is based on the American Archive of Public Broadcasting data .
The Coreference under Transformation Labeling Dataset: Entity Tracking in Procedural Texts Using Event Models (2023.findings-acl)

Copied to clipboard

Challenge: et al., 2023) show that entity coreference resolution is improved when events bring about changes in entities that are not reflected in text mentions.
Approach: They propose to perform transformation-based entity linking prior to coreference relation identification to improve entity coreference.
Outcome: The proposed model improves coreference resolution of entities mentioned under a process-oriented model of events.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations