Papers by Vincent Vandeghinste

2 papers
Challenges with Sign Language Datasets for Sign Language Recognition and Translation (2022.lrec-1)

Copied to clipboard

Challenge: Sign Languages are the primary means of communication for at least half a million people in Europe . however, the development of SL recognition and translation tools is slowed down by resource scarcity and data formats are not suitable for machine learning.
Approach: They propose a framework to unify available resources and facilitate SL research for different languages.
Outcome: The proposed framework is based on a set of ELAN files and returns textual and visual data ready to train SL recognition and translation models.
Jargon: A Suite of Language Models and Evaluation Tasks for French Specialized Domains (2024.lrec-main)

Copied to clipboard

Challenge: Pretrained language models are the de facto backbone of most state-of-the-art NLP systems.
Approach: They propose a family of domain-specific pretrained PLMs for French focusing on three important domains: transcribed speech, medicine, and law.
Outcome: The proposed models perform better on transcribed speech, medicine, and law domains than state-of-the-art models on a diverse set of tasks and datasets.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations