TheRuSLan: Database of Russian Sign Language (2020.lrec-1)

Copied to clipboard

Challenge: The database is the first of a kind for Russian sign language and is intended for use in machine learning, gesture recognition and sign language linguistics.
Approach: They present a Russian sign language multimedia database called TheRuSLan . the database includes lexical units from Russian sign languages within one subject area .
Outcome: The proposed database includes lexical units from Russian sign language within one subject area.

Similar Papers

Sign-Language Datasets at Scale: A Comprehensive Survey on Resources, Benchmarks, and Annotation Standards (2026.acl-long)

Copied to clipboard

Challenge: Existing benchmarks fail to reflect real-world communication needs and are limited in their coverage.
Approach: They present a comprehensive index of sign-language datasets, covering 120 resources across 35 sign languages.
Outcome: The proposed index covers 120 resources across 35 sign languages.
Challenges with Sign Language Datasets for Sign Language Recognition and Translation (2022.lrec-1)

Copied to clipboard

Challenge: Sign Languages are the primary means of communication for at least half a million people in Europe . however, the development of SL recognition and translation tools is slowed down by resource scarcity and data formats are not suitable for machine learning.
Approach: They propose a framework to unify available resources and facilitate SL research for different languages.
Outcome: The proposed framework is based on a set of ELAN files and returns textual and visual data ready to train SL recognition and translation models.
Collocations in Russian Lexicography and Russian Collocations Database (2020.lrec-1)

Copied to clipboard

Challenge: Existing methods for collocation extraction cannot be considered perfect, argues a new study.
Approach: They propose to build a database that will include dictionary and statistical collocations in Russian . the database will be based on dictionaries and online systems that describe collocation .
Outcome: The proposed database will include dictionary and statistical collocations in Russian . the results can be useful for machine learning and for other NLP tasks .
Sign2Vis: Automated Data Visualization from Sign Language (2025.findings-acl)

Copied to clipboard

Challenge: Existing methods to translate natural language descriptions into visualization queries focus on spoken languages, not sign languages.
Approach: They propose a sign language interface that enables the DHH community to engage more fully with data analysis.
Outcome: The proposed interface can be used by the deaf and hard-of-hearing community.
SwissSLi: The Multi-parallel Sign Language Corpus for Switzerland (2024.lrec-main)

Copied to clipboard

Challenge: Using a CC BY-NC-SA 4.0 license, this corpus contains parallel sign language videos and spoken language subtitles.
Approach: They introduce SwissSLi, the first sign language corpus that contains parallel data of all three Swiss sign languages.
Outcome: The proposed corpus contains parallel sign language videos and spoken language subtitles.
Signbank: Software to Support Web Based Dictionaries of Sign Language (L18-1)

Copied to clipboard

Challenge: Auslan Signbank is an on-line dictionary for Australian Sign Language (Auslan) it was originally built to support the Auslan signbank web dictionary, but was re-implemented using Microsoft SQL Server.
Approach: This paper describes the overall architecture of the Auslan Signbank system and its representation of lexical entries and associated entities.
Outcome: The current version of Auslan Signbank is an open-source re-implementation of the original website, with features added to allow updates to the database by researchers.
Crowdsourcing Kazakh-Russian Sign Language: FluentSigners-50 (2022.lrec-1)

Copied to clipboard

Challenge: Using crowdsourcing, we created a signer independent dataset for sign language processing.
Approach: They propose to crowdsource a signer independent Kazakh-Russian Sign Language (KRSL) dataset.
Outcome: The proposed dataset consists of 173 sentences performed by 50 signers for 43,250 video samples.
OpenHands: Making Sign Language Recognition Accessible with Pose-based Pretrained Models across Languages (2022.acl-long)

Copied to clipboard

Challenge: a new study examines the performance of pretraining for sign language recognition in low-resource settings.
Approach: They propose using pose extracted through pretrained models as the standard modality of data to reduce training time and enable efficient inference.
Outcome: The proposed model reduces training time and allows efficient inference in sign languages.
Listen, Decipher and Sign: Toward Unsupervised Speech-to-Sign Language Recognition (2023.findings-acl)

Copied to clipboard

Challenge: Existing supervised sign language recognition systems rely on well-annotated data . instead, an unsupervised speech-to-sign language recognition system learns to translate between spoken and sign languages by observing only non-parallel speech and sign-language corpora.
Approach: They propose an unsupervised speech-to-sign language recognition system that can translate between spoken and sign languages by observing only non-parallel speech and sign-language corpora.
Outcome: The proposed approach outperforms baseline models on sign language corpora by 50% . the proposed approach is available at https://github.com/cactuswiththoughts/UnsupSpeech2Sign.git .
JWSign: A Highly Multilingual Corpus of Bible Translations for more Diversity in Sign Language Processing (2023.findings-emnlp)

Copied to clipboard

Challenge: Existing sign language datasets are limited and skewed towards high-income sign languages, mainly those from high-risk countries.
Approach: They propose a large and highly multilingual dataset for sign language translation: JWSign.
Outcome: The proposed dataset consists of 2,530 hours of Bible translations in 98 sign languages, featuring more than 1,500 individual signers.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations