| Challenge: | The database is the first of a kind for Russian sign language and is intended for use in machine learning, gesture recognition and sign language linguistics. |
| Approach: | They present a Russian sign language multimedia database called TheRuSLan . the database includes lexical units from Russian sign languages within one subject area . |
| Outcome: | The proposed database includes lexical units from Russian sign language within one subject area. |
Similar Papers
Sign-Language Datasets at Scale: A Comprehensive Survey on Resources, Benchmarks, and Annotation Standards (2026.acl-long)
Copied to clipboard
| Challenge: | Existing benchmarks fail to reflect real-world communication needs and are limited in their coverage. |
| Approach: | They present a comprehensive index of sign-language datasets, covering 120 resources across 35 sign languages. |
| Outcome: | The proposed index covers 120 resources across 35 sign languages. |
Challenges with Sign Language Datasets for Sign Language Recognition and Translation (2022.lrec-1)
Copied to clipboard
Mirella De Sisto, Vincent Vandeghinste, Santiago Egea Gómez, Mathieu De Coster, Dimitar Shterionov, Horacio Saggion
| Challenge: | Sign Languages are the primary means of communication for at least half a million people in Europe . however, the development of SL recognition and translation tools is slowed down by resource scarcity and data formats are not suitable for machine learning. |
| Approach: | They propose a framework to unify available resources and facilitate SL research for different languages. |
| Outcome: | The proposed framework is based on a set of ELAN files and returns textual and visual data ready to train SL recognition and translation models. |
Collocations in Russian Lexicography and Russian Collocations Database (2020.lrec-1)
Copied to clipboard
| Challenge: | Existing methods for collocation extraction cannot be considered perfect, argues a new study. |
| Approach: | They propose to build a database that will include dictionary and statistical collocations in Russian . the database will be based on dictionaries and online systems that describe collocation . |
| Outcome: | The proposed database will include dictionary and statistical collocations in Russian . the results can be useful for machine learning and for other NLP tasks . |
Sign2Vis: Automated Data Visualization from Sign Language (2025.findings-acl)
Copied to clipboard
| Challenge: | Existing methods to translate natural language descriptions into visualization queries focus on spoken languages, not sign languages. |
| Approach: | They propose a sign language interface that enables the DHH community to engage more fully with data analysis. |
| Outcome: | The proposed interface can be used by the deaf and hard-of-hearing community. |
SwissSLi: The Multi-parallel Sign Language Corpus for Switzerland (2024.lrec-main)
Copied to clipboard
| Challenge: | Using a CC BY-NC-SA 4.0 license, this corpus contains parallel sign language videos and spoken language subtitles. |
| Approach: | They introduce SwissSLi, the first sign language corpus that contains parallel data of all three Swiss sign languages. |
| Outcome: | The proposed corpus contains parallel sign language videos and spoken language subtitles. |
Signbank: Software to Support Web Based Dictionaries of Sign Language (L18-1)
Copied to clipboard
Steve Cassidy, Onno Crasborn, Henri Nieminen, Wessel Stoop, Micha Hulsbosch, Susan Even, Erwin Komen, Trevor Johnston
| Challenge: | Auslan Signbank is an on-line dictionary for Australian Sign Language (Auslan) it was originally built to support the Auslan signbank web dictionary, but was re-implemented using Microsoft SQL Server. |
| Approach: | This paper describes the overall architecture of the Auslan Signbank system and its representation of lexical entries and associated entities. |
| Outcome: | The current version of Auslan Signbank is an open-source re-implementation of the original website, with features added to allow updates to the database by researchers. |
Crowdsourcing Kazakh-Russian Sign Language: FluentSigners-50 (2022.lrec-1)
Copied to clipboard
| Challenge: | Using crowdsourcing, we created a signer independent dataset for sign language processing. |
| Approach: | They propose to crowdsource a signer independent Kazakh-Russian Sign Language (KRSL) dataset. |
| Outcome: | The proposed dataset consists of 173 sentences performed by 50 signers for 43,250 video samples. |
OpenHands: Making Sign Language Recognition Accessible with Pose-based Pretrained Models across Languages (2022.acl-long)
Copied to clipboard
| Challenge: | a new study examines the performance of pretraining for sign language recognition in low-resource settings. |
| Approach: | They propose using pose extracted through pretrained models as the standard modality of data to reduce training time and enable efficient inference. |
| Outcome: | The proposed model reduces training time and allows efficient inference in sign languages. |
Listen, Decipher and Sign: Toward Unsupervised Speech-to-Sign Language Recognition (2023.findings-acl)
Copied to clipboard
Liming Wang, Junrui Ni, Heting Gao, Jialu Li, Kai Chieh Chang, Xulin Fan, Junkai Wu, Mark Hasegawa-Johnson, Chang Yoo
| Challenge: | Existing supervised sign language recognition systems rely on well-annotated data . instead, an unsupervised speech-to-sign language recognition system learns to translate between spoken and sign languages by observing only non-parallel speech and sign-language corpora. |
| Approach: | They propose an unsupervised speech-to-sign language recognition system that can translate between spoken and sign languages by observing only non-parallel speech and sign-language corpora. |
| Outcome: | The proposed approach outperforms baseline models on sign language corpora by 50% . the proposed approach is available at https://github.com/cactuswiththoughts/UnsupSpeech2Sign.git . |
JWSign: A Highly Multilingual Corpus of Bible Translations for more Diversity in Sign Language Processing (2023.findings-emnlp)
Copied to clipboard
| Challenge: | Existing sign language datasets are limited and skewed towards high-income sign languages, mainly those from high-risk countries. |
| Approach: | They propose a large and highly multilingual dataset for sign language translation: JWSign. |
| Outcome: | The proposed dataset consists of 2,530 hours of Bible translations in 98 sign languages, featuring more than 1,500 individual signers. |