Challenge: morphological analyzers for Quechua and Aymara have been evaluated for their performance . only a minority of these languages have been provided with adequate computational resources .
Approach: They evaluate existing morphological analyzers for Quechua and Aymara . they also examine how they handle other individual languages of the macrolanguage .
Outcome: The proposed tools perform well in Quechua and Aymara, and they can handle other languages.

Similar Papers

Indigenous Languages Spoken in Argentina: A Survey of NLP and Speech Resources (2025.coling-main)

Copied to clipboard

Challenge: Currently, no unified information on speakers and computational tools are available for these languages.
Approach: They present a systematization of the indigenous languages spoken in Argentina, along with national demographic data on the country’s Indigenous population.
Outcome: The proposed systematization of the indigenous languages spoken in Argentina, along with national demographic data on the country’s Indigenous population, is based on the Argentine population.
Juman++: A Morphological Analysis Toolkit for Scriptio Continua (D18-2)

Copied to clipboard

Challenge: a morphological analyzer is useful for languages without natural word boundaries, but it is difficult to improve it without creating costly annotations.
Approach: They propose a toolkit for developing morphological analyzers for languages without natural word boundaries using lattices and neural nets.
Outcome: The proposed morphological analyzer of Japanese achieves new SOTA on Jumandic-based corpora while being 250 times faster than the previous one.
Proceedings of the 2025 Annual Conference of the Nations of the Americas Chapter of the Association for Computational Linguistics: Human Language Technologies (Volume 5: Tutorial Abstracts) (2025.naacl-tutorial)

Copied to clipboard

Challenge: NAACL 2025 tutorial sessions are a cornerstone event of the conference . tutorials are designed to equip you with the latest insights, tools, and methodologies .
Approach: NAACL 2025 will host a tutorial session on computational linguistics and natural language processing . the tutorials are a cornerstone event of the conference .
Outcome: the tutorial sessions at NAACL 2025 are a cornerstone event of the conference . each submission received a thorough evaluation by a panel of two to three reviewers .
Fine-grained Morphosyntactic Analysis and Generation Tools for More Than One Thousand Languages (2020.lrec-1)

Copied to clipboard

Challenge: Using morphosyntactic tools, we train and distribute tools for approximately one thousand languages.
Approach: They train and distribute morphosyntactic tools for approximately one thousand languages.
Outcome: The results show that the tools generalize well across rare and common forms alike.
Exploring Linguistic Probes for Morphological Inflection (2023.emnlp-main)

Copied to clipboard

Challenge: morphological inflection models typically employ language-independent data splitting algorithms.
Approach: They propose language-specific probes to test aspects of morphological generalization . they use three morphology-distinct languages to test their generalization abilities .
Outcome: The proposed language-specific probes are used to test morphological generalization abilities on three distinct languages.
Towards Universal Dependencies for Ancash Quechua (2024.lrec-main)

Copied to clipboard

Challenge: a new corpus of Quechua morphosyntactic features are described for Ancash Quechuan, the majority variety of the Central Quechual language family . the morphology of the language is a feature of the universal dependency (UD) schema . a syntactical parser would be the first NLP tool for a Quechuang language of this family based on the UD schema based upon the typology of the languages .
Approach: They propose to describe some morphosyntactic features of Ancash Quechua . they propose to build a corpus annotated according to the universal dependency schema .
Outcome: The proposed corpus is the first bilingual and sentence-aligned digital corpus in Ancash Quechua and Spanish.
Challenges of language technologies for the indigenous languages of the Americas (C18-1)

Copied to clipboard

Challenge: Indigenous languages of the American continent are highly diverse, but have received little attention from the technological perspective.
Approach: They review the research, the digital resources and the available NLP systems for indigenous languages of the American continent . they stress the need of developing language resources and NLP tools for these languages .
Outcome: The authors review the research and the available NLP systems on indigenous languages of the Americas . they argue that the lack of resources and tools can have a negative impact on the communities which depend on these languages .
WordNet-QU: Development of a Lexical Database for Quechua Varieties (2022.coling-1)

Copied to clipboard

Challenge: Quechua is a low-resource language from south America but lacks resources to build high-performance computational systems.
Approach: They propose to include Quechua in a lexical database called wordnet . they propose a synset alignment algorithm to compare Quechuan to its nearest high-resource language .
Outcome: The proposed system compares Quechua to its nearest high-resource language, Spanish . it uses a synset alignment algorithm to find Quechuan resources in a lexical database .
Proceedings of the 2024 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (Volume 5: Tutorial Abstracts) (2024.naacl-tutorials)

Copied to clipboard

Challenge: NAACL 2024 tutorial sessions are a cornerstone event of the conference . a total of 27 tutorial submissions were received, and 6 were selected for presentation .
Approach: NAACL 2024 will host a tutorial session featuring top-notch researchers . the tutorials aim to equip attendees with the latest tools and methodologies . a total of 27 tutorial submissions were received, and 6 were selected for presentation .
Outcome: the tutorial sessions are a cornerstone event of the conference . the call, submission, reviewing, and selection of tutorials were coordinated . a total of 27 tutorial submissions were received, and 6 were selected for presentation .

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations