Papers by Cvetana Krstev

4 papers
Using English Baits to Catch Serbian Multi-Word Terminology (L18-1)

Copied to clipboard

Challenge: a new method for bilingual terminology extraction is proposed for a source language and a target language.
Approach: They propose to use a bilingual terminology extraction approach for a source language and a target language to extract the terminology for sri lanka.
Outcome: The proposed method extracts terminology for a source language and a target language from it.
Distant Reading in Digital Humanities: Case Study on the Serbian Part of the ELTeC Collection (2022.lrec-1)

Copied to clipboard

Challenge: Distant reading is a new scale of description that does not displace previous scales of literary description.
Approach: They present the Serbian part of the ELTeC multilingual corpus . they propose to test various methods and tools for distant reading .
Outcome: The Serbian part of the ELTeC multilingual corpus is being built to test various methods and tools . Several use examples show that this sub-collection is usefull for both close and distant reading approaches.
Machine Learning and Deep Neural Network-Based Lemmatization and Morphosyntactic Tagging for Serbian (2020.lrec-1)

Copied to clipboard

Challenge: The training of new tagger models for Serbian is motivated by the enhancement of the existing tagset with the grammatical category of a gender.
Approach: They propose to use TreeTagger and spaCy taggers to train new Serbian tagger models and to align Serbian morphological dictionaries with the grammatical category of a gender.
Outcome: The proposed models achieve 98% PoS-tagging precision per token, and the annotated dataset will be published.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations