Papers by Jelena Kallas

2 papers
Leveraging Domain Corpora for Enhanced Terminology: The Case of Estonian-English Remote Sensing Termbase (2024.lrec-main)

Copied to clipboard

Challenge: Termbase is a domain corpora and terminological database for remote sensing in Estonia.
Approach: They propose to develop an Estonian-English Remote Sensing Termbase from scratch . they use the Estonian Remote Sensenting Corpus 2022 as the primary data source .
Outcome: The Estonian Remote Sensing Corpus 2022 served as the primary data source for the termbase.
A Multilingual Evaluation Dataset for Monolingual Word Sense Alignment (2020.lrec-1)

Copied to clipboard

Challenge: a new dataset aims to align monolingual dictionaries with a single sense level for 15 languages . this dataset covers a wide range of languages and resources .
Approach: They propose to manually align monolingual dictionaries with possible semantic relationships . they use 15 languages to create a new baseline for the task of monolingual word sense alignment .
Outcome: The proposed dataset covers 15 languages and covers the more challenging task of linking general-purpose language.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations