Papers by Krister Lindén

4 papers
HeLI-OTS, Off-the-shelf Language Identifier for Text (2022.lrec-1)

Copied to clipboard

Challenge: Existing off-the-shelf language identification tools favor widely used languages, but Heli-OTS can be used to identify a large group of languages.
Approach: They introduce an off-the-shelf text language identification tool using the HeLI method . they compare the He LI-OTS language identifier with fastText on two different data sets .
Outcome: The proposed language identification tool is compared with fastText on two different data sets.
BabyFST - Towards a Finite-State Based Computational Model of Ancient Babylonian (2020.lrec-1)

Copied to clipboard

Challenge: morphological analyzer for Akkadian is not yet available for the extinct language . we present a general finite-state based model for Babylonian that can achieve a coverage of 97.3% and a recall of 93.7% on token level.
Approach: They propose a general finite-state based morphological model for Babylonian that can achieve a coverage of 97.3% and recall up to 93.7% on lemmatization and POS-tagging tasks.
Outcome: The proposed model can achieve coverage and recall of 97.3% on lemmatization and POS-tagging tasks on token level from a transcribed input.
Automated Phonological Transcription of Akkadian Cuneiform Text (2020.lrec-1)

Copied to clipboard

Challenge: Akkadian was an east-semitic language spoken in ancient Mesopotamia . cuneiform text does not mark the inflection for logograms, so the inflected form needs to be inferred from the sentence context.
Approach: They propose to automate phonological transcription of the transliterated Akkadian corpora . transcription is normalized according to the grammatical description of a given dialect . they find that cuneiform text does not mark the inflection for logograms .
Outcome: The proposed transcriptions show the Akkadian renderings for Sumerian logograms, while the logogram transcription is more challenging.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations