Papers by Magda Ševčíková
Towards Universal Segmentations: UniSegments 1.0 (2022.lrec-1)
Copied to clipboard
Zdeněk Žabokrtský, Niyati Bafna, Jan Bodnár, Lukáš Kyjánek, Emil Svoboda, Magda Ševčíková, Jonáš Vidra
| Challenge: | Existing data resources for morphological segmentation are limited to 32 languages . a large number of word forms exist, with some sub-parts being "recycled" many times . |
| Approach: | They propose a multilingual data resource for morphological segmentation in 32 languages . they analyze diversity of how individual linguistic phenomena are captured across them . |
| Outcome: | The proposed scheme is based on 17 existing data resources relevant for segmentation in 32 languages. |
Semi-Automatic Construction of Word-Formation Networks (for Polish and Spanish) (L18-1)
Copied to clipboard
| Challenge: | a semi-automatic method for the construction of derivational networks is proposed . the proposed method is general enough to be adopted for other languages . |
| Approach: | They propose a semi-automatic method for the construction of derivational networks using a sequential pattern mining technique. |
| Outcome: | The proposed method is general enough to be adopted for other languages. |