Papers by Gökçe Uludoğan
TURNA: A Turkish Encoder-Decoder Language Model for Enhanced Understanding and Generation (2024.findings-acl)
Copied to clipboard
| Challenge: | Recent advances in natural language processing have favored well-resourced English-centric models, resulting in a significant gap with low-resource languages. |
| Approach: | They propose a language model for the low-resource language Turkish that is capable of both natural language understanding and generation tasks. |
| Outcome: | The proposed model outperforms multilingual models in understanding and generation tasks and competes with monolingual models for understanding tasks. |
HATECAT-TR: A Hate Speech Span Detection and Categorization Dataset for Turkish (2025.findings-emnlp)
Copied to clipboard
| Challenge: | a new dataset of Turkish tweets contains 4465 hateful spans . each hateful post is directed at one of eight minority groups . |
| Approach: | They propose a span-annotated dataset of Turkish tweets containing 4465 hateful spans . each hateful spat is categorized into one of five discourse types . |
| Outcome: | The proposed dataset contains 4465 hateful spans across 2981 tweets . each span is categorized into one of five discourse types . |