Papers by Akhmed Sakip
Data Laundering: Artificially Boosting Benchmark Results through Knowledge Distillation (2025.acl-long)
Copied to clipboard
| Challenge: | Existing studies show that language model benchmarks are vulnerable to manipulation and exploitation. |
| Approach: | They propose a method that allows the covert transfer of benchmark-specific knowledge through seemingly legitimate intermediate training steps. |
| Outcome: | The proposed method can achieve significant improvements in accuracy without developing reasoning capabilities. |
KazMMLU: Evaluating Language Models on Kazakh, Russian, and Regional Knowledge of Kazakhstan (2025.acl-long)
Copied to clipboard
Mukhammed Togmanov, Nurdaulet Mukhituly, Diana Turmakhan, Jonibek Mansurov, Maiya Goloburda, Akhmed Sakip, Zhuohan Xie, Yuxia Wang, Bekassyl Syzdykov, Nurkhan Laiyk, Alham Fikri Aji, Ekaterina Kochmar, Preslav Nakov, Fajri Koto
| Challenge: | Kazakh language remains underrepresented in the field of natural language processing despite the country's population exceeding twenty million . however, there is a lack of dedicated models and benchmark evaluations specifically tailored to Kazakh languages. |
| Approach: | They propose to create a dataset specifically designed for Kazakh language with 23,000 questions sourced from authentic educational materials and manually validated by native speakers and educators. |
| Outcome: | The first MMLU-style dataset specifically designed for Kazakh language. |