Papers by Jamshidbek Mirzakhalov

4 papers
A Large-Scale Study of Machine Translation in Turkic Languages (2021.emnlp-main)

Copied to clipboard

Challenge: a large corpus covering 22 Turkic languages is included in this paper . low-resource MT evaluation has traditionally focused on European languages due to limitations of available technology and resources.
Approach: They present a case study of the practical application of MT in the Turkic language family . they propose to realize the gains of NMT for Turkic languages under high-resource to extremely low-resourced scenarios.
Outcome: The proposed study shows that the new methods can be used in the Turkic language family . the results highlight bottlenecks in building competitive systems .
Can Transformer Language Models Predict Psychometric Properties? (2021.starsem-1)

Copied to clipboard

Challenge: Transformer-based language models (LMs) are gaining popularity on many NLP benchmark tasks.
Approach: They use human responses to calculate psychometric properties of test items . they find transformer-based LMs predict psychometric property consistently well .
Outcome: The transformer-based language models are able to predict psychometric properties of test items . the models can predict psychometries well in certain categories but poorly in others .
Towards a Task-Agnostic Model of Difficulty Estimation for Supervised Learning Tasks (2020.aacl-srw)

Copied to clipboard

Challenge: Recent advances on natural language processing (NLP) benchmarks have been driven by increasingly sophisticated language models, which are pretrained on enormous amounts of data before use.
Approach: They propose to develop a task-agnostic model for problem difficulty and apply it to the Stanford Natural Language Inference dataset.
Outcome: The proposed model predicts how many annotators will answer a question correctly and then projectes the difficulty estimates onto the full SNLI train set to create the curriculum.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations