Papers by Piyumal Demotte

1 papers
BERTifying Sinhala - A Comprehensive Analysis of Pre-trained Language Models for Sinhala Text Classification (2022.lrec-1)

Copied to clipboard

Challenge: Large-scale monolingual pre-trained language models have shown promising results for high-resource as well as lowresource languages, especially for text classification.
Approach: They provide a set of recommendations for using pre-trained models for Sinhala text classification and introduce new annotated datasets useful for future research.
Outcome: The proposed models are far superior to existing models for Sinhala and set a strong baseline for text classification when fine-tuned.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations