Challenge: Existing methods to assess lexical complexity are used to evaluate the difficulty of vocabulary for language learners.
Approach: They propose to use pre-trained language models to assess the complexity of a word based on its context.
Outcome: The proposed method outperforms the best systems in SemEval-2021.

Similar Papers

ChatGPT Beyond English: Towards a Comprehensive Evaluation of Large Language Models in Multilingual Learning (2023.findings-emnlp)

Copied to clipboard

Challenge: Recent advances in natural language processing (NLP) have led to significant breakthroughs in the field.
Approach: They evaluate ChatGPT over multiple tasks with diverse languages and large datasets to provide more comprehensive information for multilingual NLP applications.
Outcome: The proposed model can process and generate texts for multiple languages due to its multilingual training data.
A Systematic Study and Comprehensive Evaluation of ChatGPT on Benchmark Datasets (2023.findings-acl)

Copied to clipboard

Challenge: Currently, the evaluation of large language models (LLMs) such as ChatGPT in academic datasets is difficult due to the difficulty of evaluating the generative outputs produced by this model against the ground truth.
Approach: They evaluate ChatGPT across 140 tasks and analyze 255K responses it generates in academic datasets.
Outcome: The proposed model performs well on 140 tasks and generates 255K responses in these datasets.
Investigating Large Language Models for Complex Word Identification in Multilingual and Multidomain Setups (2024.emnlp-main)

Copied to clipboard

Challenge: Large language models (LLMs) are popular in the Natural Language Processing community because of their versatility and capability to solve unseen tasks in zero/few-shot settings.
Approach: They investigate the use of large language models in CWI, LCP, and MWE settings by evaluating their use in zero-shot, few-shot and fine-tuning settings.
Outcome: The proposed models struggle in certain conditions or achieve comparable results against existing methods.
Simplification Using Paraphrases and Context-Based Lexical Substitution (N18-1)

Copied to clipboard

Challenge: Lexical simplification involves identifying complex words or phrases that need to be simplified and suggesting simpler meaning-preserving substitutes.
Approach: They propose a complex word identification model that exploits both lexical and contextual features and a word-embedding lexical substitution model to replace the detected complex words with simpler paraphrases.
Outcome: The proposed model detects complex words with higher accuracy than other models and proposes good substitutes in context.
Estimating Lexical Complexity from Document-Level Distributions (2024.lrec-main)

Copied to clipboard

Challenge: Existing methods for complexity estimation are limited to entire documents . health assessment tools are too short for existing methods to apply .
Approach: They propose a two-step approach for estimating lexical complexity that does not rely on pre-annotated data.
Outcome: The proposed method is tested on the Norwegian language and compares with other assessment tools.
A Word-Complexity Lexicon and A Neural Readability Ranking Model for Lexical Simplification (D18-1)

Copied to clipboard

Challenge: Current lexical simplification approaches rely on heuristics and corpus level features that do not align with human judgment.
Approach: They propose a human-rated word-complexity lexicon and a neural readability ranking model that uses human ratings to measure the complexity of any given word or phrase.
Outcome: The proposed model performs better than state-of-the-art models for lexical simplification tasks and evaluation datasets.
One Size Does Not Fit All: The Case for Personalised Word Complexity Models (2022.findings-naacl)

Copied to clipboard

Challenge: Complex word identification (CWI) aims to identify words in a text that are difficult for a reader to understand and therefore benefit from simplification.
Approach: They propose to use a novel active learning framework to tailor models to individual readers and release a dataset of complexity annotations and models as a benchmark for further research.
Outcome: The proposed model can be tailored to individual readers and released as a benchmark for future research.
Evaluating Pretrained Causal Language Models for Synonymy (2025.findings-acl)

Copied to clipboard

Challenge: Despite the scaling of causal language models, the underlying basis of complex skills remains unclear.
Approach: They propose that subjacent skills such as synonymy might be explained using linguistic concepts.
Outcome: The proposed model recognizes synonymy but struggles to generate synonyms when prompted with relevant context.
Explainable Prediction of Text Complexity: The Missing Preliminaries for Text Simplification (2021.acl-long)

Copied to clipboard

Challenge: Text simplification reduces the language complexity of professional content for accessibility purposes.
Approach: They propose that text simplification can be decomposed into a pipeline of tasks . they show that the pipeline can be used to predict whether a text needs to be simplified .
Outcome: The proposed model improves the performance of out-of-sample simplification tests on a blackbox lexical model . the proposed model reduces the complexity of professional text by a large margin .
Detecting Multiword Expression Type Helps Lexical Complexity Assessment (2020.lrec-1)

Copied to clipboard

Challenge: Multiword expressions (MWEs) represent lexemes that should be treated as single lexical units due to their idiosyncratic nature.
Approach: They re-annotate a complex word identification shared task 2018 dataset . they find that a lexical complexity assessment system benefits from the information .
Outcome: The proposed dataset provides valuable information for the text simplification community.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations