Challenge: lexicon-based simplification methods can help patients understand medical documents . but they must ensure that the content is transmitted rigorously and not creating wrong information.
Approach: They tested automatic simplification techniques using a Spanish lexicon of technical and laymen terms.
Outcome: The proposed methods improve the quantitative results and the human evaluation of medical documents.

Similar Papers

AutoMeTS: The Autocomplete for Medical Text Simplification (2020.coling-main)

Copied to clipboard

Challenge: Semi-automated text simplification approaches can be used to simplify text faster and at a higher quality.
Approach: They propose to use autocomplete to simplify medical texts using aligned English Wikipedia sentences and pretrained neural language models to analyze the additional context.
Outcome: The proposed model outperforms the best individual model by 2.1% and achieves a word prediction accuracy of 64.52%.
Multilingual Simplification of Medical Texts (2023.emnlp-main)

Copied to clipboard

Challenge: Existing work on medical text simplification has focused on monolingual settings . important findings in medicine are typically presented in technical, jargon-laden language . text simulating models can generate viable simplified texts, but there are outstanding challenges .
Approach: They propose a dataset for medical text simplification in four languages . they evaluate fine-tuned and zero-shot models across these languages based on human assessments and analyses .
Outcome: The proposed dataset evaluates models in English, Spanish, French, and Farsi . it shows that the models can generate viable simplified texts, but there are challenges .
Enhancing Sentence Simplification in Portuguese: Leveraging Paraphrases, Context, and Linguistic Features (2024.findings-acl)

Copied to clipboard

Challenge: Automated text simplification requires (paired) datasets that are scarce in languages other than English.
Approach: They propose a method that leverages paraphrases, context, and linguistic attributes to overcome the absence of paired texts in Portuguese.
Outcome: The proposed model surpasses the current state-of-the-art while competing with a Large Language Model.
MultiMSD: A Corpus for Multilingual Medical Text Simplification from Online Medical References (2025.findings-acl)

Copied to clipboard

Challenge: Medical texts contain technical terms, and non-experts often cannot use information effectively.
Approach: They propose a method for training medical text simplification models to actively paraphrase medical terms.
Outcome: The proposed method improves the performance of medical text simplification in nine languages.
Paragraph-level Simplification of Medical Texts (2021.naacl-main)

Copied to clipboard

Challenge: Existing methods for simplification of medical texts are limited due to jargon and technical content.
Approach: They propose to automate the simplification of medical texts by penalizing decoders for producing "jargon" terms.
Outcome: The proposed method improves on existing heuristics by penalizing the decoder for producing "jargon" terms.
French Biomedical Text Simplification: When Small and Precise Helps (2020.coling-main)

Copied to clipboard

Challenge: Existing studies on text simplification in English use large parallel monolingual corpora in which one complex sentence is paired with one or more simplified versions.
Approach: They use parallel sentences from existing health comparable corpora in French and WikiLarge corpus translated from English to French and a lexicon that associates medical terms with paraphrases.
Outcome: The proposed models are based on sentences from existing health comparable corpora in French and WikiLarge corpus translated from English to French.
ALEXSIS: A Dataset for Lexical Simplification in Spanish (2022.lrec-1)

Copied to clipboard

Challenge: Lexical Simplification is the process of replacing difficult words with easier synonyms while preserving the original information and meaning.
Approach: They introduce ALEXSIS, a dataset for Lexical Simplification, and use it to benchmark Lexical simplification systems in Spanish.
Outcome: The proposed dataset compares three approaches to Lexical Simplification in Spanish and a previous dataset for English.
Text Simplification from Professionally Produced Corpora (L18-1)

Copied to clipboard

Challenge: Existing approaches to Text Simplification rely on the Wikipedia-Simple Wikipedia parallel corpus, which is used for many tasks.
Approach: They propose to use the Newsela corpus to extract 550, 644 complex-simple sentence pairs from the corpus and introduce a lexical simplifier that uses the corpu to generate candidate simplifications.
Outcome: The proposed model outperforms state-of-the-art approaches and generates candidate simplifications from the newsela corpus.
Benchmarking Automated Clinical Language Simplification: Dataset, Algorithm, and Evaluation (2022.coling-1)

Copied to clipboard

Challenge: Existing studies to translate medical jargon into layperson-understandable language focus on accuracy and readability aspects of clinical language.
Approach: They propose to construct a dataset to support automated clinical language simplification and propose a model that mimics the human annotation procedure.
Outcome: The proposed model matches human annotation procedures and achieves state-of-the-art performance compared with baselines.
JEBS: A Fine-grained Biomedical Lexical Simplification Task (2025.findings-acl)

Copied to clipboard

Challenge: Existing systems for simplification of complex medical terms are limited in the scope of their topics and require massive cost and effort to keep up with the latest research.
Approach: They propose a fine-grained lexical simplification task and dataset to enable more targeted development and evaluation of systems for replacing or explaining complex biomedical terms.
Outcome: The proposed task and dataset pave the way for development and evaluation of systems for replacing or explaining complex biomedical terms.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations