Evaluating Readability Metrics for German Medical Text Simplification (2025.coling-main)
Copied to clipboard
| Challenge: | Clinical reports and scientific information sources are written for medical experts, preventing patients from understanding the main messages of these texts. |
| Approach: | They evaluated the suitability of 18 statistical, part-of-speech-based, syntactic, semantic and fluency metrics to measure readability of German medical texts. |
| Outcome: | The proposed measures are compared with standard methods on English medical texts and simplified summaries. |
Similar Papers
A Corpus for Automatic Readability Assessment and Text Simplification of German (2020.lrec-1)
Copied to clipboard
| Challenge: | Using monolingual-only data, we can automate readability assessment and text simplification of simplified language. |
| Approach: | They present a corpus for automatic readability assessment and automatic text simplification for German using parallel and monolingual data. |
| Outcome: | The proposed corpus is compiled from web sources and contains information on text structure, typography, font style, and images. |
Subjective Text Complexity Assessment for German (2022.lrec-1)
Copied to clipboard
| Challenge: | Often, readability is defined as how easily a written text is to read. |
| Approach: | They propose to use a corpus of sentences provided by a German IT service provider to assess the readability of German text. |
| Outcome: | The proposed model can predict complexity of German text by using linguistically motivated features. |
DETECT: Determining Ease and Textual Clarity of German Text Simplifications (2026.eacl-long)
Copied to clipboard
| Challenge: | Current evaluation of German automatic text simplification relies on general-purpose metrics such as SARI, BLEU, and BERTScore. |
| Approach: | They propose a German-specific metric that holistically evaluates ATS quality across all three dimensions of simplicity, meaning preservation, and fluency. |
| Outcome: | The proposed metric achieves higher correlations with human judgments than widely used ATS metrics. |
Paragraph-level Simplification of Medical Texts (2021.naacl-main)
Copied to clipboard
| Challenge: | Existing methods for simplification of medical texts are limited due to jargon and technical content. |
| Approach: | They propose to automate the simplification of medical texts by penalizing decoders for producing "jargon" terms. |
| Outcome: | The proposed method improves on existing heuristics by penalizing the decoder for producing "jargon" terms. |
Comparing and Developing Tools to Measure the Readability of Domain-Specific Texts (D19-1)
Copied to clipboard
Elissa Redmiles, Lisa Maszkiewicz, Emily Hwang, Dhruv Kuchhal, Everest Liu, Miraida Morales, Denis Peskov, Sudha Rao, Rock Stevens, Kristina Gligorić, Sean Kross, Michelle Mazurek, Hal Daumé III
| Challenge: | Despite this, we lack a thorough understanding of how to validly measure readability at scale, especially for domain-specific texts. |
| Approach: | They present a comparison of the validity of well-known readability measures and introduce a novel approach to measure readability at scale. |
| Outcome: | The proposed approach addresses shortcomings of existing measures. |
Evaluating the Evaluators: Are readability metrics good measures of readability? (2025.emnlp-main)
Copied to clipboard
| Challenge: | Plain language summarization (PLS) aims to distill complex documents into accessible summaries for non-expert audiences. |
| Approach: | They conduct a thorough survey of literature on plain language summarization (PLS) and find that traditional readability metrics are not compared to human judgments. |
| Outcome: | The proposed language models better capture deeper measures of readability, with the best-performing model achieving a Pearson correlation of 0.56 with human judgments. |
Modeling the Readability of German Targeting Adults and Children: An empirically broad analysis and its cross-corpus validation (C18-1)
Copied to clipboard
| Challenge: | a new corpus of german news broadcast subtitles is compiled and crawled . readability assessment is a task of linking a text to the appropriate target audience based on its complexity. |
| Approach: | They analyze two German educational media texts targeting adults and children . they use 400 automatically extracted measures of linguistic complexity from a wide range of linguistic domains . their most successful binary classification model for german readability shows high accuracy . |
| Outcome: | The proposed model shows high accuracy between 89.4%–98.9% for both data sets. |
Multilingual Simplification of Medical Texts (2023.emnlp-main)
Copied to clipboard
Sebastian Joseph, Kathryn Kazanas, Keziah Reina, Vishnesh Ramanathan, Wei Xu, Byron Wallace, Junyi Jessy Li
| Challenge: | Existing work on medical text simplification has focused on monolingual settings . important findings in medicine are typically presented in technical, jargon-laden language . text simulating models can generate viable simplified texts, but there are outstanding challenges . |
| Approach: | They propose a dataset for medical text simplification in four languages . they evaluate fine-tuned and zero-shot models across these languages based on human assessments and analyses . |
| Outcome: | The proposed dataset evaluates models in English, Spanish, French, and Farsi . it shows that the models can generate viable simplified texts, but there are challenges . |
Benchmarking Automated Clinical Language Simplification: Dataset, Algorithm, and Evaluation (2022.coling-1)
Copied to clipboard
| Challenge: | Existing studies to translate medical jargon into layperson-understandable language focus on accuracy and readability aspects of clinical language. |
| Approach: | They propose to construct a dataset to support automated clinical language simplification and propose a model that mimics the human annotation procedure. |
| Outcome: | The proposed model matches human annotation procedures and achieves state-of-the-art performance compared with baselines. |
MedReadMe: A Systematic Study for Fine-grained Sentence Readability in Medical Domain (2024.emnlp-main)
Copied to clipboard
| Challenge: | Using fine-grained readability measures is the first step towards making medical texts more accessible. |
| Approach: | They propose a dataset MedReadMe which measures sentences and complex spans with an annotation tool. |
| Outcome: | The proposed dataset covers 650 linguistic features and additional complex span features, and is compared against state-of-the-art methods using large language models. |