Challenge: a recent study shows that the class labels of german documents containing ADRs are imbalanced . clinical trials and physicians prescribing medications cannot cover every potential use case.
Approach: They propose to use binary annotated documents from a german patient forum to detect ADRs.
Outcome: The proposed model achieves an F1 score of 37.52 for the positive class on the German patient forum.

Similar Papers

A Dataset for Pharmacovigilance in German, French, and Japanese: Annotating Adverse Drug Reactions across Languages (2024.lrec-main)

Copied to clipboard

Challenge: Existing clinical corpora mostly revolves around scientific articles in English . existing literature is limited to only a few scientific articles .
Approach: They propose to use user-generated data sources to uncover adverse drug reactions . existing clinical corpora mostly revolves around scientific articles in english . authors provide statistics to highlight certain challenges associated with the corpus .
Outcome: The proposed corpus includes 12 entity types, four attribute types, and 13 relation types . it provides strong baselines for extracting entities and relations between entities .
Annotation of Adverse Drug Reactions in Patients’ Weblogs (2020.lrec-1)

Copied to clipboard

Challenge: Adverse drug reactions are a severe problem that significantly degrade quality of life and make the therapeutic approach unacceptable.
Approach: They crawled patient’s weblog articles shared on an online patient-networking platform and annotated the effects of drugs therein reported.
Outcome: The proposed dataset is unique for the richness of annotated information, including detailed descriptions of drug reactions with full context.
Training Data Augmentation for Detecting Adverse Drug Reactions in User-Generated Content (D19-1)

Copied to clipboard

Challenge: Existing dictionary-based, semi-supervised learning approaches are limited by the coverage and maintainability of laymen health vocabularies.
Approach: They propose a data augmentation approach that leverages variational autoencoders to learn high-quality data distributions from a large unlabeled dataset and generate a small set of labeled training sets.
Outcome: The proposed approach matches the performance of fully-supervised approaches while requiring only 25% of training data.
MedErrBench: A Fine-Grained Multilingual Benchmark for Medical Error Detection and Correction with Clinical Expert Annotations (2026.findings-acl)

Copied to clipboard

Challenge: Existing or generated clinical text may contain inaccuracies that can lead to serious adverse outcomes.
Approach: They introduce a multilingual benchmark for error detection, localization and correction . they assessed the performance of a range of general-purpose, language-specific, and medical-domain language models .
Outcome: The proposed benchmark covers English, Arabic and Chinese, with natural medical cases annotated and reviewed by domain experts.
Enhancing Adverse Drug Event Detection with Multimodal Dataset: Corpus Creation and Model Development (2024.findings-acl)

Copied to clipboard

Challenge: ADEs are a serious public health concern and cost healthcare systems billions of dollars . despite advancements in healthcare, ADE detection remains a significant challenge .
Approach: They propose a multimodal adverse drug event detection dataset that merges ADE-related textual information with visual aids to enhance patient safety.
Outcome: The proposed dataset integrates ADE-related textual information with visual aids to improve patient safety and healthcare accessibility.
Detecting Adverse Drug Reactions from Biomedical Texts with Neural Networks (P19-2)

Copied to clipboard

Challenge: Detection of adverse drug reactions in post-marketing period is a crucial challenge for pharmacology.
Approach: They propose to use social media to extract information about adverse drug reactions . they compare four state-of-the-art attention-based neural networks to the F-measure .
Outcome: The proposed methods perform better on four different benchmarks.
From Witch’s Shot to Music Making Bones - Resources for Medical Laymen to Technical Language and Vice Versa (2020.lrec-1)

Copied to clipboard

Challenge: Information we share online unveils directly or indirectly information about our lifestyle and health situation.
Approach: They propose a dataset which annotates medical laymen and technical expressions in a patient forum and a set of medical synonyms and definitions.
Outcome: The proposed dataset annotates medical laymen and technical expressions in a patient forum along with a set of medical synonyms and definitions.
Offensive language detection in Hebrew: can other languages help? (2022.lrec-1)

Copied to clipboard

Challenge: Various approaches for offensive language detection have been applied for this task . contamination of social networks with offensive content is a new reality affecting almost all of us .
Approach: They propose to use multiple supervised models and text representations to detect offensive language in three languages, including two Semitic languages.
Outcome: The proposed model can detect offensive content in two Semitic languages, including Hebrew and Arabic, and it is able to perform cross-lingual and multilingual learning.
No offence, Bert - I insult only humans! Multilingual sentence-level attack on toxicity detection networks (2023.findings-emnlp)

Copied to clipboard

Challenge: a new sentence-level attack on toxic detection models is shown to work on seven languages . toxicity detection systems are used to silence the voices of criticism, causing echo chambers .
Approach: They propose a sentence-level attack that adds positive words to a hateful message . they show the attack works on seven languages from three different language families .
Outcome: The proposed attack is shown to work on seven languages from three different language families.
A Corpus with Multi-Level Annotations of Patients, Interventions and Outcomes to Support Language Processing for Medical Literature (P18-1)

Copied to clipboard

Challenge: In 2015 alone, about 100 manuscripts describing randomized controlled trials for medical interventions were published every day.
Approach: They propose a corpus of 5,000 medical articles annotated with demarcations of text spans that describe the Patient population enrolled, the Interventions studied and to what they were Compared, and the Outcomes measured.
Outcome: The proposed corpus includes 5,000 medical articles describing clinical randomized controlled trials.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations