Distantly Supervised Named Entity Recognition using Positive-Unlabeled Learning (P19-1)
Copied to clipboard
| Challenge: | Empirical studies on four public NER datasets demonstrate the effectiveness of our proposed method. |
| Approach: | They propose a method to perform named entity recognition using unlabeled data and named entity dictionaries. |
| Outcome: | The proposed method can estimate task loss as if there is fully labeled data. |
Similar Papers
Distantly Supervised Named Entity Recognition via Confidence-Based Multi-Class Positive and Unlabeled Learning (2022.acl-long)
Copied to clipboard
| Challenge: | Existing methods for named entity recognition suffer from incomplete annotations due to incompleteness of external knowledge bases. |
| Approach: | They propose a method to solve the named entity recognition problem under distant supervision using dictionaries and knowledge bases. |
| Outcome: | The proposed method outperforms existing methods on two benchmark datasets labeled by various knowledge bases. |
Distantly-Supervised Named Entity Recognition with Noise-Robust Learning and Language Model Augmented Self-Training (2021.emnlp-main)
Copied to clipboard
| Challenge: | Named entity recognition models require abundant high-quality annotations to train . distant supervision may induce incomplete and noisy labels, making supervised learning ineffective. |
| Approach: | They propose a noise-robust learning scheme for training named entity recognition models using only distantly-labeled data and a self-training method that uses contextualized augmentations created by pre-trained language models. |
| Outcome: | The proposed method outperforms existing supervised NER models on three datasets by significant margins. |
Named Entity Recognition without Labelled Data: A Weak Supervision Approach (2020.acl-main)
Copied to clipboard
| Challenge: | Named Entity Recognition (NER) performance often degrades when applied to target domains that differ from the texts observed during training. |
| Approach: | They propose a method to learn NER models in the absence of labelled data through weak supervision by using a broad spectrum of labelling functions to automatically annotate texts from the target domain. |
| Outcome: | The proposed approach improves on two English datasets and shows that it improves by 7 percentage points on entity-level F1 scores compared to an out-of-domain neural NER model. |
Noise-Robust Training with Dynamic Loss and Contrastive Learning for Distantly-Supervised Named Entity Recognition (2023.findings-acl)
Copied to clipboard
| Challenge: | Named entity recognition (NER) is a task in natural language processing that aims at locating entity mentions in a given sentence and assigning them to certain types. |
| Approach: | They propose to use a dynamic loss function to better adapt to the changing noise during the training process and incorporate token level contrastive learning to fully utilize the noisy data. |
| Outcome: | The proposed method outperforms existing NER models on three benchmark datasets and outperformed existing models by significant margins. |
Re-Examine Distantly Supervised NER: A New Benchmark and a Simple Approach (2025.coling-main)
Copied to clipboard
| Challenge: | Existing DS-NER approaches rely on large validation sets and test set for tuning inappropriately. |
| Approach: | They propose a method where training data is annotated using domain dictionaries and test data is analyzed by domain experts. |
| Outcome: | The proposed method reduces the need for labor-intensive manual annotations but rely on large human labeled validation set. |
Toward Recognizing More Entity Types in NER: An Efficient Implementation using Only Entity Lexicons (2020.findings-emnlp)
Copied to clipboard
| Challenge: | Existing named entity recognition systems require large scale labeled data to perform, while annotation of NER data is laborious and time-consuming. |
| Approach: | They propose to adjust an existing named entity recognition system to recognize entity types not defined in the system. |
| Outcome: | The proposed method can be quickly adjusted to a named entity recognition system. |
Better Modeling of Incomplete Annotations for Named Entity Recognition (N19-1)
Copied to clipboard
| Challenge: | Existing approaches to named entity recognition (NER) assume that the training data is fully annotated with named entity information. |
| Approach: | They propose a supervised setup for named entity recognition where annotated data is assumed to be available during training. |
| Outcome: | The proposed approach is able to recognize named entities with incomplete annotations. |
Reinforcement-based denoising of distantly supervised NER with partial annotation (D19-61)
Copied to clipboard
| Challenge: | Existing named entity recognition systems rely on large amounts of human-labeled data for supervision, but the result is noisy. |
| Approach: | They propose to use partial annotation to address false negative cases and implement a reinforcement learning strategy to identify false positive instances. |
| Outcome: | The proposed model reduces the amount of manually annotated data required to perform NER in a new domain. |
Label Refinement via Contrastive Learning for Distantly-Supervised Named Entity Recognition (2022.findings-naacl)
Copied to clipboard
| Challenge: | Existing methods to locate and classify entities using knowledge bases and unlabeled corpus are expensive and limited application. |
| Approach: | They propose to use a method to directly learn the distant label refinement knowledge by imitating annotations of different qualities and comparing them in contrastive learning frameworks. |
| Outcome: | The proposed method can give modified suggestions on distant data without additional supervised labels and thus reduces the requirement on the quality of the knowledge bases. |
Sampling Better Negatives for Distantly Supervised Named Entity Recognition (2023.findings-acl)
Copied to clipboard
| Challenge: | Existing supervised named entity recognition approaches rely on human annotations. |
| Approach: | They propose a method to select negative samples with high similarities with positive samples . they propose to use automatically labeled training data instead of human annotations . |
| Outcome: | The proposed method achieves consistent performance improvements on four distantly supervised NER datasets. |