Papers by Roser Morante

10 papers
Systems’ Agreements and Disagreements in Temporal Processing: An Extensive Error Analysis of the TempEval-3 Task (L18-1)

Copied to clipboard

Challenge: Temporal Processing systems are crucial for timelines and storylines . TempEval-3 is the latest evaluation campaign on open-domain TP in english .
Approach: They present a Temporal Processing system that incorporates high level lexical semantic features and uses them to evaluate temporal relation classification.
Outcome: The proposed system achieves the best scores for event detection and temporal relation classification from raw text, but the errors are not as robust as previous systems.
Annotating Perspectives on Vaccination (2020.lrec-1)

Copied to clipboard

Challenge: Vaccination corpus is a corpus of texts related to the online vaccination debate . it contains documents from the Internet which reflect different views on vaccinations .
Approach: They present a corpus of texts related to the online vaccination debate annotated with perspectives about attribution, claims and opinions.
Outcome: The Vaccination Corpus contains 294 documents from the Internet which reflect different views on vaccinations.
Detecting Negation Cues and Scopes in Spanish (2020.lrec-1)

Copied to clipboard

Challenge: Negation is a phenomenon that "relates an expression e to another expression with a meaning that is in some way opposed to the meaning of e" previous work on negation in English has focused mostly and only recently on annotation tasks.
Approach: They propose a machine learning system that processes negation in Spanish . they use a corpus from the SFU corpus to perform two tasks .
Outcome: The proposed system outperforms state-of-the-art in negation cue detection and scope identification.
Must Children be Vaccinated or not? Annotating Modal Verbs in the Vaccination Debate (2020.lrec-1)

Copied to clipboard

Challenge: In this paper we analyze the use of modal verbs in a corpus of texts related to the vaccination debate.
Approach: They analyze the use of modal verbs in a corpus of texts related to the vaccination debate.
Outcome: The use of modal verbs in the vaccination debate is analysed using modal auxiliaries and a corpus of texts.
Scoring and Classifying Implicit Positive Interpretations: A Challenge of Class Imbalance (C18-1)

Copied to clipboard

Challenge: a reimplementation of a system on detecting implicit positive meaning from negated statements is reported . a baseline taking the mean score or most frequent class is hard to beat because of class imbalance in the dataset.
Approach: They propose a system to detect implicit positive meaning from negated statements . they convert the scores into classes and report their results on regression and classification tasks .
Outcome: The proposed system is hard to beat because of class imbalance in the dataset.
A Web Portal about the State of the Art of NLP Tasks in Spanish (2024.lrec-main)

Copied to clipboard

Challenge: a web portal has been created with information about the state of the art of natural language processing tasks in Spanish.
Approach: They propose a web portal that provides information about the state of the art of natural language processing tasks in Spanish.
Outcome: The portal provides information about forums, competitions, tasks and datasets in Spanish that would otherwise be spread in multiple articles and web sites.
Identifying Copied Fragments in a 18th Century Dutch Chronicle (2022.lrec-1)

Copied to clipboard

Challenge: We use stylometric methods to identify which fragments of the manuscript represent the author’s own original work and which show signs of external source use.
Approach: They apply computational stylometric techniques to an 18th century Dutch chronicle to determine which fragments represent the author's original work and which show signs of external source use.
Outcome: The proposed method is effective for authorship verification of the Dutch chronicle, but less effective when personal writing style is masked by author independent styles or when applied to paraphrased text.
A review of Spanish corpora annotated with negation (C18-1)

Copied to clipboard

Challenge: Existing corpora annotated with negation information are small and not always compatible . negation is a linguistic phenomenon that is not addressed in English .
Approach: They review existing corpora annotated with negation in Spanish and analyze compatibility . they propose to develop a supervised negation processing system for Spanish .
Outcome: The proposed system will not be able to merge the small corpora in Spanish due to lack of compatibility in annotations.
Resource Interoperability for Sustainable Benchmarking: The Case of Events (L18-1)

Copied to clipboard

Challenge: Despite efforts to improve interoperability, there are still problems with benchmark corpora that are hampered by too laborious conversion steps.
Approach: They assess aspects of interoperability at the document-level across 20 annotated corpora and compare their compatibility and consistency across the corpors.
Outcome: The proposed framework enables the analysis of document intersections between the corpora and shows their compatibility and consistency across the corpus.
Bilingual Evaluation of Language Models on General Knowledge in University Entrance Exams with Minimal Contamination (2025.coling-main)

Copied to clipboard

Challenge: Existing benchmarks for Large Language Models have been proposed as single-task evaluations, but they are not fully comprehensive.
Approach: They present a bilingual dataset that contains 1003 multiple-choice questions in Spanish and English.
Outcome: The proposed model ranking is almost identical to the one obtained with MMLU .

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations