Challenge: Comparative study of two different approaches to build an automatic classification system for Modality values in the Portuguese language.
Approach: They propose to use a single multi-class classifier with the full Portuguese language dataset that includes eleven modal verbs and a weighted average approach to build different classifiers for each verb.
Outcome: The proposed system is based on a Portuguese language dataset with 11 modal verbs and two different classifiers, one for each verb.

Similar Papers

Multilingual and Multimodal Learning for Brazilian Portuguese (2022.lrec-1)

Copied to clipboard

Challenge: Existing models that learn multimodal and multilingual representations perform better in many natural language tasks.
Approach: They use a multimodal and multilingual corpus to test its generalization ability for other languages . they achieve a BLEU score of 51.8 and a METEOR score of 78.0 on the test set .
Outcome: The proposed model outperforms the existing model on a Portuguese-English multimodal translation task.
Finely Tuned, 2 Billion Token Based Word Embeddings for Portuguese (L18-1)

Copied to clipboard

Challenge: A distributional semantics model is instrumental to improve the performance of many applications and processing tasks for any language.
Approach: They propose to develop an advanced distributional model for Portuguese with the largest vocabulary and best evaluation scores published so far.
Outcome: The proposed model has the largest vocabulary and the best evaluation scores published so far.
Using Eye-tracking Data to Predict the Readability of Brazilian Portuguese Sentences in Single-task, Multi-task and Sequential Transfer Learning Approaches (2020.coling-main)

Copied to clipboard

Challenge: Sentence complexity assessment is a relatively new task in Natural Language Processing.
Approach: They propose to use Brazilian Portuguese to evaluate sentences with linguistic features to improve readability.
Outcome: The proposed model reaches the state-of-the-art for Brazilian Portuguese with 97.8% accuracy with linguistic features.
Evaluating Methods for Extraction of Aspect Terms in Opinion Texts in Portuguese - the Challenges of Implicit Aspects (2022.lrec-1)

Copied to clipboard

Challenge: In aspect-based sentiment analysis, the implicit mention of aspects is difficult to identify and may require world knowledge to do so.
Approach: They evaluate frequency-based, hybrid, and machine learning methods to extract aspect terms from opinionated texts in Portuguese.
Outcome: The proposed methods show that they are more efficient and more efficient than previous methods.
Comparing Feature-Engineering and Feature-Learning Approaches for Multilingual Translationese Classification (2021.emnlp-main)

Copied to clipboard

Challenge: Traditional hand-crafted features have been used for distinguishing between translated and original non-translated texts.
Approach: They compare a feature-engineering-based approach to a features-learning-based one and use pre-trained neural word embeddings to train neural architectures.
Outcome: The proposed approach outperforms other approaches by more than 20 accuracy points and the BERT-based model performs the best in both monolingual and multilingual settings.
DORE: A Dataset for Portuguese Definition Generation (2024.lrec-main)

Copied to clipboard

Challenge: Definition modelling (DM) is the task of automatically generating a dictionary definition of a specific word.
Approach: They propose to create a dataset for definition modelling for Portuguese with 100,000 definitions and evaluate several deep learning based DM models on the dataset.
Outcome: The proposed dataset will facilitate research and study of Portuguese in wider contexts.
ModaFact: Multi-paradigm Evaluation for Joint Event Modality and Factuality Detection (2025.coling-main)

Copied to clipboard

Challenge: NLP studies have mostly dealt with factuality and modality separately . linguistic modality conveys the relationship a situation is supposed to have with respect to wishes, norms, goals, authority, etc.
Approach: They propose a resource with joint factuality and modality information for event-denoting expressions in Italian.
Outcome: The proposed resource is consistent with existing ones and compares classification systems trained on italy's ModaFact dataset and best-performing model.
Framed Multi30K: A Frame-Based Multimodal-Multilingual Dataset (2024.lrec-main)

Copied to clipboard

Challenge: Recent advances in image-captioning datasets combine image and language to solve a diverse range of tasks.
Approach: They propose a Brazilian Portuguese multimodal-multilingual dataset that extends the Multi30K dataset with 158,915 original Brazilian Portuguese descriptions and 30,104 Brazilian Portuguese translations.
Outcome: The proposed dataset adds 2,677,613 frame evocation labels to the 158,915 English descriptions and to the ones created for Brazilian Portuguese.
A Study of Syntactic Multi-Modality in Non-Autoregressive Machine Translation (2022.naacl-main)

Copied to clipboard

Challenge: Non-autoregressive translation models suffer from the multi-modality problem when a source sentence corresponds to multiple correct translations.
Approach: They propose to decompose the syntactic multi-modality problem into short- and long-range models and evaluate them on synthesized and real datasets.
Outcome: The proposed loss functions can handle short- and long-range syntactic multi-modalities better than existing models.
Multimodal Frame Identification with Multilingual Evaluation (N18-1)

Copied to clipboard

Challenge: FrameNet Semantic Role Labeling aims to disambiguate situations around predicates using textual representations.
Approach: They extend a frame identification task to leverage multimodal representations to improve FrameNet Semantic Role Labeling.
Outcome: The proposed system outperforms its unimodal counterpart on the English frameNet and its German counterpart on IMAGINED words.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations