Papers by Mariane Carvalho

2 papers
Framed Multi30K: A Frame-Based Multimodal-Multilingual Dataset (2024.lrec-main)

Copied to clipboard

Challenge: Recent advances in image-captioning datasets combine image and language to solve a diverse range of tasks.
Approach: They propose a Brazilian Portuguese multimodal-multilingual dataset that extends the Multi30K dataset with 158,915 original Brazilian Portuguese descriptions and 30,104 Brazilian Portuguese translations.
Outcome: The proposed dataset adds 2,677,613 frame evocation labels to the 158,915 English descriptions and to the ones created for Brazilian Portuguese.
Frame2: A FrameNet-based Multimodal Dataset for Tackling Text-image Interactions in Video (2024.lrec-main)

Copied to clipboard

Challenge: et al., 2016) describe a multimodal dataset built from a Brazilian travel TV show . frameNet is composed of frames and their associated roles in a network of typed frame-to-frame relations.
Approach: They present a multimodal dataset built from a Brazilian travel TV show annotated for FrameNet categories for both text and image communicative modes.
Outcome: The proposed dataset includes 230 minutes of video annotated for FrameNet categories . the model can be applied to other communicative modes, i.e., images .

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations