Papers by Lorenzo Lupo
Divide and Rule: Effective Pre-Training for Context-Aware Multi-Encoder Translation Models (2022.acl-long)
Copied to clipboard
| Challenge: | Multi-encoder models aim to improve translation quality by encoding document-level contextual information alongside the current sentence. |
| Approach: | They propose to pre-train contextual parameters over split sentence pairs to improve contextual encoding . they propose four different splitting methods to improve learning of contextual parameters . |
| Outcome: | The proposed model improves learning of contextual parameters, both in low and high resource settings. |
DADIT: A Dataset for Demographic Classification of Italian Twitter Users and a Comparison of Prediction Methods (2024.lrec-main)
Copied to clipboard
| Challenge: | Social scientists increasingly use demographically stratified social media data to study attitudes, beliefs, and behavior of the general public. |
| Approach: | They validated the DADIT dataset of 30M tweets of 20k Italian Twitter users, along with their bios and profile pictures. |
| Outcome: | The best XLM-based classifier improves upon the commonly used competitor M3 by up to 53% F1. |