Papers by Dan Oneață
Multilingual Multimodal Learning with Machine Translated Text (2022.findings-emnlp)
Copied to clipboard
| Challenge: | Currently, most vision-and-language pretraining research focuses on English tasks due to the availability of datasets. |
| Approach: | They propose a framework for machine translating English multimodal data to improve training data . they propose two metrics to prevent models from learning from low-quality translated text . |
| Outcome: | The proposed framework can be applied to any multimodal dataset and model. |