Papers by Zehong Yan
TRUST-VL: An Explainable News Assistant for General Multimodal Misinformation Detection (2025.emnlp-main)
Copied to clipboard
| Challenge: | Existing methods focus on a single type of distortion and struggle to generalize to unseen scenarios. |
| Approach: | They propose a vision-language model that combines a question-aware visual amplifier module with a large-scale instruction dataset to support training. |
| Outcome: | The proposed model is able to generalize to multiple distortion types while requiring task-specific skills. |