Papers by Xianpeng Guo
MagicBench: Diagnosing Visual Agency Loss and Semantic Dependency in Multimodal LLMs (2026.acl-long)
Copied to clipboard
| Challenge: | MLLMs assume linguistic context invariably enhances visual understanding . a diagnostic benchmark is used to evaluate ML models under hierarchical linguistic interference . |
| Approach: | They propose a diagnostic benchmark to evaluate MLLMs under hierarchical linguistic interference. |
| Outcome: | The proposed benchmark compared 402 videos with a physical constraint set to evaluate MLLMs under hierarchical linguistic interference. |