Papers by Jiajian Huang
Retrieving to Recover: Towards Incomplete Audio-Visual Question Answering via Semantic-consistent Purification (2026.acl-long)
Copied to clipboard
| Challenge: | Recent audio-visual question answering methods lack effective mechanisms for handling missing modalities, leading to performance degradation in real-world scenarios with data interruptions. |
| Approach: | They propose a framework that shifts the paradigm of missing modality handling to retrieval-based recovery . they leverage cross-modal retrieval via unified semantic embeddings to acquire missing domain-specific knowledge. |
| Outcome: | The proposed framework improves AVQA and enhances robustness in modal-incomplete scenarios. |