Papers by Hui Jiao

2 papers
DiVE: Decoupling Intra-layer Visual Evidence for Mitigating Hallucinations in Large Vision-Language Models (2026.acl-long)

Copied to clipboard

Challenge: Existing decoding-based approaches do not explicitly decouple visual evidence from mixed vision–language representations.
Approach: They propose to decouple visual evidence from mixed vision–language representations by dynamically identifying layers enriched with visual information and performing intra-layer decoupling to extract aggregated visual evidence.
Outcome: Experiments show that DiVE achieves state-of-the-art performance on multiple benchmarks.
Cross-modality Data Augmentation for End-to-End Sign Language Translation (2023.findings-emnlp)

Copied to clipboard

Challenge: End-to-end sign language translation (SLT) aims to convert sign language videos into spoken language texts without intermediate representations.
Approach: They propose a cross-modality data-augmented framework to transfer gloss-to-text translation capabilities to end-to end sign language translation.
Outcome: The proposed framework outperforms baseline models on two widely used SLT datasets.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations