Papers by Ronghang Hu

2 papers
Are You Looking? Grounding to Multiple Modalities in Vision-and-Language Navigation (P19-1)

Copied to clipboard

Challenge: Existing models that ground language into visual appearance and route structure are outperforming their visual counterparts in unseen new environments.
Approach: They propose to decompose the grounding procedure into a set of expert models with access to different modalities and ensemble them at prediction time.
Outcome: The proposed model outperforms models with only route structure and visual features on the benchmark Room-to-Room dataset.
No Free Lunch: Retrieval-Augmented Generation Undermines Fairness in LLMs, Even for Vigilant Users (2025.findings-emnlp)

Copied to clipboard

Challenge: Retrieval-augmented generation is widely adopted for its effectiveness and cost-efficiency in mitigating hallucinations.
Approach: They propose a practical three-level threat model from the perspective of user fairness awareness.
Outcome: The proposed model shows that RAG can undermine fairness alignment without fine-tuning or retraining.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations