Papers by Luzhe Sun
HALP: Detecting Hallucinations in Vision-Language Models without Generating a Single Token (2026.eacl-long)
Copied to clipboard
| Challenge: | Existing methods for detection of hallucinations operate after text generation, making intervention costly and untimely. |
| Approach: | They examine whether hallucination risk can instead be predicted before any token is generated by probing a model's internal representations in a single forward pass. |
| Outcome: | The proposed model can detect hallucinations before token generation, while query-token representations can be more accurate. |