Papers with REFORM

2 papers
Teach a Reward Model to Correct Itself: Reward Guided Adversarial Failure Discovery for Robust Reward Modeling (2026.acl-long)

Copied to clipboard

Challenge: Existing failure discovery methods rely on prior knowledge of preference attributes . Existing methods do not scale to new models or data.
Approach: They propose a preference distribution agnostic procedure that uses the reward model itself to guide controlled decoding toward mis specified responses while preserving the underlying preference class.
Outcome: The proposed procedure improves robustness without degrading reward quality across models.
Cultivating Forensic Reasoning for Generalizable Multimodal Manipulation Detection (2026.acl-long)

Copied to clipboard

Challenge: Existing methods for manipulation detection and grounding focus on manipulator type classification under result-oriented supervision.
Approach: They propose a reasoning-driven framework that shifts learning from outcome fitting to process modeling.
Outcome: The proposed framework achieves state-of-the-art with superior generalization on large-scale datasets.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations