Papers with EST
Detecting Proxy Gaming in RL and LLM Alignment via Evaluator Stress Tests (2026.findings-acl)
Copied to clipboard
| Challenge: | Proxy optimization is a challenge spanning reinforcement learning and LLM alignment. |
| Approach: | They propose an invariance-based framework that detects proxy gaming by separating exploitable sensitivity from content-driven improvements using semantic validity audits. |
| Outcome: | The proposed framework achieves 78.4% precision and 81.7% recall across 15 environments and 5 algorithms. |
Evolving Beyond Snapshots: Harmonizing Structure and Sequence via Entity State Tuning for Temporal Knowledge Graph Forecasting (2026.acl-long)
Copied to clipboard
| Challenge: | Temporal knowledge graphs (TKGs) require predicting future facts by modeling structural dependencies within each snapshot and temporal evolution across snapshots. |
| Approach: | They propose an encoder-agnostic framework that provides persistent entity states . EST maintains a global state buffer and aligns structural evidence with sequential signals . |
| Outcome: | Experiments show that EST improves diverse backbones and achieves state-of-the-art performance. |