Papers by Hari Bandi
Debate, Deliberate, Decide (D3): A Cost-Aware Adversarial Framework for Reliable and Interpretable LLM Evaluation (2026.eacl-long)
Copied to clipboard
| Challenge: | Existing evaluation tools for Large Language Models (LLMs) are inconsistency, bias, and lack of transparent decision criteria. |
| Approach: | They propose a cost-aware, adversarial multi-agent framework that orchestrates structured debate among role-specialized agents to produce reliable and interpretable evaluations. |
| Outcome: | The proposed framework orchestrates structured debate among role-specialized agents to produce reliable and interpretable evaluations. |