Papers by Mazda Moayeri

1 papers
DyePack: Provably Flagging Test Set Contamination in LLMs Using Backdoors (2025.emnlp-main)

Copied to clipboard

Challenge: Open benchmarks are essential for evaluating large language models, but their accessibility makes them likely targets of test set contamination.
Approach: They propose a framework that leverages backdoor attacks to flag models that used benchmark test sets during training.
Outcome: The proposed framework detects models that trained on benchmark test sets without loss of logits or internal details . it can prevent false accusations while providing strong evidence for every detected case of contamination.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations