Papers by Mahmoud Khalil

2 papers
The Unintended Trade-off of AI Alignment: Balancing Hallucination Mitigation and Safety in LLMs (2026.findings-eacl)

Copied to clipboard

Challenge: Hallucination in large language models has been studied, but a side effect remains unrecognized . a new study examines the trade-off between truthfulness and safety alignment .
Approach: They propose a method that disentangles hallucination from hallucinian features using sparse autoencoders.
Outcome: The proposed method preserves refusal behavior and task utility while maintaining safety alignment.
An Information-theoretic Approach to Prompt Engineering Without Ground Truth Labels (2022.acl-long)

Copied to clipboard

Challenge: Existing prompt engineering methods require labeled data and access to model parameters . a new method for selecting prompt templates without labeles and without direct access to the model is needed.
Approach: They propose a method for selecting prompt templates without labeled examples and without direct access to the model.
Outcome: The proposed method performs at almost oracle levels, without labels, on 7 datasets representing 7 different NLP tasks.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations