Papers by Abhishek Lalwani

1 papers
Alternate Preference Optimization for Unlearning Factual Knowledge in Large Language Models (2025.coling-main)

Copied to clipboard

Challenge: Existing methods for large language models rely on negative feedback to suppress responses related to the forget set, which often results in nonsensical or inconsistent outputs, diminishing model utility and posing potential privacy risks.
Approach: They propose an approach which combines negative feedback with in-domain positive feedback on the forget set and introduces new evaluation metrics to assess the quality of responses related to the forget sets.
Outcome: The proposed approach avoids undesirable model behaviors while maintaining overall model performance.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations