Papers by Marc Versage

1 papers
InfoPO: On Mutual Information Maximization for Large Language Model Alignment (2025.naacl-long)

Copied to clipboard

Challenge: Recent studies have shown that direct preference optimization and its variants can be useful for fine-tuning large language models with human preferences data.
Approach: They propose a preference fine-tuning algorithm that effectively and efficiently aligns large language models using preference data.
Outcome: Extensive experiments show that the proposed algorithm outperforms established baselines on reasoning tasks.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations