Papers by JoonHo Lee
Preference Consistency Matters: Enhancing Preference Learning in Language Models with Automated Self-Curation of Training Corpora (2025.naacl-long)
Copied to clipboard
| Challenge: | Existing methods to address inconsistencies in preference learning datasets rely on heuristics to achieve alignment. |
| Approach: | They propose a method that preprocesses annotated datasets by leveraging proxy models trained directly on them to detect and select consistent annotations. |
| Outcome: | The proposed method shows performance improvements of up to 33% across learning algorithms and proxy capabilities. |