Papers by Youngjun Choi

2 papers
FinHarmBench: Financial Jailbreak Benchmark and Unsupervised Safety Fine-Tuning via Refusal Steering Distillation (2026.acl-industry)

Copied to clipboard

Challenge: Existing safety benchmarks focus on general harms and lack the granularity needed to capture domain-specific financial threats.
Approach: They propose a benchmark to evaluate financially harmful and confusable benign prompts.
Outcome: The proposed framework improves refusal behavior without annotating refusal responses.
LBC: Language-Based-Classifier for Out-Of-Variable Generalization (2025.naacl-long)

Copied to clipboard

Challenge: Large Language Models (LLMs) have great success in natural language processing tasks such as response generation, but their performance on tabular data tasks has been limited due to their inferior performance compared to traditional machine learning models (TMLs).
Approach: They propose a Language-Based-Classifier (LBC) that maximizes the benefits of LLMs to outperform TMLs on OOV tasks.
Outcome: The proposed model outperforms TMLs on OOV tasks by using three key methods.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations