Papers by Changbong Kim
SLM as Guardian: Pioneering AI Safety with Small Language Model (2024.emnlp-industry)
Copied to clipboard
Ohjoon Kwon, Donghyeon Jeon, Nayoung Choi, Gyu-Hwung Cho, Hwiyeol Jo, Changbong Kim, Hyunwoo Lee, Inho Kang, Sun Kim, Taiwoo Park
| Challenge: | Prior safety research on large language models focused on aligning them to safety requirements, but internalizing such safeguard features into larger models brought challenges of higher training cost and unintended degradation of helpfulness. |
| Approach: | They propose a multi-task learning mechanism that integrates harmful query detection and safeguard response into a single model. |
| Outcome: | The proposed approach outperforms the publicly available LLMs in harmful query detection and safeguard response generation. |
Taxonomy and Analysis of Sensitive User Queries in Generative AI Search System (2025.findings-naacl)
Copied to clipboard
Hwiyeol Jo, Taiwoo Park, Hyunwoo Lee, Nayoung Choi, Changbong Kim, Ohjoon Kwon, Donghyeon Jeon, Eui Hyeon Lee, Kyoungho Shin, Lim Sun Suk, Kyungmi Kim, Lee Jihye, Sun Kim
| Challenge: | generative LLMs have been used by industries for various purposes, but limited resources and limited experience hinder their deployment and maintenance. |
| Approach: | They propose a taxonomy for sensitive search queries and outline approaches to generating generative LLMs. |
| Outcome: | The proposed model can be used to analyze sensitive queries from real users. |