Papers by Yerang Kim
K/DA: Automated Data Generation Pipeline for Detoxifying Implicitly Offensive Language in Korean (2025.acl-long)
Copied to clipboard
| Challenge: | Language detoxification involves removing toxicity from offensive language. |
| Approach: | They propose an automated pipeline to generate offensive language with implicit offensiveness and trend-aligned slang. |
| Outcome: | The proposed dataset exhibits high pair consistency and greater implicit offensiveness compared to existing Korean datasets and demonstrates applicability to other languages. |