Papers by Yuhan Ke
RoleLLM: Benchmarking, Eliciting, and Enhancing Role-Playing Abilities of Large Language Models (2024.findings-acl)
Copied to clipboard
Noah Wang, Z.y. Peng, Haoran Que, Jiaheng Liu, Wangchunshu Zhou, Yuhan Wu, Hongcheng Guo, Ruitong Gan, Zehao Ni, Jian Yang, Man Zhang, Zhaoxiang Zhang, Wanli Ouyang, Ke Xu, Wenhao Huang, Jie Fu, Junran Peng
| Challenge: | Large Language Models (LLMs) have paved the way for complex tasks such as role-playing. |
| Approach: | They propose a framework to benchmark, elicit, and enhance role-playing abilities in Large Language Models. |
| Outcome: | The proposed framework improves role-playing abilities with 168,093 samples. |
EntroBench: Evaluating LLM Watermarking Under Multi-Entropy Scenarios and Practical User Operations (2026.findings-acl)
Copied to clipboard
| Challenge: | Existing evaluations of large language models (LLMs) watermarking are limited to fixed entropy settings. |
| Approach: | They propose a benchmark for LLM watermarking that systematically covers three entropy levels and seven representative tasks. |
| Outcome: | The proposed framework covers three entropy levels and seven representative tasks. |