Papers by Yuhan Ke

2 papers
RoleLLM: Benchmarking, Eliciting, and Enhancing Role-Playing Abilities of Large Language Models (2024.findings-acl)

Copied to clipboard

Challenge: Large Language Models (LLMs) have paved the way for complex tasks such as role-playing.
Approach: They propose a framework to benchmark, elicit, and enhance role-playing abilities in Large Language Models.
Outcome: The proposed framework improves role-playing abilities with 168,093 samples.
EntroBench: Evaluating LLM Watermarking Under Multi-Entropy Scenarios and Practical User Operations (2026.findings-acl)

Copied to clipboard

Challenge: Existing evaluations of large language models (LLMs) watermarking are limited to fixed entropy settings.
Approach: They propose a benchmark for LLM watermarking that systematically covers three entropy levels and seven representative tasks.
Outcome: The proposed framework covers three entropy levels and seven representative tasks.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations