Papers by Kuang-Da Wang
Benchmarking Agentic Newswriting via Journalistic Workflows (2026.findings-acl)
Copied to clipboard
| Challenge: | Recent advances in autonomous digital agents highlight their potential for structured tasks through autonomous decision-making and task decomposition, but it remains unclear how well such systems support real-world information-intensive workflows. |
| Approach: | They propose a benchmark to evaluate how journalists can use agents to organize and organize information from the web. |
| Outcome: | The proposed system can be used to iterate and evaluate newswriting tasks in real-world situations. |
Extending Automatic Machine Translation Evaluation to Book-Length Documents (2025.emnlp-main)
Copied to clipboard
Kuang-Da Wang, Shuoyang Ding, Chao-Han Huck Yang, Ping-Chun Hsieh, Wen-Chih Peng, Vitaly Lavrukhin, Boris Ginsburg
| Challenge: | Large Language Models (LLMs) have superior translation performance and long-context capabilities, but evaluation methodologies remain constrained to sentence-level assessment due to dataset limitations and token number restrictions in metrics. |
| Approach: | They propose an evaluation scheme that extends existing automatic metrics to long-document translation by treating documents as continuous text and applying sentence segmentation and alignment methods. |
| Outcome: | The proposed evaluation scheme outperforms existing long-form document evaluation schemes while accounting for under-/over-translations and varied sentence boundaries. |