Papers with Beemo
Beemo: Benchmark of Expert-edited Machine-generated Outputs (2025.naacl-long)
Copied to clipboard
Ekaterina Artemova, Jason S Lucas, Saranya Venkatraman, Jooyoung Lee, Sergei Tilga, Adaku Uchendu, Vladislav Mikhailov
| Challenge: | Existing benchmarks for machine-generated texts (MGTs) include single-author texts (human-written and machine-generated). |
| Approach: | They propose to benchmark machine-generated outputs (Beemo) which includes 6.5k texts written by humans, generated by ten instruction-finetuned LLMs, and edited by experts for various use cases. |
| Outcome: | The proposed benchmark includes 6.5k texts written by humans, generated by ten instruction-finetuned LLMs, and edited by experts for various use cases, ranging from creative writing to summarization. |