Papers by Hannah Babe

1 papers
StudentEval: A Benchmark of Student-Written Prompts for Large Language Models of Code (2024.findings-acl)

Copied to clipboard

Challenge: Existing CodeLLM benchmarks rely on a single expert-written prompt per problem . a growing body of work shows their utility to professional programmers .
Approach: They propose a natural-language-to-code benchmark of prompts written by non-experts . student prompts are written by 80 students who have only completed one introductory Python course .
Outcome: The proposed model is better discriminator of student prompt descriptions than existing benchmarks.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations