Challenge: Student Evaluations of Teaching (SETs) are used in colleges and universities to assess student perceptions about their courses.
Approach: They propose a system that leverages sentiment analysis, aspect extraction, summarization and visualization techniques to provide organized illustrations of SET findings to instructors and other reviewers.
Outcome: The proposed system can be used by 10 professors from diverse departments to analyze SET results.

Similar Papers

ReflectSumm: A Benchmark for Course Reflection Summarization (2024.lrec-main)

Copied to clipboard

Challenge: Existing research has focused on standard summarization benchmarks within domains like news, scientific articles, and opinions.
Approach: They propose a summarization dataset specifically designed for summarizing students’ reflective writing.
Outcome: The proposed summarization dataset can be used in opinion summarizing scenarios and in educational domains.
BOOKSUM: A Collection of Datasets for Long-form Narrative Summarization (2022.findings-emnlp)

Copied to clipboard

Challenge: Existing text summarization datasets include short-form source documents that lack long-range causal and temporal dependencies and contain strong layout and stylistic biases.
Approach: They propose a dataset for long-form narrative summarization that uses human written summaries on three levels of difficulty.
Outcome: The proposed dataset covers documents from the literature domain, such as novels, plays and stories, and includes highly abstractive, human written summaries on three levels of difficulty.
WikiSum: Coherent Summarization Dataset for Efficient Human-Evaluation (2021.acl-short)

Copied to clipboard

Challenge: Existing summarization datasets are limited in their ability to evaluate output . a human evaluation is necessary to understand and improve summarizing systems .
Approach: They propose a dataset based on how-to articles and coherent paragraph summaries written in plain language.
Outcome: The proposed dataset makes human evaluation easier and more effective . the authors compare the proposed dataset to existing ones on PubMed and the literature.
ACLSum: A New Dataset for Aspect-based Summarization of Scientific Publications (2024.naacl-long)

Copied to clipboard

Challenge: Existing statistical phrasal or hierarchical machine translation systems relies on a large set of translation rules which results in engineering challenges.
Approach: They propose to use factorized grammar from the field of linguistics as more general translation rules from XTAG English Grammar to generate a manually crafted summarization dataset.
Outcome: The proposed method outperforms existing methods on low-resource language translation tasks with less training data.
Summary Explorer: Visualizing the State of the Art in Text Summarization (2021.emnlp-demo)

Copied to clipboard

Challenge: Automatic text summarization is the task of generating a summary of a long text by condensing it to its most important parts.
Approach: They propose a tool to visually explore document summarization systems based on three well-known summary quality criteria .
Outcome: The proposed tool compiles outputs of 55 state-of-the-art document summarization approaches and visually explores them during a qualitative assessment.
ForumSum: A Multi-Speaker Conversation Summarization Dataset (2021.findings-emnlp)

Copied to clipboard

Challenge: Abstractive summarization quality has been improved but there is a lack of data for conversation summarizing applications.
Approach: They propose to build a conversation summarization dataset with human written summaries from internet forums.
Outcome: The proposed dataset can be easily expanded to improve conversation summarization applications.
From Information to Insight: Leveraging LLMs for Open Aspect-Based Educational Summarization (2025.acl-long)

Copied to clipboard

Challenge: a novel dataset summarizes student reflections on STEM lectures . ReflectASP eases the exploration of open-aspect-based summarization (OABS) despite the limitations of current datasets, it is still under-explored.
Approach: They propose a dataset that summarizes student reflections on STEM lectures . they propose two refinement methods to improve summaries .
Outcome: The proposed dataset summarizes student reflections on STEM lectures using automatic and human evaluations.
CaseSumm: A Large-Scale Dataset for Long-Context Summarization from U.S. Supreme Court Opinions (2025.findings-naacl)

Copied to clipboard

Challenge: CaseSumm is a dataset for long-context summarization in the legal domain . human groundtruth summaries are often not available for legal summarizing .
Approach: They propose a dataset for long-context summarization that includes SCOTUS opinions and their official summaries.
Outcome: The proposed dataset is the largest open legal case summarization dataset . it outperforms larger models on automatic metrics and human evaluation .
BillSum: A Corpus for Automatic Summarization of US Legislation (D19-54)

Copied to clipboard

Challenge: In the US Congress, over 10,000 bills are introduced each year, with state legislatures introducing tens of thousands of bills.
Approach: They introduce the first dataset for summarizing US Congressional and California state bills . they demonstrate that models built on Congressional bills can be used to summarize California billa .
Outcome: The proposed summarization methods can be applied to states without human-written summaries.
Automatic Pyramid Evaluation Exploiting EDU-based Extractive Reference Summaries (D18-1)

Copied to clipboard

Challenge: Existing methods for evaluating content are not accurate because they only confirm if the summary contains small textual fragments.
Approach: They propose to transform human-made reference summaries into extractive reference sums and weight them using elementary discourse units.
Outcome: The proposed method strongly correlates with manual evaluations on DUC and TAC data sets.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations