Annotation and Analysis of Extractive Summaries for the Kyutech Corpus (L18-1)

Copied to clipboard

Challenge: Summarization of multi-party conversation requires corpora to analyze characteristics of conversations and construct a method for summary generation.
Approach: They propose to annotate a Japanese conversation corpus for a decision-making task . they compare extractive summarization methods with the annotated extractive summary .
Outcome: The proposed corpus is the first annotated for conversation summarization tasks and freely available to anyone.

Similar Papers

Extractive Summarisation for German-language Data: A Text-level Approach with Discourse Features (2022.coling-1)

Copied to clipboard

Challenge: Using RST, extractive summarisation involves using select phrases and sentences as a summary, which still remains a strong method for producing summaries despite its simple nature.
Approach: They propose to use RST-based features to analyse the connection between summary sentences and several RST features and transfer these insights to various automated summarisation models.
Outcome: The proposed models are based on the best features proposed over the last 20+ years and incorporate the best ones into the proposed models.
Relational Summarization for Corpus Analysis (N18-1)

Copied to clipboard

Challenge: Existing methods for summarizing textual content are often ignored . relationshipal questions are ubiquitous and varied.
Approach: They propose a method which generates a natural language summary of the relationship between two lexical items in a corpus without reference to a knowledge base.
Outcome: The proposed method generates a natural language summary of the relationship between two lexical items in a corpus without reference to a knowledge base.
SAMSum Corpus: A Human-annotated Dialogue Dataset for Abstractive Summarization (D19-54)

Copied to clipboard

Challenge: Existing work on abstractive dialogue summarizations has focused on news summarizing but there is no such comprehensive dataset.
Approach: They propose to use a chat-dialogues corpus with abstractive dialogue summaries to generate a short version of text that covers the main points succinctly.
Outcome: The proposed dataset achieves higher ROUGE scores than the model-generated summaries of news, compared with human evaluators' judgement.
Situation-Based Multiparticipant Chat Summarization: a Concept, an Exploration-Annotation Tool and an Example Collection (2021.acl-srw)

Copied to clipboard

Challenge: Currently, text chat does not offer navigation or full-featured search, although the high volumes of messages demand it.
Approach: They propose a data annotation tool for situation-based summarization that can be used to extract messages from chat logs.
Outcome: The proposed tool is the first to be developed for situation-based summarization.
Abstractive Meeting Summarization: A Survey (2023.tacl-1)

Copied to clipboard

Challenge: Recent advances in deep learning have improved language generation systems, opening the door to improved forms of abstractive summarization.
Approach: They propose to use neural encoder-decoder architectures to generate abstractive meeting summarizations that are particularly well-suited for multi-party conversation.
Outcome: The proposed system could be used in a wide variety of real-world contexts, from business meetings to medical consultations to customer service calls.
Beyond Generic Summarization: A Multi-faceted Hierarchical Summarization Corpus of Large Heterogeneous Data (L18-1)

Copied to clipboard

Challenge: Automated summarization has focused on ten to twenty documents, typically news articles, but could in theory analyze hundreds of documents from a wide range of sources and provide an overview to the interested reader.
Approach: They propose a method for creating hierarchical summarization corpora from large, heterogeneous document collections by crowdsourcing relevant content and asking trained annotators to order the relevant information hierarchically.
Outcome: The proposed method can be used to develop and evaluate hierarchical summarization systems.
Facet-Aware Evaluation for Extractive Summarization (2020.acl-main)

Copied to clipboard

Challenge: lexical overlap is a common evaluation metric for extractive summarization, but recent studies reveal its limitations.
Approach: They propose a facet-aware evaluation setup for better assessment of information coverage in extractive summaries.
Outcome: The proposed evaluation setup improves human correlation with extractive summarization datasets and improves comparative analysis.
Extractive Summarization via ChatGPT for Faithful Summary Generation (2023.findings-emnlp)

Copied to clipboard

Challenge: Abstractive summarization methods struggle with generating ungrammatical or even nonfactual contents.
Approach: They evaluate ChatGPT's performance on extractive summarization and compare it with traditional fine-tuning methods on benchmark datasets.
Outcome: The proposed pipeline performs better than abstractive methods on summary faithfulness and in-context learning.
Improving Crowdsourcing-Based Annotation of Japanese Discourse Relations (L18-1)

Copied to clipboard

Challenge: Discourse parsing is an important task in natural language processing, but few languages have corpora annotated with discourse relations . crowdsourcing-based annotations are of poor quality and require expensive and time-consuming . et al. (2009) evaluated the quality of annotations using expert annotations.
Approach: They construct a Japanese corpus with discourse annotations through crowdsourcing . they propose improvement techniques based on language tests .
Outcome: The proposed methods improve the quality of the annotations, and will make them publicly available.
Frame Semantic-Enhanced Sentence Modeling for Sentence-level Extractive Text Summarization (2021.emnlp-main)

Copied to clipboard

Challenge: Sentence-level extractive text summarization is difficult to model the importance of sentences.
Approach: They propose a Frame Semantic-Enhanced Sentence Modeling for Extractive Summarization that leverages Frame semantics to model sentences from both intra-sentence level and inter-sentent level.
Outcome: The proposed model outperforms six state-of-the-art methods on two benchmark corpus datasets.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations