Annotation and Quantitative Analysis of Speaker Information in Novel Conversation Sentences in Japanese (L18-1)
Copied to clipboard
| Challenge: | a qualitative lexicological analysis of conversation sentences in novels is performed . gender and age of conversation sentence information is not used as actual speech . |
| Approach: | They performed a quantitative lexicological analysis using attributed speaker information . they also examined the differences between Japanese novels and translations of foreign novels . |
| Outcome: | The results show that conversation sentences in novels are representative of spoken language . the authors conclude that conversation sentence data are not useful as speech materials . |
Similar Papers
Identifying Speakers and Addressees in Dialogues Extracted from Literary Fiction (L18-1)
Copied to clipboard
Adam Ek, Mats Wirén, Robert Östling, Kristina N. Björkenstam, Gintarė Grigonytė, Sofia Gustafson Capková
| Challenge: | Using a sequence labeling approach, it is possible to identify speakers and addressees in dialogues extracted from literary fiction using a small amount of training data. |
| Approach: | They propose to use a sequence labeling approach applied to a given set of characters to identify speakers and addressees in dialogues extracted from literary fiction. |
| Outcome: | The proposed method allows for enriched search facilities and construction of social networks from the corpora. |
An Annotated Dataset of Discourse Modes in Hindi Stories (2020.lrec-1)
Copied to clipboard
Swapnil Dhanwal, Hritwik Dutta, Hitesh Nankani, Nilay Shrivastava, Yaman Kumar, Junyi Jessy Li, Debanjan Mahata, Rakesh Gosangi, Haimin Zhang, Rajiv Ratn Shah, Amanda Stent
| Challenge: | Using a new corpus of sentences from Hindi short stories, we analyze the annotations for five different discourse modes argumentative, narrative, descriptive, dialogic and informative. |
| Approach: | They propose to annotate sentences from Hindi short stories for five different discourse modes argumentative, narrative, descriptive, dialogic and informative. |
| Outcome: | The proposed corpus has a high inter-annotator agreement (0.87 k-alpha) and is able to capture the nuances of the embedded discourse structures. |
Annotation and Analysis of Extractive Summaries for the Kyutech Corpus (L18-1)
Copied to clipboard
| Challenge: | Summarization of multi-party conversation requires corpora to analyze characteristics of conversations and construct a method for summary generation. |
| Approach: | They propose to annotate a Japanese conversation corpus for a decision-making task . they compare extractive summarization methods with the annotated extractive summary . |
| Outcome: | The proposed corpus is the first annotated for conversation summarization tasks and freely available to anyone. |
A Document-Level Text Simplification Dataset for Japanese (2024.lrec-main)
Copied to clipboard
| Challenge: | Document-level text simplification tasks combine summarization and intra-sentence simplification. |
| Approach: | They devised a Japanese document-level text simplification dataset based on newspaper articles and Wikipedia. |
| Outcome: | The proposed dataset compared Japanese document-level text simplification models with English models and newspaper articles. |
Creating a Data Set of Abstractive Summaries of Turn-labeled Spoken Human-Computer Conversations (2022.lrec-1)
Copied to clipboard
| Challenge: | Digital recorded written and spoken dialogues are becoming more available due to the growing popularity of online messenger services and chatbots. |
| Approach: | They propose to use Dutch spoken human-computer conversations, an annotation layer of turn labels, and conversational abstractive summaries of user answers to build a conversational agent. |
| Outcome: | The proposed system can be integrated into a conversational agent. |
A Conversation-Analytic Annotation of Turn-Taking Behavior in Japanese Multi-Party Conversation and its Preliminary Analysis (2020.lrec-1)
Copied to clipboard
| Challenge: | a new conversation-analytic annotation scheme is proposed for multi-party conversations . current systems do not take a turn like a human even in simple two-party conversation . |
| Approach: | They propose a conversation-analytic annotation scheme for turn-taking behavior in multi-party conversations . they analyze how syntactic and prosodic features of utterances vary across four selection types . |
| Outcome: | The proposed model is based on Japanese multi-party conversations. |
Construction of the Corpus of Everyday Japanese Conversation: An Interim Report (L18-1)
Copied to clipboard
Hanae Koiso, Yasuharu Den, Yuriko Iseki, Wakako Kashino, Yoshiko Kawabata, Ken’ya Nishikawa, Yayoi Tanaka, Yasuyuki Usuda
| Challenge: | a new corpus of everyday conversations is being developed in the field of everyday conversation . the corpus is based on 94 hours of recordings of everyday Japanese conversations . |
| Approach: | They propose to build a large-scale corpus of everyday Japanese conversation in a balanced manner. |
| Outcome: | The proposed corpus will be published in 2022 and consist of more than 200 hours of recordings. |
A Japanese Corpus for Analyzing Customer Loyalty Information (L18-1)
Copied to clipboard
| Challenge: | a corpus of customer loyalty information is used to analyze customer's voice . a variety of studies have focused on analyzing attitudes, opinions, sentiments of text data . |
| Approach: | They present a corpus for analyzing customer loyalty information . they use voice of customer to capture customer's behaviors, needs and feedbacks . |
| Outcome: | The proposed corpus analyzes customer loyalty by analyzing their voice . the study is based on a corpus of customer loyalty information . |
CLAUSE-ATLAS: A Corpus of Narrative Information to Scale up Computational Literary Analysis (2024.lrec-main)
Copied to clipboard
| Challenge: | XIX and XX century English novels annotated automatically contain 41,715 labeled clauses . a new approach to analyze novels based on clauses captures structural patterns within books, as well as qualitative differences between them. |
| Approach: | They propose to use a corpus of XIX and XX century English novels annotated automatically to study stories as sequences of eventive, subjective and contextual information. |
| Outcome: | The proposed method captures structural patterns within books, as well as qualitative differences between them. |
Japanese Dialogue Corpus of Information Navigation and Attentive Listening Annotated with Extended ISO-24617-2 Dialogue Act Tags (L18-1)
Copied to clipboard
| Challenge: | Large-scale conventional dialogue corpora are mainly built for specified tasks with specially designed dialogue states. |
| Approach: | They propose to annotate large-scale dialogue data with an extended ISO-24617-2 dialogue act tag-set to model a natural conversation with machines. |
| Outcome: | The proposed corpus covers a wider range of dialogue tasks than existing task-oriented systems or text-chat systems. |