Papers by Kazunori Komatani
Collection of Multimodal Dialog Data and Analysis of the Result of Annotation of Users’ Interest Level (L18-1)
Copied to clipboard
Masahiro Araki, Sayaka Tomimasu, Mikio Nakano, Kazunori Komatani, Shogo Okada, Shinya Fujie, Hiroaki Sugiyama
| Challenge: | a group of researchers is building a corpus for evaluating elements of multimodal dialogue systems. |
| Approach: | They propose to build a corpus for evaluating elements of the multimodal dialogue system . they use the Wizard of Oz method to record chat dialogue data between a human and a virtual agent . |
| Outcome: | The proposed method annotates chat dialogue data between a human and a virtual agent and measures their interest level in the data. |
Collection and Analysis of Travel Agency Task Dialogues with Age-Diverse Speakers (2022.lrec-1)
Copied to clipboard
| Challenge: | Using deep neural networks, task-oriented dialogue systems can be used to generate an appropriate response to users' inputs. |
| Approach: | They collected a multimodal dialogue corpus with a wide range of speaker ages and set up a dialogue task based on travel . results suggest adult speakers have more independent opinions, older speakers express opinions more frequently compared with other age groups, and operators expressed a smile more frequently to minor speakers. |
| Outcome: | The results show that adult speakers have more independent opinions, the older speakers express their opinions more frequently compared with other age groups, and the operators expressed a smile more frequently to the minor speakers. |
Collecting Human-Agent Dialogue Dataset with Frontal Brain Signal toward Capturing Unexpressed Sentiment (2024.lrec-main)
Copied to clipboard
| Challenge: | Multimodal information such as text and audiovisual data has been used for emotion/sentiment estimation during human-agent dialogues. |
| Approach: | They present a method for dealing with eye-blink noise for frontal EEGs denoising. |
| Outcome: | The proposed method improves sentiment estimation performance when used with other modalities by multimodal fusion, although it only has three channels. |