Yuya Chiba
Other people with similar names: Yuya Chiba
Unverified author pages with similar names: Yuya Chiba
2026
From Felt Sense to Words: Construction and Analysis of a Focusing Dialogue Dataset for Verbalization
Yuiko Tsunomori | Yuya Chiba | Yosuke Koshikawa
Proceedings of the 27th Annual Meeting of the Special Interest Group on Discourse and Dialogue
Yuiko Tsunomori | Yuya Chiba | Yosuke Koshikawa
Proceedings of the 27th Annual Meeting of the Special Interest Group on Discourse and Dialogue
Verbalization, the process of expressing one’s internal states in words, has been shown to deepen self-understanding and improve well-being. However, this can be difficult to achieve for some individuals. One approach to supporting verbalization is Focusing-Oriented Psychotherapy. Focusing facilitates verbalization by guiding a speaker’s attention to their felt sense, an internal state that has not yet been verbalized, and helping them to find the words that appropriately express this felt sense through dialogue. Toward developing Focusing dialogue systems that support user verbalization, we constructed a dataset containing Focusing dialogues between professionally trained listeners and speakers. This dataset consists of 50 dialogues (762 minutes, 17,986 utterances) and contains ratings of verbalization progress, subjective evaluations, and dialogue act (DA) annotations. To clarify the characteristics of Focusing dialogues, we analyzed associations among verbalization progress, subjective evaluations, and DA usage. Based on these analyses, we discussed design implications for Focusing dialogue systems.
Sensor-Augmented Voice Activity Projection for Enhancing Turn-Taking Prediction
Satoki Hamanaka | Yasue Kishino | Yuiko Tsunomori | Shin Mizutani | Yuya Chiba | Tadashi Okoshi | Jin Nakazawa
Proceedings of the 27th Annual Meeting of the Special Interest Group on Discourse and Dialogue
Satoki Hamanaka | Yasue Kishino | Yuiko Tsunomori | Shin Mizutani | Yuya Chiba | Tadashi Okoshi | Jin Nakazawa
Proceedings of the 27th Annual Meeting of the Special Interest Group on Discourse and Dialogue
Voice Activity Projection (VAP) has been actively studied to enable natural turn-taking in spoken dialogue systems, relying primarily on acoustic features. Visual cues such as head movements are also known to contribute to turn-taking prediction; however, camera-based approaches are affected by placement and lighting conditions and are not always reliably available to dialogue systems. As a camera-independent approach for directly capturing head motion, earable devices offer a promising solution. In this study, we propose Sensor-Augmented VAP, a framework that integrates in-ear inertial measurement unit (IMU) signals with a pre-trained VAP model via a lightweight residual fusion module. To validate our proposed method, we collected a dataset pairing conversational audio with in-ear IMU data, comprising 12 dyadic Japanese dialogues recorded using microphones and earbuds. Experiments in speaker-independent and speaker-dependent settings demonstrate that IMU fusion consistently improves weighted F1 score for shift detection and reduces VAP loss over the audio-only baseline. These results confirm that head-motion cues are effective for enhancing turn-taking prediction.