Human-robot conversation with multiple participants in noisy public spaces

📅 2026-08-31
📈 Citations: 0
Influential: 0
📄 PDF
🤖 AI Summary
研究提出了一种音频系统,使用单个多通道麦克风阵列增强嘈杂环境中的多人对话信号,提高语音识别和远程操作的沉浸感。
📝 Abstract
For noisy real-world environments such as those in open public spaces, spoken dialogue systems for both autonomous robots and avatars should be carefully designed to provide enhanced speech signals. These signals can be used either for speech recognition or, in the case of an avatar system, transmitted as clean speech to a remote operator. This work proposes an audio system that can be used for both these scenarios and was demonstrated as a proof-of-concept at the 2025 World Expo in Osaka. The first scenario is an attentive listening system with the android ERICA, and the second is a conversation support system with mobile Teleco robots, with one of them acting as an avatar for a remote operator. Both systems feature multi-party conversation and use a single multi-channel microphone array. We describe how our audio system not only enhances the speech of multiple speakers in a noisy environment, but provides a form of spatial audio which allows for more immersiveness in avatar-based conversational interactions.
Problem

Research questions and friction points this paper is trying to address.

noisy public spaces
speech signals
multi-party conversation
Innovation

Methods, ideas, or system contributions that make the work stand out.

audio system
multi-party conversation
noisy environment
spatial audio
💼 Related Jobs
No related jobs found.
Divesh Lala
Divesh Lala
Kyoto University
Artificial intelligencehuman-computer interactionvirtual agentsandroidsdialogue systems
Y
Yogeeswaran Muthukumaran
National University of Singapore
V
Vincent Fernandes
Osaka University, Graduate School of Engineering Science; ENSEA
K
Kazushi Kato
Kyoto University Graduate School of Informatics
S
Shota Fujiki
Osaka University, Graduate School of Engineering Science
Z
Zihao Chi
Osaka University, Graduate School of Engineering Science
M
Masaya Iwasaki
Osaka University, Graduate School of Engineering Science
T
Taiken Shintani
Osaka University, Graduate School of Engineering Science
M
Megumi Kawata
Osaka University, Graduate School of Engineering Science
K
Kazuki Sakai
Osaka University, Graduate School of Engineering Science
Koji Inoue
Koji Inoue
Kyoto University
Spoken Dialogue SystemHuman-Robot InteractionTurn-Taking
Y
Yuicihiro Yoshikawa
Osaka University, Graduate School of Engineering Science
Tatsuya Kawahara
Tatsuya Kawahara
Professor, School of Informatics, Kyoto University
Speech Processingspeech recognitionNatural Language Processingdialogue