Multi-modal integration for personalized conversation: Towards a humanoid in daily life

Shinya Fujie, Daichi Watanabe, Yuhi Ichikawa, Hikaru Taniyama, Kosuke Hosoya, Yoichi Matsuyama, Tetsunori Kobayashi

研究成果: Conference contribution

3 引用 (Scopus)

抜粋

Humanoid with spoken language communication ability is proposed and developed. To make humanoid live with people, spoken language communication is fundamental because we use this kind of communication everyday. However, due to difficulties of speech recognition itself and implementation on the robot, a robot with such an ability has not been developed. In this study, we propose a robot with the technique implemented to overcome these problems. This proposed system includes three key features, image processing, sound source separation, and turn-taking timing control. Processing image captured with camera mounted on the robot's eyes enables to find and identify whom the robot should talked to. Sound source separation enables distant speech recognition, so that people need no special device, such as head-set microphones. Turn-taking timing control is often lacked in many conventional spoken dialogue system, but this is fundamental because the conversation proceeds in realtime. The effectiveness of these elements as well as the example of conversation are shown in experiments.

元の言語English
ホスト出版物のタイトル2008 8th IEEE-RAS International Conference on Humanoid Robots, Humanoids 2008
ページ617-622
ページ数6
DOI
出版物ステータスPublished - 2008 12 1
イベント2008 8th IEEE-RAS International Conference on Humanoid Robots, Humanoids 2008 - Daejeon, Korea, Republic of
継続期間: 2008 12 12008 12 3

出版物シリーズ

名前2008 8th IEEE-RAS International Conference on Humanoid Robots, Humanoids 2008

Conference

Conference2008 8th IEEE-RAS International Conference on Humanoid Robots, Humanoids 2008
Korea, Republic of
Daejeon
期間08/12/108/12/3

ASJC Scopus subject areas

  • Artificial Intelligence
  • Computer Vision and Pattern Recognition
  • Human-Computer Interaction

フィンガープリント Multi-modal integration for personalized conversation: Towards a humanoid in daily life' の研究トピックを掘り下げます。これらはともに一意のフィンガープリントを構成します。

  • これを引用

    Fujie, S., Watanabe, D., Ichikawa, Y., Taniyama, H., Hosoya, K., Matsuyama, Y., & Kobayashi, T. (2008). Multi-modal integration for personalized conversation: Towards a humanoid in daily life. : 2008 8th IEEE-RAS International Conference on Humanoid Robots, Humanoids 2008 (pp. 617-622). [4756014] (2008 8th IEEE-RAS International Conference on Humanoid Robots, Humanoids 2008). https://doi.org/10.1109/ICHR.2008.4756014