Real-Time multiple speaker tracking by multi-modal integration for mobile robots

Kazuhiro Nakadai, Ken Ichi Hidai, Hiroshi G. Okuno, Hiroaki Kitano

研究成果: Conference contribution

15 引用 (Scopus)

抜粋

In this paper, real-Time multiple speaker tracking is addressed, because it is essential in robot perception and humanrobot social interaction. The difficulty lies in treating a mixture of sounds, occlusion (some talkers are hidden) and real-Time processing. Our approach consists of three components; (1) the extraction of the direction of each speaker by using interaural phase difference and interaural intensity difference, (2) the resolution of each speaker's direction by multi-modal integration of audition, vision and motion with canceling inevitable motor noises in motion in case of an unseen or silent speaker, and (3) the distributed implementation to three PCs connected by TCP/IP network to attain real-Time processing. As a result, we attain robust real-Time speaker tracking with 200 ms delay in a non-Anechoic room, even when multiple speakers exist and the tracking person is visually occluded.

元の言語English
ホスト出版物のタイトルEUROSPEECH 2001 - SCANDINAVIA - 7th European Conference on Speech Communication and Technology
出版者International Speech Communication Association
ページ1193-1196
ページ数4
ISBN(電子版)8790834100, 9788790834104
出版物ステータスPublished - 2001
外部発表Yes
イベント7th European Conference on Speech Communication and Technology - Scandinavia, EUROSPEECH 2001 - Aalborg, Denmark
継続期間: 2001 9 32001 9 7

Other

Other7th European Conference on Speech Communication and Technology - Scandinavia, EUROSPEECH 2001
Denmark
Aalborg
期間01/9/301/9/7

    フィンガープリント

ASJC Scopus subject areas

  • Communication
  • Linguistics and Language
  • Computer Science Applications
  • Software

これを引用

Nakadai, K., Hidai, K. I., Okuno, H. G., & Kitano, H. (2001). Real-Time multiple speaker tracking by multi-modal integration for mobile robots. : EUROSPEECH 2001 - SCANDINAVIA - 7th European Conference on Speech Communication and Technology (pp. 1193-1196). International Speech Communication Association.