Online Multimodal Speaker Detection for Humanoid Robots

In this paper we address the problem of audio-visual speaker detection. We introduce an online system working on the humanoid robot NAO. The scene is perceived with two cameras and two microphones. A multimodal Gaussian mixture model (mGMM) fuses the information extracted from the auditory and visual sensors and detects the most probable audio-visual object, e.g., a person emitting a sound, in the 3D space. The system is implemented on top of a platform-independent middleware and it is able to process the information online (17Hz). A detailed description of the system and its implementation are provided, with special emphasis on the online processing issues and the proposed solutions. Experimental validation, performed with five different scenarios, show that that the proposed method opens the door to robust human-robot interaction scenarios.

Domains

Computer Vision and Pattern Recognition [cs.CV] Signal and Image processing Signal and Image Processing

Fichier principal

Sanchez-Humanoids2012.pdf (755.72 Ko)

Origin : Files produced by the author(s)

Perception team : Connect in order to contact the contributor

https://inria.hal.science/hal-00768764

Submitted on : Sunday, December 23, 2012-7:44:32 PM

Last modification on : Thursday, April 4, 2024-9:19:05 PM

Long-term archiving on: Sunday, March 24, 2013-3:51:06 AM

Dates and versions

hal-00768764 , version 1 (23-12-2012)

Identifiers

HAL Id : hal-00768764 , version 1
DOI : 10.1109/HUMANOIDS.2012.6651509

Cite

Jordi Sanchez-Riera, Xavier Alameda-Pineda, Johannes Wienke, Antoine Deleforge, Soraya Arias, et al.. Online Multimodal Speaker Detection for Humanoid Robots. Humanoids 2012 - IEEE International Conference on Humanoid Robotics, Nov 2012, Osaka, Japan. pp.126-133, ⟨10.1109/HUMANOIDS.2012.6651509⟩. ⟨hal-00768764⟩

Export

BibTeX XML-TEI Dublin Core DC Terms EndNote DataCite

Collections

UNIV-RENNES1 UGA CNRS INRIA IRISA LJK LJK_GI LJK_GI_PERCEPTION INRIA2 UR1-MATH-STIC UR1-UFR-ISTIC UNIV-RENNES UR1-MATH-NUM

518 View

363 Download