Xavier Alameda-Pineda
University of Trento, Institut national de recherche en sciences et technologies du numérique, Centre Inria de l'Université Grenoble Alpes, Université Grenoble Alpes, Institut polytechnique de Grenoble, Directorate-General for Interpretation, Centre National de la Recherche Scientifique
Papers
18
Total Citations
362
H-Index
10
About
Xavier Alameda-Pineda is a researcher specializing in multimodal perception, human-robot interaction, and probabilistic machine learning, with a particular focus on audio-visual fusion for understanding human behavior in complex environments. His work has made significant contributions to the challenge of tracking, detecting, and localizing people using synchronized auditory and visual data — a problem central to building socially aware robotic systems. Among his most influential contributions are variational Bayesian frameworks for multi-person tracking in cluttered scenes, which together have garnered over 100 citations and represent a rigorous probabilistic approach to handling occlusions, appearance changes, and variable numbers of subjects. His early work on active-speaker detection using stereoscopic cameras and microphone arrays embedded in a robotic head (39 citations) demonstrated the practical power of sensory fusion in real-world robotics. The RAVEL corpus and his sound classification benchmarks laid important groundwork for training and evaluating domestic robots in realistic acoustic conditions. Across his career, Alameda-Pineda has consistently championed the complementarity of audio and visual modalities, showing that integrating both streams yields systems far more robust than either alone. With over 300 cumulative citations, his research has shaped modern approaches to perceptual intelligence in humanoid and companion robots.
Research Focus
Key Achievements
Top Papers
- 1Tracking Multiple Persons Based on a Variational Bayesian Model68 citations · 2016
- 2
- 3
- 4RAVEL: an annotated corpus for training robots with audiovisual abilities35 citations · 2012
- 5Sound representation and classification benchmark for domestic robots29 citations · 2014
- 6Sound-event recognition with a companion humanoid28 citations · 2012
- 7
- 8Finding audio-visual events in informal social gatherings22 citations · 2011
- 9Tracking a varying number of people with a visually-controlled robotic head18 citations · 2017
- 10Online multimodal speaker detection for humanoid robots16 citations · 2012