Multi-modal and multi-camera attention in smart environments

Boris Schauerte, Jan Richarz, Thomas Plötz, Christian Thurau, Gernot A. Fink

发表年份: 2009
引用次数: 17

摘要

This paper considers the problem of multi-modal saliency and attention. Saliency is a cue that is often used for directing attention of a computer vision system, e.g., in smart environments or for robots. Unlike the majority of recent publications on visual/audio saliency, we aim at a well grounded integration of several modalities. The proposed framework is based on fuzzy aggregations and offers a flexible, plausible, and efficient way for combining multi-modal saliency information. Besides incorporating different modalities, we extend classical 2D saliency maps to multi-camera and multi-modal 3D saliency spaces. For experimental validation we realized the proposed system within a smart environment. The evaluation took place for a demanding setup under real-life conditions, including focus of attention selection for multiple subjects and concurrently active modalities.

关键词

ModalitiesComputer scienceModalFocus (optics)Artificial intelligenceComputer visionModality (human–computer interaction)Fuzzy logicHuman–computer interactionRobot

Multi-modal and multi-camera attention in smart environments

摘要

关键词

相关论文

Statistical Learning Theory

Artificial intelligence: a modern approach

Applied Nonlinear Control

A new optimizer using particle swarm theory