Perceiving user's intention-for-interaction: A probabilistic multimodal data fusion scheme

2015 IEEE International Conference on Multimedia and Expo (ICME) Pub Date : 2015-06-29 DOI:10.1109/ICME.2015.7177514

C. Mollaret, Alhayat Ali Mekonnen, I. Ferrané, J. Pinquier, F. Lerasle

引用次数: 13

Abstract

Understanding people's intention, be it action or thought, plays a fundamental role in establishing coherent communication amongst people, especially in non-proactive robotics, where the robot has to understand explicitly when to start an interaction in a natural way. In this work, a novel approach is presented to detect people's intention-for-interaction. The proposed detector fuses multimodal cues, including estimated head pose, shoulder orientation and vocal activity detection, using a probabilistic discrete state Hidden Markov Model. The multimodal detector achieves up to 80% correct detection rates improving purely audio and RGB-D based variants.

查看原文本刊更多论文

感知用户交互意图:一种概率多模态数据融合方案

理解人们的意图，无论是行动还是思想，在人与人之间建立连贯的沟通中起着至关重要的作用，尤其是在非主动机器人中，机器人必须明确地理解何时以自然的方式开始互动。在这项工作中，提出了一种新的方法来检测人们的互动意图。该检测器使用概率离散状态隐马尔可夫模型融合多模态线索，包括估计的头部姿势、肩部方向和声音活动检测。多模态检测器实现了高达80%的正确检测率，改善了纯音频和RGB-D变体。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

求助全文

约1分钟内获得全文求助全文

来源期刊

2015 IEEE International Conference on Multimedia and Expo (ICME)

自引率

0.00%

发文量