首页 /研究 /Barge-in-able robot audition based on ICA and missing feature theory under semi-blind situation
OTHER

Barge-in-able robot audition based on ICA and missing feature theory under semi-blind situation

Ryu Takeda, Kazuhiro Nakadai, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno

发表年份
2008
引用次数
14

摘要

This paper describes a robot audition system that allows the user to barge-in; that is, the user can speak simultaneously when the robot is speaking. Our ldquobarge-in-ablerdquo system consists of two stages: (1) cancellation of robot speech and (2) recognition of the separated user speech under the ldquosemi-blind situationrdquo. The semi-blind situation is where a robotpsilas speech signal is known but a userpsilas speech signal is not. The first stage is achieved by using an adaptive filter based on time-frequency domain Independent Component Analysis, because that can separate robot speech more robustly against noise than conventional echo cancellers. To improve performance in online processing, we utilized known source normalization and the exponentially weighted stepsize method. The second stage is achieved by automatic speech recognition (ASR) based on the missing feature theory which provides robust recognition by exploiting the reliability of speech features distorted due to noise and/or separation. The semi-blind situation simplifies the estimation of such reliabilities. Experiments demonstrated that our system improved word correctness of ASR by 10.0%.

关键词

Computer scienceSpeech recognitionNormalization (sociology)CorrectnessRobotArtificial intelligenceIndependent component analysisFeature (linguistics)Speech processingPattern recognition (psychology)

相关论文

查看 OTHER 分类全部论文