首页 /研究 /Automatic Speech Recognition for Human-Robot Interaction Using an Under-Resourced Language
HRI

Automatic Speech Recognition for Human-Robot Interaction Using an Under-Resourced Language

Juho Leinonen

发表年份
2015
引用次数
5
访问权限
开放获取

摘要

Automatic speech recognition will soon be a part of everyday life. Even today many people use the speech recognizer in their smartphones, whether it is Google Now or Siri. Commercial applications have existed for years for automatic dictation, and command-based voice user interfaces. The abundance of software divides languages in two; in well-resourced languages there is no shortage of products, while under-resourced languages might not even receive academic interest.
\n
\nIn this thesis, an automatic speech recognizer is built for North Sami, which is a morphologically rich under-resourced language in the Uralic family. These properties create challenges for the recognition process, of which this thesis will concentrate on the issue of out-of-vocabulary words. The use of whole words is compared with word fragments, morphs, and tests are conducted to optimize other language model variables such as vocabulary size and context length.
\n
\nThe experiments show that morph-based language models solve the problem of out-of-vocabulary words and significantly improve the recognition results without slowing the process too much. In addition, increasing context length improves the morph models, while adding supervision to generating them does not. As such, this thesis recommends a high order morph model generated with unsupervised methods to be used with North Sami.

关键词

Human–robot interactionComputer scienceSpeech recognitionNatural language processingRobotArtificial intelligenceHuman–computer interactionCommunicationPsychology

相关论文

查看 HRI 分类全部论文