发明名称 Modelling and processing filled pauses and noises in speech recognition
摘要 A speech recognition system recognizes filled pause utterances made by a speaker. In one embodiment, an ergodic model is used to acoustically model filled pauses that provides flexibility allowing varying utterances of the filled pauses to be made. The ergodic HMM model can also be used for other types of noise such as but limited to breathing, keyboard operation, microphone noise, laughter, door openings and/or closings, or any other noise occurring in the environment of the user or made by the user. Similarly, silence can be modeled using an ergodic HMM model. Recognition can be used with N-gram, context-free grammar or hybrid language models.
申请公布号 US7076422(B2) 申请公布日期 2006.07.11
申请号 US20030388259 申请日期 2003.03.13
申请人 MICROSOFT CORPORATION 发明人 HWANG MEI-YUH
分类号 G10L15/20;G10L15/14;G10L21/02 主分类号 G10L15/20
代理机构 代理人
主权项
地址