Patent · US Expired

Modelling and processing filled pauses and noises in speech recognition

US7076422B2 · kind B2 · utility

5Cited by

7References

22Claims

0Family size

Assignee

Microsoft Corporation · US

Inventor

Mei-Yuh Hwang · Sammamish, US

Key dates

Filing date	Mar 13, 2003
Grant date	Jul 11, 2006
Priority date	—
Expiry date	Mar 11, 2024

Classification

Technology area (CPC G)Physics
CPC primaryG10L2021/02168
WIPO fieldComputer technology
WIPO sectorElectrical engineering

Abstract

A speech recognition system recognizes filled pause utterances made by a speaker. In one embodiment, an ergodic model is used to acoustically model filled pauses that provides flexibility allowing varying utterances of the filled pauses to be made. The ergodic HMM model can also be used for other types of noise such as but limited to breathing, keyboard operation, microphone noise, laughter, door openings and/or closings, or any other noise occurring in the environment of the user or made by the user. Similarly, silence can be modeled using an ergodic HMM model. Recognition can be used with N-gram, context-free grammar or hybrid language models.

Source: USPTO / EPO open patent data. Objective bibliographic and citation counts.