Patent · US Expired

Method of real-time speaker change point detection, speaker tracking and speaker model construction

US7181393B2 · kind B2 · utility

6Cited by

2References

28Claims

0Family size

Assignee

Microsoft Corporation · US

Inventors

Lie Lu · Beijing, CN
Hong-Jiang Zhang · Beijing, CN

Key dates

Filing date	Nov 29, 2002
Grant date	Feb 20, 2007
Priority date	—
Expiry date	Apr 30, 2025

Classification

Technology area (CPC G)Physics
CPC primaryG10L17/20
WIPO fieldComputer technology
WIPO sectorElectrical engineering

Abstract

A method is provided for real-time speaker change detection and speaker tracking in a speech signal. The method is a “coarse-to-refine” process, which consists of two stages: pre-segmentation and refinement. In the pre-segmentation process, the covariance of a feature vector of each segment of speech is built initially. A distance is determined based on the covariance of the current segment and a previous segment; and the distance is used to determine if there is a potential speaker change between these two segments. If there is no speaker change, the model of current identified speaker model is updated by incorporating data of the current segment. Otherwise, if there is a speaker change, a refinement process is utilized to confirm the potential speaker change point.

Source: USPTO / EPO open patent data. Objective bibliographic and citation counts.