Method of real-time speaker change point detection, speaker tracking and speaker model construction
US7181393B2 · kind B2 · utility
Assignee
Inventors
Key dates
| Filing date | Nov 29, 2002 |
| Grant date | Feb 20, 2007 |
| Priority date | — |
| Expiry date | Apr 30, 2025 |
Classification
- Technology area (CPC G)Physics
- CPC primaryG10L17/20
- WIPO fieldComputer technology
- WIPO sectorElectrical engineering
Abstract
A method is provided for real-time speaker change detection and speaker tracking in a speech signal. The method is a “coarse-to-refine” process, which consists of two stages: pre-segmentation and refinement. In the pre-segmentation process, the covariance of a feature vector of each segment of speech is built initially. A distance is determined based on the covariance of the current segment and a previous segment; and the distance is used to determine if there is a potential speaker change between these two segments. If there is no speaker change, the model of current identified speaker model is updated by incorporating data of the current segment. Otherwise, if there is a speaker change, a refinement process is utilized to confirm the potential speaker change point.
Source: USPTO / EPO open patent data. Objective bibliographic and citation counts.