Patent · US Active

Training acoustic models using connectionist temporal classification

US10803855B1 · kind B1 · utility

4Cited by

61References

26Claims

0Family size

Assignee

Google LLC · US

Inventors

Kanury Kanishka Rao · Santa Clara, US
Andrew W. Senior · New York, US
Hasim Sak · New York, US

Key dates

Filing date	Jan 25, 2019
Grant date	Oct 13, 2020
Priority date	—
Expiry date	Mar 29, 2039

Classification

Technology area (CPC G)Physics
CPC primaryG10L2015/022
WIPO fieldComputer technology
WIPO sectorElectrical engineering

Abstract

Methods, systems, and apparatus, including computer programs encoded on a computer storage medium, for training acoustic models and using the trained acoustic models. A connectionist temporal classification (CTC) acoustic model is accessed, the CTC acoustic model having been trained using a context-dependent state inventory generated from approximate phonetic alignments determined by another CTC acoustic model trained without fixed alignment targets. Audio data for a portion of an utterance is received. Input data corresponding to the received audio data is provided to the accessed CTC acoustic model. Data indicating a transcription for the utterance is generated based on output that the accessed CTC acoustic model produced in response to the input data. The data indicating the transcription is provided as output of an automated speech recognition service.

Source: USPTO / EPO open patent data. Objective bibliographic and citation counts.