Audio event detection with window-based prediction
US11948599B2 · kind B2 · utility
Assignee
Inventors
Key dates
| Filing date | Jan 6, 2022 |
| Grant date | Apr 2, 2024 |
| Priority date | — |
| Expiry date | Oct 9, 2042 |
Classification
- Technology area (CPC G)Physics
- CPC primaryG10L25/57
- WIPO fieldComputer technology
- WIPO sectorElectrical engineering
Abstract
A computing system for a plurality of classes of audio events is provided, including one or more processors configured to divide a run-time audio signal into a plurality of segments and process each segment of the run-time audio signal in a time domain to generate a normalized time domain representation of each segment. The processor is further configured to feed the normalized time domain representation of each segment to an input layer of a trained neural network. The processor is further configured to generate, by the neural network, a plurality of predicted classification scores and associated probabilities for each class of audio event contained in each segment of the run-time input audio signal. In post-processing, the processor is further configured to generate smoothed predicted classification scores, associated smoothed probabilities, and class window confidence values for each class for each of a plurality of candidate window sizes.
Source: USPTO / EPO open patent data. Objective bibliographic and citation counts.