Patent · US Active

Audio event detection with window-based prediction

US12272377B2 · kind B2 · utility

0Cited by
14References
20Claims
0Family size

Assignee

Inventors

Key dates

Filing dateMar 5, 2024
Grant dateApr 8, 2025
Priority date
Expiry dateMar 5, 2044

Classification

  • Technology area (CPC G)Physics
  • CPC primaryG10L25/57
  • WIPO fieldComputer technology
  • WIPO sectorElectrical engineering

Abstract

A computing system for a plurality of classes of audio events is provided, including one or more processors configured to divide a run-time audio signal into a plurality of segments and process each segment of the run-time audio signal in a time domain to generate a normalized time domain representation of each segment. The processor is further configured to feed the normalized time domain representation of each segment to an input layer of a trained neural network. The processor is further configured to generate, by the neural network, a plurality of predicted classification scores and associated probabilities for each class of audio event contained in each segment of the run-time input audio signal. In post-processing, the processor is further configured to generate smoothed predicted classification scores, associated smoothed probabilities, and class window confidence values for each class for each of a plurality of candidate window sizes.

Source: USPTO / EPO open patent data. Objective bibliographic and citation counts.