IP Library Patent Application 16555251
Patent Application
App. No. 16/555,251

AUDIO DATA AUGMENTATION FOR MACHINE LEARNING OBJECT CLASSIFICATION

Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US None
App. No.
16/555,251
Abstract

This application discloses a computing system to receive audio data corresponding to sounds emitted by objects capable of being identified in an environment. The audio data includes labels identifying types of the objects emitting the sounds. The computing system alters the audio data to generate an augmented audio data set by modifying a timing of the audio data, adjusting an amplitude of the audio data, incorporating noise corresponding to different traffic environments into the audio data, or dampening noise in the audio data. The computing system can label the augmented audio data set with the labels of the audio data altered to generate the augmented audio data set. A machine-learning classification system can be trained with the augmented audio data set, which configures machine-learning classification system to classify objects from audio measurements captured with one or more audio devices configured to sense the environment.

Claims (29)

1 . A method comprising:

receiving, by a computing system, audio data corresponding to sounds emitted by objects capable of being identified in an environment;

altering, by the computing system, the audio data to generate an augmented audio data set having multiple different versions of the audio data; and

training a machine-learning classification system with the augmented audio data set, wherein the training configures the machine-learning classification system to classify sound measurements captured from the environment with one or more audio devices.

2 . The method of claim 1 , wherein altering the audio data to generate augmented audio data set further comprises modifying a timing of the audio data to simulate objects emitting sounds at a different temporal pace.

3 . The method of claim 1 , wherein altering the audio data to generate augmented audio data set further comprises adjusting an amplitude of the audio data to simulate objects emitting sounds at different distances.

4 . The method of claim 1 , wherein altering the audio data to generate augmented audio data set further comprises incorporating noise corresponding to different traffic environments into the audio data to simulate objects emitting sounds in a traffic situation.

5 . The method of claim 1 , wherein altering the audio data to generate augmented audio data set further comprises dampening noise in the audio data to simulate objects emitting sounds within different environments.

6 . The method of claim 1 , wherein the audio data includes labels identifying types of the objects emitting the sounds associated with the audio data, and further comprising labeling, by the computing system, the augmented audio data set with the labels of the audio data altered to generate the augmented audio data set.

7 . The method of claim 1 , further comprising classifying, by the machine-learning classification system trained with the augmented audio data set, audio measurements captured with the one or more audio devices as corresponding to a type of object in the environment around a vehicle, wherein a control system for the vehicle is configured to control operation of the vehicle based, at least in part, on the type of object corresponding to the classified audio measurements.

8 . An apparatus comprising at least one memory device storing instructions configured to cause one or more processing devices to perform operations comprising:

receiving audio data corresponding to sounds emitted by objects capable of being identified in an environment; and

altering the audio data to generate an augmented audio data set having multiple different versions of the audio data, wherein a machine-learning classification system trained with the augmented audio data set is configured to classify sound measurements captured from the environment with one or more audio devices.

9 . The apparatus of claim 8 , wherein altering the audio data to generate augmented audio data set further comprises modifying a timing of the audio data to simulate objects emitting sounds at a different temporal pace.

10 . The apparatus of claim 8 , wherein altering the audio data to generate augmented audio data set further comprises adjusting an amplitude of the audio data to simulate objects emitting sounds at different distances.

11 . The apparatus of claim 8 , wherein altering the audio data to generate augmented audio data set further comprises incorporating noise corresponding to different traffic environments into the audio data to simulate objects emitting sounds in a traffic situation.

12 . The apparatus of claim 8 , wherein altering the audio data to generate augmented audio data set further comprises dampening noise in the audio data to simulate objects emitting sounds within different environments.

13 . The apparatus of claim 8 , wherein the audio data includes labels identifying types of the objects emitting the sounds associated with the audio data, and wherein the instructions are further configured to cause the one or more processing devices to perform operations comprising labeling the augmented audio data set with the labels of the audio data altered to generate the augmented audio data set.

14 . The apparatus of claim 8 , wherein the machine-learning classification system trained with the augmented audio data set is configured to classify audio measurements captured with the one or more audio devices as corresponding to a type of object in the environment around a vehicle, wherein a control system for the vehicle is configured to control operation of the vehicle based, at least in part, on the type of object corresponding to the classified audio measurements.

15 . A system comprising:

a memory device configured to store machine-readable instructions; and

a computing system including one or more processing devices, in response to executing the machine-readable instructions, configured to:

receive audio data corresponding to sounds emitted by objects capable of being identified in an environment; and

alter the audio data to generate an augmented audio data set having multiple different versions of the audio data, wherein a machine-learning classification system trained with the augmented audio data set is configured to classify sound measurements captured from the environment with one or more audio devices.

16 . The system of claim 15 , wherein the one or more processing devices, in response to executing the machine-readable instructions, are configured to alter the audio data by modifying a timing of the audio data to simulate objects emitting sounds at a different temporal pace.

17 . The system of claim 15 , wherein the one or more processing devices, in response to executing the machine-readable instructions, are configured to alter the audio data by adjusting an amplitude of the audio data to simulate objects emitting sounds at different distances.

18 . The system of claim 15 , wherein the one or more processing devices, in response to executing the machine-readable instructions, are configured to alter the audio data by incorporating noise corresponding to different traffic environments into the audio data to simulate objects emitting sounds in a traffic situation or by dampening noise in the audio data to simulate objects emitting sounds within different environments.

19 . The system of claim 15 , wherein the audio data includes labels identifying types of the objects emitting the sounds associated with the audio data, and wherein the one or more processing devices, in response to executing the machine-readable instructions, are configured to label the augmented audio data set with the labels of the audio data altered to generate the augmented audio data set.

20 . The system of claim 15 , wherein the machine-learning classification system trained with the augmented audio data set is configured to classify audio measurements captured with the one or more audio devices as corresponding to a type of object in the environment around a vehicle, wherein a control system for the vehicle is configured to control operation of the vehicle based, at least in part, on the type of object corresponding to the classified audio measurements.

Assignments (3)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 26, 2022
From: SIEMENS INDUSTRY SOFTWARE INC.
To: SIEMENS ELECTRONIC DESIGN AUTOMATION GMBH
Reel/Frame 058778/0341 →
MERGER AND CHANGE OF NAME Recorded Jul 9, 2021
From: MENTOR GRAPHICS CORPORATION; SIEMENS INDUSTRY SOFTWARE INC.
To: SIEMENS INDUSTRY SOFTWARE INC.
Reel/Frame 056803/0385 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 5, 2021
From: SALLEM, NIZAR; BARAK, OHAD
To: MENTOR GRAPHICS CORPORATION
Reel/Frame 056143/0743 →