System and method for data augmentation of feature-based voice data
A method, computer program product, and computing system for receiving feature-based voice data associated with a first acoustic domain. One or more gain-based augmentations may be performed on at least a portion of the feature-based voice data, thus defining gain-augmented feature-based voice data.
1. A computer-implemented method, executed on a computing device, comprising:
receiving feature-based voice data associated with a first acoustic domain, wherein the feature-based voice data is converted from a signal in the first acoustic domain to a feature domain;
performing one or more gain-based augmentations on at least a portion of the feature-based voice data converted from the signal in the first acoustic domain to the feature domain, thus defining gain-augmented feature-based voice data;
receiving a selection of a target acoustic domain;
determining a distribution of gain levels from training data associated with the target acoustic domain varies over time for one or more of particular frequencies and particular frequency bands;
mapping the gain-augmented feature-based voice data from the first acoustic domain to the target acoustic domain, and
training a speech processing system for the target acoustic domain based on the gain-augmented feature-based voiced data mapped from the first acoustic domain to the target acoustic domain.
2. The computer-implemented method of claim 1 , further comprising:
receiving the selection of the target acoustic domain via at least one of a user interface and a database.
3. The computer-implemented method of claim 2 , wherein performing the one or more gain-based augmentations to the at least a portion of the feature-based voice data includes performing the one or more gain-based augmentations to the at least a portion of the feature-based voice data based upon, at least in part, the target acoustic domain.
4. The computer-implemented method of claim 2 , further comprising:
determining a distribution of gain levels associated with the target acoustic domain.
5. The computer-implemented method of claim 4 , wherein performing the one or more gain-based augmentations to the at least a portion of the feature-based voice data includes performing the one or more gain-based augmentations to the at least a portion of the feature-based voice data based upon, at least in part, the distribution of gain levels associated with the target acoustic domain.
6. The computer-implemented method of claim 1 , wherein performing the one or more gain-based augmentations to the at least a portion of the feature-based voice data includes amplifying the at least a portion of the feature-based voice data.
7. The computer-implemented method of claim 1 , wherein performing the one or more gain-based augmentations to the at least a portion of the feature-based voice data includes attenuating the at least a portion of the feature-based voice data.
8. A computer program product residing on a non-transitory computer readable medium having a plurality of instructions stored thereon which, when executed by a processor, cause the processor to perform operations comprising:
receiving feature-based voice data associated with a first acoustic domain, wherein the feature-based voice data is converted from a signal in the first acoustic domain to a feature domain;
performing one or more gain-based augmentations on at least a portion of the feature-based voice data converted from the signal in the first acoustic domain to the feature domain, thus defining gain-augmented feature-based voice data;
receiving a selection of a target acoustic domain;
determining a distribution of gain levels from training data associated with the target acoustic domain varies over time for one or more of particular frequencies and particular frequency bands;
mapping the gain-augmented feature-based voice data from the first acoustic domain to the target acoustic domain, and
training a speech processing system for the target acoustic domain based on the gain-augmented feature-based voiced data mapped from the first acoustic domain to the target acoustic domain.
9. The computer program product of claim 8 , wherein the operations further comprise:
receiving the selection of the target acoustic domain via at least one of a user interface and a database.
10. The computer program product of claim 9 , wherein performing the one or more gain-based augmentations to the at least a portion of the feature-based voice data includes performing the one or more gain-based augmentations to the at least a portion of the feature-based voice data based upon, at least in part, the target acoustic domain.
11. The computer program product of claim 9 , wherein the operations further comprise:
determining a distribution of gain levels associated with the target acoustic domain.
12. The computer program product of claim 11 , wherein performing the one or more gain-based augmentations to the at least a portion of the feature-based voice data includes performing the one or more gain-based augmentations to the at least a portion of the feature-based voice data based upon, at least in part, the distribution of gain levels associated with the target acoustic domain.
13. The computer program product of claim 8 , wherein performing the one or more gain-based augmentations to the at least a portion of the feature-based voice data includes amplifying the at least a portion of the feature-based voice data.
14. The computer program product of claim 8 , wherein performing the one or more gain-based augmentations to the at least a portion of the feature-based voice data includes attenuating the at least a portion of the feature-based voice data.
15. A computing system comprising:
a memory; and
a processor configured to receive feature-based voice data associated with a first acoustic domain, wherein the feature-based voice data is converted from a signal in the first acoustic domain to a feature domain, wherein the processor is further configured to perform one or more gain-based augmentations on at least a portion of the feature-based voice data converted from the signal in the first acoustic domain to the feature domain, thus defining gain-augmented feature-based voice data, wherein the processor is further configured to receive a selection of a target acoustic domain, wherein the processor is further configured to determine a distribution of gain levels from training data associated with the target acoustic domain varies over time for one or more of particular frequencies and particular frequency bands; wherein the processor is further configured to map the gain-augmented feature-based voice data from the first acoustic domain to the target acoustic domain; and wherein the processor is further configured to train a speech processing system for the target acoustic domain based on the gain-augmented feature-based voiced data mapped from the first acoustic domain to the target acoustic domain.
16. The computing system of claim 15 , wherein the processor is further configured to:
receive the selection of the target acoustic domain via at least one of a user interface and a database.
17. The computing system of claim 16 , wherein performing the one or more gain-based augmentations to the at least a portion of the feature-based voice data includes performing the one or more gain-based augmentations to the at least a portion of the feature-based voice data based upon, at least in part, the target acoustic domain.
18. The computing system of claim 16 , wherein the processor is further configured to:
determine a distribution of gain levels associated with the target acoustic domain.
19. The computing system of claim 18 , wherein performing the one or more gain-based augmentations to the at least a portion of the feature-based voice data includes performing the one or more gain-based augmentations to the at least a portion of the feature-based voice data based upon, at least in part, the distribution of gain levels associated with the target acoustic domain.
20. The computing system of claim 15 , wherein performing the one or more gain-based augmentations to the at least a portion of the feature-based voice data includes one or more of:
amplifying the at least a portion of the feature-based voice data; and
attenuating the at least a portion of the feature-based voice data.