IP Library Granted Patent US 9,691,392
Granted Patent B1
US 9,691,392 · App. 15/071,258 · Granted Jun 27, 2017

System and method for improved audio consistency

Inventor: Umesh Sachdev (Chennai, IN)
Assignee: UNIPHORE SOFTWARE SYSTEMS
G10L17/04G10L21/0208G10L25/90G10L25/93
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,691,392
App. No.
15/071,258
Granted
Jun 27, 2017
Kind
B1
Abstract

A voice biometrics system adapted to authenticate a user based on speech diagnostics is provided. The system includes a pre-processing module to receive and pre-process an input voice sample. The pre-processing module includes a clipping module to clip the input voice sample based on a clipping threshold and a voice activity detection module to apply a detection model on the input voice sample to determine an audible region and a non-audible region in the input voice sample. The pre-processing module includes a noise reduction module to apply a noise reduction model to remove noise components from the input voice sample. The voice biometrics system includes a feature extraction module to extract features from the pre-processed input voice sample. The voice biometrics system also includes an authentication module to authenticate the user by comparing a plurality of features extracted from the pre-processed input voice sample to a plurality of enrollment features.

Claims (50)

1. A voice biometrics system adapted to authenticate a user based on speech diagnostics, the voice biometrics system comprising:

a pre-processing module configured to receive an input voice sample and pre-process the input voice sample by:

a clipping module configured to clip the input voice sample based on a clipping threshold;

a voice activity detection module configured to apply a detection model on the input voice sample to determine an audible region and a non-audible region in the input voice sample; and

a noise reduction module configured to apply a noise reduction model to remove noise components from the input voice sample;

a feature extraction module configured to extract features from the pre-processed input voice sample; and

an authentication model configured to authenticate the user by comparing a plurality of features extracted from the pre-processed input voice sample to a plurality of enrolment features, wherein the voice activity detection module further comprises a zero crossing module configured to detect the polarity of the input voice sample across a time.

2. The voice biometrics system of claim 1 , wherein the pre-processing module further comprises:

a pre-emphasis module configured to remove the low frequency components from the input voice sample; and

an amplification module configured to amplify the magnitude of the input voice sample.

3. The voice biometrics system of claim 1 , further comprising a post-processing module configured to apply a gaussian mixture model to detect the input channel and/or device through which the features from the voice samples are entered.

4. The voice biometrics system of claim 1 , wherein the voice activity detection module further comprises a short time energy module configured to classify the audible region and the non-audible region of the input voice sample.

5. The voice biometrics system of claim 1 , wherein the voice activity detection module further comprises a pitch detection module configured to estimate a pitch level of the input voice sample.

6. The voice biometrics system of claim 1 , wherein the voice activity detection module further comprises voice activity detection sub-system configured to detect plurality of speech frames comprising speech and non-speech frames of the input voice sample.

7. The voice biometrics system of claim 1 , wherein the pre-processing module further comprises a feature normalization module configured to apply a mean and variance normalization model to remove noise components from the input voice sample caused by the input channel and/or device.

8. A voice biometrics system adapted to authenticate a user based on speech diagnostics, the voice biometrics system comprising:

a pre-processing module configured to receive an input voice sample and pre-process the input voice sample by:

a clipping module configured to clip the input voice sample based on a clipping threshold;

a voice activity detection module configured to apply a detection model on the input voice sample to determine an audible region and a non-audible region in the input voice sample; and

a noise reduction module configured to apply a noise reduction model to remove noise components from the input voice sample;

a feature extraction module configured to extract features from the pre-processed input voice sample; and

an authentication model configured to authenticate the user by comparing a plurality of features extracted from the pre-processed input voice sample to a plurality of enrolment features, further comprising a post-processing module configured to apply a gaussian mixture model to detect the input channel and/or device through which the features from the voice samples are entered.

9. A voice biometrics system adapted to authenticate a user based on speech diagnostics, the voice biometrics system comprising:

a pre-processing module configured to receive an input voice sample and pre-process the input voice sample by:

a clipping module configured to clip the input voice sample based on a clipping threshold;

a voice activity detection module configured to apply a detection model on the input voice sample to determine an audible region and a non-audible region in the input voice sample; and

a noise reduction module configured to apply a noise reduction model to remove noise components from the input voice sample;

a feature extraction module configured to extract features from the pre-processed input voice sample; and

an authentication model configured to authenticate the user by comparing a plurality of features extracted from the pre-processed input voice sample to a plurality of enrolment features, wherein the voice activity detection module further comprises a pitch detection module configured to estimate a pitch level of the input voice sample.

10. A voice biometrics system adapted to authenticate a user based on speech diagnostics, the voice biometrics system comprising:

a pre-processing module configured to receive an input voice sample and pre-process the input voice sample by:

a clipping module configured to clip the input voice sample based on a clipping threshold;

a voice activity detection module configured to apply a detection model on the input voice sample to determine an audible region and a non-audible region in the input voice sample; and

a noise reduction module configured to apply a noise reduction model to remove noise components from the input voice sample;

a feature extraction module configured to extract features from the pre-processed input voice sample; and

an authentication model configured to authenticate the user by comparing a plurality of features extracted from the pre-processed input voice sample to a plurality of enrolment features, wherein the voice activity detection module further comprises voice activity detection sub-system configured to detect plurality of speech frames comprising speech and non-speech frames of the input voice sample.

11. A voice biometrics system adapted to authenticate a user based on speech diagnostics, the voice biometrics system comprising:

a pre-processing module configured to receive an input voice sample and pre-process the input voice sample by:

a clipping module configured to clip the input voice sample based on a clipping threshold;

a voice activity detection module configured to apply a detection model on the input voice sample to determine an audible region and a non-audible region in the input voice sample; and

a noise reduction module configured to apply a noise reduction model to remove noise components from the input voice sample;

a feature extraction module configured to extract features from the pre-processed input voice sample; and

an authentication model configured to authenticate the user by comparing a plurality of features extracted from the pre-processed input voice sample to a plurality of enrolment features, wherein the pre-processing module further comprises a feature normalization module configured to apply a mean and variance normalization model to remove noise components from the input voice sample caused by the input channel and/or device.

12. A voice biometrics system adapted to authenticate a user based on speech diagnostics, the voice biometrics system comprising:

a pre-processing module configured to receive an input voice sample and pre-process the input voice sample by:

a clipping module configured to clip the input voice sample based on a clipping threshold;

a voice activity detection module configured to apply a detection model on the input voice sample to determine an audible region and a non-audible region in the input voice sample; and

a noise reduction module configured to apply a noise reduction model to remove noise components from the input voice sample;

a feature extraction module configured to extract features from the pre-processed input voice sample; and

an authentication model configured to authenticate the user by comparing a plurality of features extracted from the pre-processed input voice sample to a plurality of enrolment features, wherein the voice activity detection module further comprises a short time energy module configured to classify the audible region and the non-audible region of the input voice sample.

Assignments (4)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 9, 2026
From: UNIPHORE SOFTWARE SYSTEMS
To: UNIPHORE TECHNOLOGIES INC.
Reel/Frame 074017/0631 →
SECURITY INTEREST Recorded Jan 20, 2023
From: UNIPHORE TECHNOLOGIES INC.; UNIPHORE TECHNOLOGIES NORTH AMERICA INC.; UNIPHORE SOFTWARE SYSTEMS INC.; COLABO, INC.
To: HSBC VENTURES USA INC.
Reel/Frame 062440/0619 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 21, 2022
From: UNIPHORE SOFTWARE SYSTEMS
To: UNIPHORE TECHNOLOGIES INC.
Reel/Frame 061841/0541 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 29, 2016
From: SACHDEV, UMESH
To: UNIPHORE SOFTWARE SYSTEMS
Reel/Frame 039292/0467 →
Priority Claims (1)
IN 6580/CHE/2015 · Dec 9, 2015 · national