IP Library Granted Patent US 11,270,721
Granted Patent B2
US 11,270,721 · App. 16/418,828 · Granted Mar 8, 2022

Systems and methods of pre-processing of speech signals for improved speech recognition

Inventors: Youhong Lu (Irvine, CA); Arun Rajasekaran (Saratoga, CA)
Assignee: PLANTRONICS, INC.
G10L25/90G10L15/22G10L25/24G06F3/167
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,270,721
App. No.
16/418,828
Granted
Mar 8, 2022
Kind
B2
Abstract

Pre-processing systems, methods of pre-processing, and speech processing systems for improved Automated Speech Recognition are provided. Some pre-processing systems for improved speech recognition of a speech signal are provided, which systems comprise a pitch estimation circuit; and a pitch equalization processor. The pitch estimation circuit is configured to receive the speech signal to determine a pitch index of the speech signal, and the pitch equalization processor is configured to receive the speech signal and pitch information, to equalize a speech pitch of the speech signal using the pitch information, and to provide a pitch-equalized speech signal.

Claims (63)

1. A pre-processing system for improved speech recognition of a speech signal, the system comprising at least

a pitch estimation circuit; and

a pitch equalization processor; wherein

the pitch estimation circuit is configured to receive the speech signal to determine a pitch index of the speech signal; and wherein

the pitch equalization processor is configured to

receive the speech signal and pitch information;

equalize a speech pitch of the speech signal using the pitch information through removal of the speech pitch of the speech signal and re-synthesis of the speech pitch of the speech signal with a predefined speech recognition pitch index; and to

provide a pitch-equalized speech signal.

2. The pre-processing system of claim 1 , further comprising a pitch averaging circuit; which pitch averaging circuit is configured to receive the pitch index from the pitch estimation circuit, to determine an average pitch index, and to provide the pitch information to the pitch equalization processor, which pitch information corresponds to the average pitch index.

3. The pre-processing system of claim 1 , wherein the pitch equalization processor is configured to equalize by normalizing the speech pitch of the speech signal.

4. The pre-processing system of claim 3 , wherein the pitch equalization processor is configured to normalize the speech pitch of the speech signal to the predefined speech recognition pitch index.

5. The pre-processing system of claim 1 , wherein the pitch equalization processor is configured to filter the speech pitch.

6. The pre-processing system of claim 1 , wherein the pitch equalization processor is configured to provide the pitch-equalized speech signal to one or more of an automatic speech recognition system and an acoustic model system.

7. A method of pre-processing of a speech signal for improved speech recognition, comprising the steps of

receiving the speech signal;

determining a pitch index of the speech signal;

equalizing a speech pitch of the speech signal using pitch information through removal of the speech pitch of the speech signal and re-synthesis of the speech pitch of the speech signal with a predefined speech recognition pitch index; and

providing a pitch-equalized speech signal.

8. The method of claim 7 , further comprising determining the pitch information by calculating an average pitch index from the determined pitch index.

9. A speech processing system, comprising

a pre-processor for improved speech recognition of a speech signal; and

one or more of a speech recognition processor and an acoustic modeler; wherein

the pre-processor comprising at least

a pitch estimation circuit; and

a pitch equalization processor; wherein

the pitch estimation circuit is configured to receive the speech signal to determine a pitch index of the speech signal; and wherein

the pitch equalization processor is configured to receive the speech signal and pitch information;

to equalize a speech pitch of the speech signal using the pitch information through removal of the speech pitch of the speech signal and re-synthesis of the speech pitch of the speech signal with a predefined speech recognition pitch index; and to

provide a pitch-equalized speech signal to one or more of the speech recognition processor and the acoustic modeler.

10. A headset system with a speech processing system of claim 9 .

11. A pre-processing system for improved speech recognition of a speech signal, the system comprising at least

a pitch estimation circuit; and

a pitch classification processor; wherein

the pitch estimation circuit is configured to:

receive the speech signal to determine a pitch index of the speech signal;

to equalize a speech pitch of the speech signal using the pitch information through removal of the speech pitch of the speech signal and re-synthesis of the speech pitch of the speech signal with a predefined speech recognition pitch index; and wherein

the pitch classification processor is configured to

receive pitch information; and to

determine classification information of the speech signal using the pitch information.

12. The pre-processing system of claim 11 , further comprising a pitch averaging circuit; which pitch averaging circuit is configured to receive the pitch index from the pitch estimation circuit, to determine an average pitch index, and to provide the pitch information corresponding to the averaged pitch index to the pitch classification processor.

13. The pre-processing system of claim 11 , wherein the pitch classification processor is configured with multiple pitch bins and is configured to determine the classification information by determining a correlation between one of the pitch bins and the speech signal.

14. The pre-processing system of claim 13 , wherein each pitch bin is associated with one of a plurality of acoustic models of an automatic speech recognition system.

15. The pre-processing system of claim 11 , wherein the pitch classification processor is configured to provide at least the classification information to one or more of an automatic speech recognition system and an acoustic model system.

16. A method of pre-processing of a speech signal for improved speech recognition, comprising the steps of

receiving the speech signal;

determining a pitch index of the speech signal;

equalizing a speech pitch of the speech signal using the pitch information through removal of the speech pitch of the speech signal and re-synthesis of the speech pitch of the speech signal with a predefined speech recognition pitch index; and determining classification information of the speech signal using pitch information.

17. The method of claim 16 , further comprising determining the pitch information by calculating an average pitch index from the determined pitch index.

18. A speech processing system, comprising

a pre-processor for improved speech recognition of a speech signal; and

one or more of a speech recognition processor and an acoustic modeler; wherein

the pre-processor comprising at least

a pitch estimation circuit configured to equalize a speech pitch of the speech signal using the pitch information through removal of the speech pitch of the speech signal and re-synthesis of the speech pitch of the speech signal with a predefined speech recognition pitch index; and

a pitch classification processor; wherein

the pitch estimation circuit is configured to receive the speech signal to determine a pitch index of the speech signal; and wherein

the pitch classification processor is configured to

receive pitch information;

determine classification information of the speech signal using the pitch information; and to

provide the classification information to one or more of the speech recognition processor and the acoustic modeler.

19. A headset system with a speech processing system of claim 18 .

20. The pre-processing system of claim 1 , wherein the pitch index is of a fundamental frequency.

21. The pre-processing system of claim 1 , wherein the pitch estimation circuit is further configured to equalize the speech pitch of the speech signal using the pitch information through application of a fixed pitch filter with the predefined speech recognition pitch index after an average pitch is inverse-filtered.

22. The pre-processing system of claim 1 , wherein the pitch estimation circuit is further configured to equalize the speech pitch to have a fixed pitch.

Assignments (5)
NUNC PRO TUNC ASSIGNMENT Recorded Nov 13, 2023
From: PLANTRONICS, INC.
To: HEWLETT-PACKARD DEVELOPMENT COMPANY, L.P.
Reel/Frame 065549/0065 →
RELEASE OF PATENT SECURITY INTERESTS Recorded Aug 30, 2022
From: WELLS FARGO BANK, NATIONAL ASSOCIATION
To: PLANTRONICS, INC.; POLYCOM, INC.
Reel/Frame 061356/0366 →
SUPPLEMENTAL SECURITY AGREEMENT Recorded Oct 16, 2020
From: PLANTRONICS, INC.; POLYCOM, INC.
To: WELLS FARGO BANK, NATIONAL ASSOCIATION
Reel/Frame 054090/0467 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 31, 2019
From: LU, YOUHONG
To: PLANTRONICS, INC.
Reel/Frame 049920/0016 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 22, 2019
From: RAJASEKARAN, ARUN
To: PLANTRONICS, INC.
Reel/Frame 049259/0619 →
Continuity (2)
Provisional Application 62674226 · May 21, 2018
Related Publication 20190355385A1 · Nov 21, 2019
Cited By (1)
US 12,505,830