IP Library Granted Patent US 8,650,028
Granted Patent B2
US 8,650,028 · App. 12/229,324 · Granted Feb 11, 2014

Multi-mode speech encoding system for encoding a speech signal used for selection of one of the speech encoding modes including multiple speech encoding rates

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,650,028
App. No.
12/229,324
Granted
Feb 11, 2014
Kind
B2
Abstract

A method comprises analyzing each frame of a plurality of frames of the speech signal to determine one or more speech parameters for the speech signal; deciding, for each frame of the plurality of frames of the speech signal, based on the one or more speech parameters of the speech signal, to select one of a plurality of encoding modes including a first encoding mode and a second encoding mode for encoding each frame of the plurality of frames of the speech signal; encoding each frame of the plurality of frames of the speech signal according to the selected one of the plurality of encoding modes for each frame of the plurality of frames in the deciding; the first encoding mode supports a first encoding rate and the second encoding mode supports a second encoding rate, wherein the first encoding rate is the same encoding rate as the encoding rate.

Claims (32)

1. A method of encoding a speech signal, the method comprising:

analyzing each frame of a plurality of frames of the speech signal to determine one or more speech parameters for the speech signal, wherein one parameter of the one or more speech parameters includes one or more pitch lags;

deciding, for each frame of the plurality of frames of the speech signal, based on the one or more speech parameters of the speech signal, to select one of a plurality of encoding modes including a first encoding mode, a second encoding mode and a third encoding mode for encoding each frame of the plurality of frames of the speech signal;

converting the speech signal into an encoded speech by encoding each frame of the plurality of frames of the speech signal according to the selected one of the plurality of encoding modes for each frame of the plurality of frames in the deciding;

wherein the first encoding mode supports a first encoding rate, the second encoding mode supports a second encoding rate and the third encoding mode supports a third encoding rate, wherein the first encoding rate is the same encoding rate as the second encoding rate, wherein the third encoding rate is different than the first encoding rate and the second encoding rate, and wherein the converting of the speech signal to the encoded speech signal for each frame of the plurality of frames of the speech signal further comprises encoding a single pitch lag of the one or more pitch lags if the encoding mode is one of the second encoding mode and the third encoding mode;

wherein the converting of the speech signal to an encoded speech signal for each frame of the plurality of frames of the speech signal comprises encoding of a single pitch lag of the one or more pitch lags if the encoding mode is the second encoding mode or the third encoding mode.

2. The method of claim 1 , wherein the first encoding rate and the second encoding rate are both at 6.65 kbps, and the third encoding rate is 5.80 kbps.

3. The method of claim 1 , wherein the first encoding mode is long-term prediction mode (LTP_mode) and the second encoding mode and the third encoding mode are pitch preprocessing mode (PP_mode).

4. The method of claim 1 , wherein the deciding is based on the one or more speech parameters of the speech signal including a pitch lag parameter.

5. The method of claim 1 , wherein the deciding is based on the one or more speech parameters of the speech signal including a pitch gain parameter.

6. The method of claim 1 , wherein the deciding is based on the one or more speech parameters of the speech signal including a line spectrum frequency LSF parameter.

7. The method of claim 1 , wherein the deciding is based on the one or more speech parameters of the speech signal including a pitch correlation parameter.

8. The method of claim 1 , wherein the deciding is based on the one or more speech parameters of the speech signal including linear prediction analysis parameters.

9. The method of claim 8 , wherein the deciding is based on the one or more speech parameters of the speech signal including a distance measure between linear prediction analysis parameters.

10. The method of claim 1 , wherein the first encoding mode supports a plurality of encoding rates including the first encoding rate and the second encoding mode supports a plurality of encoding rates including the second encoding rate.

11. A speech encoding system for encoding a speech signal, the speech encoding system comprising:

an encoder processing circuit configured to:

analyze each frame of a plurality of frames of the speech signal to determine one or more speech parameters for the speech signal, wherein one parameter of the one or more speech parameters includes one or more pitch lags;

decide, for each frame of the plurality of frames of the speech signal, based on the one or more speech parameters of the speech signal, to select one of a plurality of encoding modes including a first encoding, a second encoding mode and a third encoding mode for encoding each frame of the plurality of frames of the speech signal;

convert the speech signal into an encoded speech by encoding each frame of the plurality of frames of the speech signal according to the selected one of the plurality of encoding modes for each frame of the plurality of frames, thereby converting the speech signal into an encoded speech;

wherein the first encoding mode supports a first encoding rate, the second encoding mode supports a second encoding rate and the third encoding mode supports a third encoding rate, wherein the first encoding rate is the same encoding rate as the second encoding rate, wherein the third encoding rate is different than the first encoding rate and the second encoding rate, wherein converting the speech signal to the encoded speech signal for each frame of the plurality of frames of the speech signal further comprises encoding a single pitch lag of the one or more pitch lags if the encoding mode is one of the second encoding mode and the third encoding mode.

12. The speech encoding system of claim 11 , wherein the first encoding rate and the second encoding rate are both at 6.65 kbps, and the third encoding rate is 5.80 kbps.

13. The speech encoding system of claim 11 , wherein the first encoding mode is long-term prediction mode (LTP_mode) and the second encoding mode and the third encoding mode are pitch preprocessing mode (PP_mode).

14. The speech encoding system of claim 11 , wherein the encoder processing circuit is configured to decide based on the one or more speech parameters of the speech signal including a pitch lag parameter.

15. The speech encoding system of claim 11 , wherein the encoder processing circuit is configured to decide based on the one or more speech parameters of the speech signal including a pitch gain parameter.

16. The speech encoding system of claim 11 , wherein the encoder processing circuit is configured to decide based on the one or more speech parameters of the speech signal including a line spectrum frequency LSF parameter.

17. The speech encoding system of claim 11 , wherein the encoder processing circuit is configured to decide based on the one or more speech parameters of the speech signal including a pitch correlation parameter.

18. The speech encoding system of claim 11 , wherein the encoder processing circuit is configured to decide based on the one or more speech parameters of the speech signal including linear prediction analysis parameters.

19. The speech encoding system of claim 18 , wherein the encoder processing circuit is configured to decide based on the one or more speech parameters of the speech signal including a distance measure between linear prediction analysis parameters.

20. The speech encoding system of claim 11 , wherein the first encoding mode supports a plurality of encoding rates including the first encoding rate and the second encoding mode supports a plurality of encoding rates including the second encoding rate.

21. The method of claim 1 , wherein the converting of the speech signal to the encoded speech signal for each frame of the plurality of frames of the speech signal uses Code Excited Linear Prediction (CELP) if the encoding mode is the first encoding mode.

22. The speech encoding system of claim 11 , wherein converting the speech signal to the encoded speech signal for each frame of the plurality of frames of the speech signal uses Code Excited Linear Prediction (CELP) if the encoding mode is the first encoding mode.

Assignments (9)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 7, 2017
From: MINDSPEED TECHNOLOGIES, LLC
To: MACOM TECHNOLOGY SOLUTIONS HOLDINGS, INC.
Reel/Frame 044791/0600 →
CHANGE OF NAME Recorded Aug 19, 2016
From: MINDSPEED TECHNOLOGIES, INC.
To: MINDSPEED TECHNOLOGIES, LLC
Reel/Frame 039755/0323 →
CHANGE OF NAME Recorded Aug 10, 2016
From: MINDSPEED TECHNOLOGIES, INC.
To: MINDSPEED TECHNOLOGIES, LLC
Reel/Frame 039645/0264 →
SECURITY INTEREST Recorded May 9, 2014
From: M/A-COM TECHNOLOGY SOLUTIONS HOLDINGS, INC.; MINDSPEED TECHNOLOGIES, INC.; BROOKTREE CORPORATION
To: GOLDMAN SACHS BANK USA
Reel/Frame 032859/0374 →
RELEASE OF SECURITY INTEREST Recorded May 9, 2014
From: JPMORGAN CHASE BANK, N.A.
To: MINDSPEED TECHNOLOGIES, INC.
Reel/Frame 032861/0617 →
SECURITY INTEREST Recorded Mar 21, 2014
From: MINDSPEED TECHNOLOGIES, INC.
To: JPMORGAN CHASE BANK, N.A., AS ADMINISTRATIVE AGENT
Reel/Frame 032495/0177 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 28, 2011
From: WIAV SOLUTIONS LLC
To: MINDSPEED TECHNOLOGIES, INC
Reel/Frame 025717/0356 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 20, 2008
From: CONEXANT SYSTEM, INC.
To: MINDSPEED TECHNOLOGIES, INC.
Reel/Frame 021479/0947 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 20, 2008
From: GAO, YANG; SU, HUAN-YU
To: CONEXANT SYSTEMS, INC.
Reel/Frame 021479/0307 →