IP Library Granted Patent US 7,191,122
Granted Patent B1
US 7,191,122 · App. 11/112,394 · Granted Mar 13, 2007

Speech compression system and method

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 7,191,122
App. No.
11/112,394
Granted
Mar 13, 2007
Kind
B1
Abstract

The invention improves the encoding and decoding of speech by focusing the encoding on the perceptually important characteristics of speech. The system analyzes selected features of an input speech signal, and first performing a common frame based speech coding of an input speech signal. The system then performs a speech coding based on either a first speech coding mode or a second speech coding mode. The selection of a mode is based on characteristics of the input speech signal. The first speech coding mode uses a first framing structure and the second speech coding mode uses a second framing structure.

Claims (30)

1. A speech compression system for processing a speech signal, the speech compression system comprising:

a mode selection module configured to select one of a first framing structure and a second framing structure for encoding parameters of a frame of the speech signal; and

a prediction module configured to predict a fixed codebook characteristic for each of a plurality of subframes when the second framing structure is selected, the prediction module is further configured to predict the fixed codebook characteristic as a function of prediction coefficients associated with each subframe and a fixed codebook characteristic from each of plurality of subframes of a previous frame;

wherein a fixed codebook gain for each of the subframe is represented with respective predicted fixed codebook characteristic when the second framing structure is selected;

wherein the speech compression system derives a pitch gain for each of a plurality of subframes during pitch pre-processing, quantizes the pitch gain of each of the subframes, and performs a delayed joint quantization of the fixed codebook gains for each of the subframes as a function of the stored quantized pitch gain of each of the subframes.

2. The speech compression system of claim 1 , wherein the first framing structure has a first bit allocation and the second framing structure has a second bit allocation, wherein the first bit allocation is different than the second bit allocation, and wherein each of the first bit allocation and the second bit allocation includes a bit allocation indicative of the selected framing structure.

3. The speech compression system of claim 1 , wherein the prediction module applies a third order moving average prediction.

4. The speech compression system of claim 1 , wherein the prediction coefficients comprise a first subframe predictor coefficient represented as {0.6, 0.3, 0.1}, a second subframe predictor coefficient represented as {0.4, 0.25, 0.1}, and a third subframe predictor coefficient represented as {0.3, 0.015, 0.075}.

5. The speech compression system of claim 1 , wherein the fixed codebook characteristic comprises fixed codebook energy.

6. The speech compression system of claim 1 , wherein a fixed codebook characteristic for each of a plurality of subframes is not predicted when the first framing structure is selected, and wherein a fixed codebook gain for each of the subframe is not represented with respective predicted fixed codebook characteristic when the first framing structure is selected.

7. A method of processing a speech signal, the method comprising:

selecting one of a first framing structure and a second framing structure for encoding parameters of a frame of the speech signal;

predicting a fixed codebook characteristic for each of a plurality of subframes when the second framing structure is selected, wherein the predicting is performed as a function of prediction coefficients associated with each subframe and a fixed codebook characteristic from each of plurality of subframes of a previous frame;

representing a fixed codebook gain for each of the subframe with respective predicted fixed codebook characteristic when the second framing structure is selected;

deriving a pitch gain for each of a plurality of subframes during pitch pre-processing;

quantizing the pitch gain of each of the subframes; and

performing a delayed joint quantization of the fixed codebook gains for each of the subframes as a function of the stored quantized pitch gain of each of the subframes.

8. The method of claim 7 , wherein the predicting and the representing are not performed when the first framing structure is selected.

9. The method of claim 7 , wherein the predicting the fixed codebook characteristic comprises applying a third order moving average prediction.

10. The method of claim 7 , wherein the prediction coefficients comprise a first subframe predictor coefficient represented as {0.6, 0.3, 0.1}, a second subframe predictor coefficient represented as {0.4, 0.25, 0.1}, and a third subframe predictor coefficient represented as {0.3, 0.015, 0.075}.

11. The method of claim 7 , wherein the fixed codebook characteristic comprises fixed codebook energy.

12. The method of claim 7 , wherein the first framing structure has a first bit allocation and the second framing structure has a second bit allocation, wherein the first bit allocation is different than the second bit allocation, and wherein each of the first bit allocation and the second bit allocation includes a bit allocation indicative of the selected framing structure.

13. A method of processing a speech signal, the method comprising:

selecting one of a first framing structure and a second framing structure for encoding parameters of a frame of the speech signal;

predicting a fixed codebook characteristic for each of a plurality of subframes when the second framing structure is selected, wherein the predicting is performed as a function of prediction coefficients associated with each subframe and a fixed codebook characteristic from each of plurality of subframes of a previous frame; and

representing a fixed codebook gain for each of the subframe with respective predicted fixed codebook characteristic when the second framing structure is selected;

wherein the prediction coefficients comprise a first subframe predictor coefficient represented as {0.6, 0.3, 0.1}, a second subframe predictor coefficient represented as {0.4, 0.25, 0.1}, and a third subframe predictor coefficient represented as {0.3, 0.015, 0.075}.

14. The method of claim 13 , wherein the predicting the fixed codebook characteristic comprises applying a third order moving average prediction.

15. The method of claim 13 , wherein the fixed codebook characteristic comprises fixed codebook energy.

16. The method of claim 13 , wherein the first framing structure has a first bit allocation and the second framing structure has a second bit allocation, wherein the first bit allocation is different than the second bit allocation, and wherein each of the first bit allocation and the second bit allocation includes a bit allocation indicative of the selected framing structure.

Assignments (10)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 3, 2020
From: INTELLECTUAL VENTURES ASSETS 142
To: DIGIMEDIA TECH, LLC
Reel/Frame 051463/0365 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 10, 2019
From: NYTELL SOFTWARE LLC
To: INTELLECTUAL VENTURES ASSETS 142 LLC
Reel/Frame 050963/0872 →
MERGER Recorded Nov 24, 2015
From: O'HEARN AUDIO LLC
To: NYTELL SOFTWARE LLC
Reel/Frame 037136/0356 →
CORRECTIVE ASSIGNMENT TO CORRECT THE GRANT LANGUAGE WITHIN THE ASSIGNMENT DOCUMENT PREVIOUSLY RECORDED ON REEL 018871 FRAME 0350. ASSIGNOR(S) HEREBY CONFIRMS THE ASSIGNMENT DOCUMENT. Recorded Dec 4, 2012
From: CONEXANT SYSTEMS, INC.
To: MINDSPEED TECHNOLOGIES, INC.
Reel/Frame 029405/0398 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 23, 2012
From: MINDSPEED TECHNOLOGIES, INC.
To: O'HEARN AUDIO LLC
Reel/Frame 029343/0322 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 28, 2011
From: WIAV SOLUTIONS LLC
To: MINDSPEED TECHNOLOGIES, INC
Reel/Frame 025717/0311 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 1, 2007
From: SKYWORKS SOLUTIONS INC.
To: WIAV SOLUTIONS LLC
Reel/Frame 019899/0305 →
EXCLUSIVE LICENSE Recorded Aug 6, 2007
From: CONEXANT SYSTEMS, INC.
To: SKYWORKS SOLUTIONS, INC.
Reel/Frame 019649/0544 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 8, 2007
From: GAO, YANG; BENYASSINE, ADIL; THYSSEN, JES; SHLOMOT, EYAL; SU, HUAN-YU
To: CONEXANT SYSTEMS, INC.
Reel/Frame 018871/0270 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 8, 2007
From: CONEXANT SYSTEMS, INC.
To: MINDSPEED TECHNOLOGIES, INC.
Reel/Frame 018871/0350 →