IP Library Granted Patent US 8,712,785
Granted Patent B2
US 8,712,785 · App. 13/618,414 · Granted Apr 29, 2014

Method and system for reduction of quantization-induced block-discontinuities and general purpose audio codec

Inventors: Shuwu Wu (Foothill Ranch, CA); John Mantegna (Irvine, CA); Keren Perlmutter (Newport Beach, CA)
Assignee: Facebook, Inc.
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,712,785
App. No.
13/618,414
Granted
Apr 29, 2014
Kind
B2
Abstract

A method and system for reduction of quantization-induced block-discontinuities arising from lossy compression and decompression of continuous signals, especially audio signals. One embodiment encompasses a general purpose, ultra-low latency, efficient audio codec algorithm. More particularly, the invention includes a method and apparatus for compression and decompression of audio signals using a novel boundary analysis and synthesis framework to substantially reduce quantization-induced frame or block discontinuity; a novel adaptive cosine packet transform (ACPT) as the transform of choice to effectively capture the input audio characteristics; a signal-residue classifier to separate the strong signal clusters from the noise and weak signal components (collectively called residue); an adaptive sparse vector quantization (ASVQ) algorithm for signal components; a stochastic noise model for the residue; and an associated rate control algorithm. The invention further includes corresponding computer program implementations of these and other algorithms.

Claims (40)

1. A method comprising:

receiving audio data;

splitting, using at least one processor, a frame of the received audio data into a subframe;

applying a cosine packet transform to the subframe; and

determining optimal transform coefficients for the subframe.

2. The method as recited in claim 1 , wherein the optimal transform coefficients capture one or more characteristics of the received audio data.

3. The method as recited in claim 1 , further comprising performing a boundary analysis on the received audio data, wherein the boundary analysis comprises applying boundary exclusion and interpolation to the received audio data.

4. The method as recited in claim 3 , wherein applying boundary exclusion and interpolation on the received audio data comprises applying residue quantization to the received audio data.

5. The method as recited in claim 3 , further comprising normalizing output of the boundary analysis on the received audio data.

6. The method as recited in claim 1 , further comprising applying rate control to achieve a target bit rate.

7. The method as recited in claim 6 , further comprising modifying one or more parameters used by a signal and residue classifier.

8. The method as recited in claim 6 , further comprising modifying one or more parameters used by a quantization function.

9. The method as recited in claim 1 , further comprising:

applying a quantization function to strong signal components of the received audio data;

applying a stochastic noise analysis to weak signal components of the received audio data; and

formatting output of the quantization function and output of the stochastic noise analysis into a bit-stream format.

10. A non-transitory computer-readable medium including a set of instructions that, when executed by at least one processor, cause a computer system to perform steps comprising:

receiving audio data;

splitting a frame of the received audio data into a subframe;

applying a cosine packet transform to the subframe; and

determining optimal transform coefficients for the subframe.

11. The computer-readable storage medium as recited in claim 10 , further comprising instructions that, when executed, cause at least one processor to perform steps comprising:

performing a boundary analysis on the received audio data;

applying a signal and residue classifier to identify strong signal components and weak signal components of the received audio data;

applying a quantization function to strong signal components of the received audio data;

applying a stochastic noise analysis to weak signal components of the received audio data; and

formatting output of the quantization function and output of the stochastic noise analysis into a bit-stream format.

12. A method comprising:

receiving a bit stream;

generating, using at least one processor, cosine packet coefficients based on the received bit stream;

synthesizing a time-domain signal from the cosine packet coefficients; and

generating audio data based on the time-domain signal.

13. The method as recited in claim 12 , further comprising separating the received bit stream into signal components and noise components.

14. The method as recited in claim 13 , further comprising applying a stochastic noise synthesis to the noise components.

15. The method as recited in claim 13 , wherein generating cosine packet coefficients based on the received bit stream comprises applying an inverse quantization function to the signal components.

16. The method as recited in claim 15 , wherein the inverse quantization function comprises an adaptive sparse vector quantization type function.

17. The method as recited in claim 15 , wherein synthesizing the time-domain signal from the cosine packet coefficients comprises applying an inverse transform function to the cosine packet coefficients.

18. The method as recited in claim 12 , further comprising renormalizing the audio data, wherein the audio data comprises a combined signal that consists of the time-domain signal and a noise signal.

19. The method as recited in claim 12 , further comprising applying a boundary synthesis function to the audio data, wherein the audio data comprises a combined signal that consists of the time-domain signal and a noise signal.

20. The method as recited in claim 19 , further comprising clipping the audio data using one of a soft clipping technique or a hard clipping technique.

Assignments (5)
CHANGE OF NAME Recorded Dec 20, 2021
From: FACEBOOK, INC.
To: META PLATFORMS, INC.
Reel/Frame 058961/0436 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 19, 2013
From: WU, SHUWU; MANTEGNA, JOHN; PERLMUTTER, KEREN
To: AMERICA ONLINE, INC.
Reel/Frame 031033/0713 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 19, 2013
From: AOL LLC
To: AOL INC.
Reel/Frame 031033/0723 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 19, 2013
From: AOL INC.
To: FACEBOOK, INC.
Reel/Frame 031033/0827 →
CHANGE OF NAME Recorded Aug 19, 2013
From: AMERICA ONLINE, INC.
To: AOL LLC
Reel/Frame 031034/0432 →
Continuity (7)
Continuation 13191496 · Jul 27, 2011
Division 12197645 · Aug 25, 2008
Division 11609081 · Dec 11, 2006
Division 11075440 · Mar 9, 2005
Division 10061310 · Feb 4, 2002
Division 09321488 · May 27, 1999
Related Publication 20130173272A1 · Jul 4, 2013