IP Library Granted Patent US 10,244,321
Granted Patent B2
US 10,244,321 · App. 15/498,384 · Granted Mar 26, 2019

Audio decoder for audio channel reconstruction

Inventors: Heiko Purnhagen (Sundbyberg, SE); Lars Villemoes (Järfälla, SE); Jonas Engdegard (Stockholm, SE); Jonas Roeden (Solna, SE); Kristofer Kjoerling (Solna, SE)
Assignee: Dolby International AB
H04R5/00G10L19/008G10L19/0204G10L19/032G10L19/167G10L19/26H04S3/02H04S5/00H04S2400/01H04S2400/03
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,244,321
App. No.
15/498,384
Granted
Mar 26, 2019
Kind
B2
Abstract

A method performed by an audio decoder for reconstructing N audio channels from an audio signal containing M audio channels is disclosed. The method includes receiving a bitstream containing an encoded audio signal having M audio channels and a set of spatial parameters, the set of spatial parameters including an inter-channel intensity difference parameter and an inter-channel coherence parameter. The encoded audio bitstream is then decoded to obtain a decoded frequency domain representation of the M audio channels, and at least a portion of the frequency domain representation is decorrelated with an all-pass filter having a fractional delay. The all-pass filter is attenuated at locations of a transient. A matrixed version of the decorrelated signals are summed with a matrixed version of the decoded frequency domain representation to obtain N audio signals that collectively having N audio channels where M is less than N.

Claims (27)

1. A method performed in an audio decoder for reconstructing N audio channels from M audio channels, the method comprising:

receiving an encoded audio bitstream, the encoded audio bitstream including a downmixed audio signal and surround data, the downmixed audio signal having M audio channels and the surround data including a set of spatial parameters, the set of spatial parameters including at least one inter-channel intensity difference parameter and at least one inter-channel coherence parameter;

decoding the surround data to produce decoded surround data;

decoding the downmixed audio signal having M audio channels to obtain a decoded frequency domain representation of the M audio channels, wherein the decoded frequency domain representation of the M audio channels includes a plurality of frequency bands, and each frequency band includes one or more spectral components;

reconstructing a frequency domain representation of the N audio channels from the decoded frequency domain representation of the M audio channels, downmixing information used to generate the downmixed audio signal, and the decoded surround data; and

synthesizing, with one or more synthesis filterbanks, the frequency domain representation of the N audio channels to create a time domain representation of the N audio channels; and

outputting the time domain representation of the N audio channels;

wherein M is one or more, M is less than N;

wherein the inter-channel coherence parameter is difference coded over frequency and the audio decoder is implemented at least in part with hardware.

2. The method of claim 1 , wherein the inter-channel coherence parameter is determined based on a dissimilarity of a first channel and a second channel.

3. The method of claim 1 , wherein the method further includes an analysis filterbank for decomposing the decoded representation of the M audio channels.

4. The method of claim 1 , wherein the set of spatial parameters further includes an inter-channel time or phase difference parameter.

5. The method of claim 1 , wherein the inter-channel intensity difference parameter is a ratio between the energy or level of a first channel and a second channel.

6. The method of claim 5 , wherein the first channel is a left channel, the second channel is a right channel, M=1 and N=2.

7. The method of claim 1 , wherein the M audio channels are a linear down mix of the N audio channels.

8. The method of claim 1 , wherein the decoding is performed by an MPEG-4 High Efficiency AAC decoder.

9. The method of claim 1 , wherein the synthesizing is performed with N synthesis filterbanks.

10. The method of claim 1 , wherein the synthesizing is perform with a QMF synthesis filterbank.

11. A non-transitory, computer readable storage medium containing instructions that when executed by a processor perform the method of claim 1 .

12. An audio decoder for reconstructing N audio channels from M audio channels, the audio decoder comprising:

an input interface for receiving an encoded audio bitstream, the encoded audio bitstream including a downmixed audio signal and surround data, the downmixed audio signal having M audio channels and the surround data including a set of spatial parameters, the set of spatial parameters including at least one inter-channel intensity difference parameter and at least one inter-channel coherence parameter;

a first decoder for decoding the surround data to produce decoded surround data;

a second decoder for decoding the downmixed audio signal having M audio channels to obtain a decoded frequency representation of the M audio channels, wherein the decoded frequency representation of the M audio channels includes a plurality of frequency bands, and each frequency band includes one or more spectral components;

a third decoder for reconstructing a frequency domain representation of the N audio channels from the decoded frequency domain representation of the M audio channels, downmixing information used to generate the downmixed audio signal, and the decoded surround data; and

one or more synthesis filterbanks for synthesizing, with one or more synthesis filterbanks, the frequency domain representation of the N audio channels to create a time domain representation of the N audio channels; and

wherein M is one or more, M is less than N;

wherein the inter-channel coherence parameter is difference coded over frequency.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 27, 2017
From: PURNHAGEN, HEIKO; VILLEMOES, LARS; ENGDEGARD, JONAS; ROEDEN, JONAS; KJOERLING, KRISTOFER
To: DOLBY INTERNATIONAL AB
Reel/Frame 042164/0393 →
Priority Claims (1)
SE 0400998 · Apr 16, 2004 · national
Continuity (5)
Continuation 13866947 · Apr 19, 2013
Continuation 12882894 · Sep 15, 2010
Division 11549963 · Oct 16, 2006
Continuation PCTEP2005003849 · Apr 12, 2005
Related Publication 20170229128A1 · Aug 10, 2017