IP Library › Granted Patent US 12,192,734
Granted Patent B2
US 12,192,734 · App. 18/525,910 · Granted Jan 7, 2025

Parametric stereo upmix apparatus, a parametric stereo decoder, a parametric stereo downmix apparatus, a parametric stereo encoder

Inventor: Erik G. P. Schuijers (Breda, NL)
Assignee: Koninklijke Philips N.V.
H04S5/00G10L19/008H04S3/02H04S2400/03H04S2420/03
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,192,734
App. No.
18/525,910
Granted
Jan 7, 2025
Kind
B2
Abstract

A parametric stereo upmix method for generating a left signal and a right signal from a mono downmix signal based on spatial parameters includes predicting a difference signal comprising a difference between the left signal and the right signal based on the mono downmix signal scaled with a prediction coefficient. The prediction coefficient is derived from the spatial parameters. The method further includes deriving the left signal and the right signal based on a sum and a difference of the mono downmix signal and said difference signal.

Claims (309)

1. A method, comprising:

splitting an input bitstream into a mono bitstream and a parameter bitstream;

decoding the mono bitstream into a mono downmix signal;

decoding the parameter bitstream into spatial parameters;

scaling the mono downmix signal with a prediction coefficient (α) to produce a scaled mono downmix signal;

predicting a first difference signal, wherein the predicting is based on the scaled mono downmix signal;

adding a scaled decorrelated mono downmix signal to the first difference signal to form a second difference signal, wherein the scaled decorrelated mono downmix signal is formed by scaling a decorrelated mono downmix signal by a scaling factor (β);

forming the left signal based on a sum of the mono downmix signal and the second difference signal; and

forming the right signal based on a difference between the mono downmix signal and the second difference signal wherein the scaling factor (β) is:

β

=

iid

+

1

-

2

·

cos

⁡

(

ipd

)

·

icc

·

iid

iid

+

1

+

2

·

cos

⁡

(

ipd

)

·

icc

·

iid

-

❘

"\[LeftBracketingBar]"

α

❘

"\[RightBracketingBar]"

2

wherein iid, ipd, and icc are the spatial parameters,

wherein iid is an interchannel intensity difference,

wherein ipd is an interchannel phase difference, and

wherein icc is an interchannel coherence.

2. The method of claim 1 , wherein the prediction coefficient (α) is:

α

=

iid

-

1

-

j

·

2

·

sin

⁡

(

ipd

)

·

icc

·

iid

iid

+

1

+

2

·

cos

⁡

(

ipd

)

·

icc

·

iid

.

3. The method of claim 1 wherein the scaling factor (β) is derived from spatial parameters.

4. The method of claim 1 , wherein the prediction residual signal has substantially zero correlation with the mono downmix signal.

5. The method of claim 1 wherein the scaling factor (β) compensates for a prediction energy loss.

6. The method of claim 1 , wherein the prediction coefficient (α) is based on waveform matching the downmix signal onto the first difference signal.

7. A computer program stored on a non-transitory medium, wherein the computer program when executed on a processor performs the method as claimed in claim 1 .

8. A method, comprising:

splitting an input bitstream into a mono bitstream and a parameter bitstream;

extracting a prediction residual bitstream from the input bitstream;

decoding the mono bitstream into a mono downmix signal;

decoding a prediction residual signal from the prediction residual bitstream;

decoding the parameter bitstream into spatial parameters;

scaling the mono downmix signal with a prediction coefficient (α) to produce a scaled mono downmix signal;

predicting a first difference signal, wherein the predicting is based on the scaled mono downmix signal;

adding a scaled decorrelated mono downmix signal to the first difference signal to form a second difference signal, wherein the scaled decorrelated mono downmix signal is formed by scaling a decorrelated mono downmix signal by a scaling factor (β);

forming a first portion of the left signal based on a sum of the mono downmix signal, the first difference signal, and the prediction residual signal;

forming a second portion of the left signal based on a sum of the mono downmix signal and the second difference signal;

forming a first portion of the right signal based on a difference between the mono downmix signal, and a sum of the first difference signal and the prediction residual signal; and

forming a second portion of the right signal based on a difference between the mono downmix signal and the second difference signal

wherein the scaling factor (β) is:

β

=

iid

+

1

-

2

·

cos

⁡

(

ipd

)

·

icc

·

iid

iid

+

1

+

2

·

cos

⁡

(

ipd

)

·

icc

·

iid

-

❘

"\[LeftBracketingBar]"

α

❘

"\[RightBracketingBar]"

2

wherein iid, ipd, and icc are the spatial parameters,

wherein iid is an interchannel intensity difference,

wherein ipd is an interchannel phase difference, and

wherein icc is an interchannel coherence.

9. The method of claim 8 ,

wherein the first portion is a first frequency subband,

wherein the second portion is a second frequency subband,

wherein the first frequency subband is different from the second frequency subband.

10. The method of claim 8 ,

wherein the first portion comprises a first frequency subband,

wherein the second portion comprises a second frequency subband,

wherein the first frequency subband does not overlap the second frequency subband.

11. The method of claim 8 , wherein the prediction coefficient (α) is:

α

=

iid

-

1

-

j

·

2

·

sin

⁡

(

ipd

)

·

icc

·

iid

iid

+

1

+

2

·

cos

⁡

(

ipd

)

·

icc

·

iid

.

12. The method of claim 8 wherein the scaling factor (β) is derived from spatial parameters.

13. The method of claim 8 , wherein the prediction residual signal has substantially zero correlation with the mono downmix signal.

14. The method of claim 8 wherein the scaling factor (β) compensates for a prediction energy loss.

15. The method of claim 8 , wherein the prediction coefficient (α) is based on waveform matching the downmix signal onto the first difference signal.

16. A computer program stored on a non-transitory medium, wherein the computer program when executed on a processor performs the method as claimed in claim 8 .

17. A method, comprising:

splitting an input bitstream into a mono bitstream and a parameter bitstream, wherein the input bitstream comprises a plurality of subbands;

extracting a prediction residual bitstream from the input bitstream, wherein the prediction residual bitstream comprises a third portion of plurality of subbands;

decoding the mono bitstream into a mono downmix signal,

wherein the mono downmix signal comprises mono downmix subband signals,

wherein the mono downmix subband signal comprises a fourth portion of the plurality of subbands;

decoding a prediction residual signal,

wherein the prediction residual signal comprises prediction residual subband signals,

wherein the prediction residual subband signals comprise a fifth portion of the third portion of the plurality of subbands;

decoding the parameter bitstream into spatial parameters for at least one subband of the plurality of subbands;

scaling the mono downmix subband signal with a prediction coefficient (α) to produce a scaled mono downmix subband signal for at least one subband of the plurality of subbands;

predicting a first difference subband signal for at least one subband of the plurality of subbands, wherein the predicting is based on the scaled mono downmix subband signal;

adding a scaled decorrelated mono downmix subband signal to the first difference subband signal for at least one subband of the plurality of subbands to form a second difference subband signal, wherein the scaled decorrelated mono downmix subband signal is formed by scaling a decorrelated mono downmix subband signal by a scaling factor (β);

forming a first portion of the left signal,

wherein the first portion of the left signal comprises one or more subbands,

wherein each subband is based on a sum of the mono downmix subband signal, the first difference subband signal, and the prediction residual subband signal;

forming a second portion of the left signal,

wherein the second portion of the left signal comprises one or more subbands,

wherein each subband is based on a sum of the mono downmix subband signal and the second difference subband signal;

forming a first portion of the right signal,

wherein the first portion of the right signal comprises one or more subbands,

wherein each subband is based on a difference between the mono downmix subband signal, and a sum of the first difference subband signal and the prediction residual subband signal; and

forming a second portion of the right signal,

wherein the second portion of the right signal comprises one or more subbands,

wherein each subband is based on a difference between the mono downmix subband signal and the second difference subband signal

wherein the scaling factor (β) is:

β

=

iid

+

1

-

2

·

cos

⁡

(

ipd

)

·

icc

·

iid

iid

+

1

+

2

·

cos

⁡

(

ipd

)

·

icc

·

iid

-

❘

"\[LeftBracketingBar]"

α

❘

"\[RightBracketingBar]"

2

wherein iid, ipd, and icc are the spatial parameters,

wherein iid is an interchannel intensity difference,

wherein ipd is an interchannel phase difference, and

wherein icc is an interchannel coherence.

18. The method of claim 17 , wherein the prediction coefficient (α) is:

α

=

iid

-

1

-

j

·

2

·

sin

⁡

(

ipd

)

·

icc

·

iid

iid

+

1

+

2

·

cos

⁡

(

ipd

)

·

icc

·

iid

.

19. The method of claim 17 wherein the scaling factor (β) is derived from spatial parameters.

20. The method of claim 17 , wherein the prediction residual signal has substantially zero correlation with the mono downmix signal.

21. The method of claim 17 wherein the scaling factor (β) compensates for a prediction energy loss.

22. The method of claim 17 , wherein the prediction coefficient (α) is based on waveform matching the downmix signal onto the first difference signal.

23. A computer program stored on a non-transitory medium, wherein the computer program when executed on a processor performs the method as claimed in claim 17 .

Priority Claims (1)
EP 08156801 · May 23, 2008 · regional
Continuity (6)
Continuation 17324420 · May 19, 2021
Continuation 16166496 · Oct 22, 2018
Division 15411127 · Jan 20, 2017
Division 14330498 · Jul 14, 2014
Division 12992317
Related Publication 20240121567A1 · Apr 11, 2024
References Cited (30)
US 5434948A · Holt · 1995 [cited by examiner]
US 5717764A · Johnston et al. · 1998 [cited by applicant]
US 7391870B2 · Herre · 2008 [cited by examiner]
US 7573912B2 · Lindblom · 2009 [cited by examiner]
US 7933415B2 · Breebaart · 2011 [cited by applicant]
US 8811621B2 · Schuijers · 2014 [cited by applicant]
US 9591425B2 · Schuijers · 2017 [cited by applicant]
US 10136237B2 · Schuijers · 2018 [cited by applicant]
US 11871205B2 · Schuijers · 2024 [cited by examiner]
US 20060133618A1 · Mllemoes · 2006 [cited by applicant]
US 20070280485A1 · Villemoes · 2007 [cited by examiner]
US 20080031462A1 · Walsh · 2008 [cited by examiner]
US 20080199014A1 · Kurniawati et al. · 2008 [cited by applicant]
US 20100094631A1 · Engdegard · 2010 [cited by examiner]
US 20110022402A1 · Engderard · 2011 [cited by applicant]
JP H04506141A · 1992 [cited by applicant]
JP 20089026914A · 2008 [cited by applicant]
KR 20070107615A · 2007 [cited by applicant]
TW 1303411B · 2008 [cited by applicant]
TW 1427621B · 2014 [cited by applicant]
WO 09016136A1 · 1990 [cited by applicant]
WO 2003090206A1 · 2003 [cited by applicant]
WO 2006048815A1 · 2006 [cited by applicant]
WO 2006060279A1 · 2006 [cited by applicant]
WO 2006108573A1 · 2006 [cited by applicant]
WO 2007010451A1 · 2007 [cited by applicant]
Ekstrand, P.: “Bandwidth Extension of Audio Signals by Spectral Band Replication”; Proceedings of the 1SR IEEE Benelux Workshop on Model Based Processing and Coding of Audio (MPCA-2002), Leuven, Belgium, Nov. 2002, pp. … [cited by applicant]
Breebaart et al: “Parametric Coding of Stereo Audio”; EURASIP Journal on Applied Signal Processing, 2005, vol. 9, pp. 1305-1322. [cited by applicant]
Breebaart et al: “MPEG Spatial Audio Coding/MPEG Surround: Overview and Current Status”; Audio Engineering Soscity Convention Paper, 119th Convention, Oct. 2005, pp. 1-17. [cited by applicant]
Kontola et al: “AMR-WB+:Low Bit Rate Audio Coding for Mobile Multimedia”; IEEE Symposium on Broadband Multimedia Systems and Broadcasting, 2006, pp. 1-6. [cited by applicant]