IP Library › Granted Patent US 12,456,478
Granted Patent B2
US 12,456,478 · App. 17/501,993 · Granted Oct 28, 2025

Apparatus, method or computer program for generating an output downmix representation

Inventors: Franz Reutelhuber (Erlangen, DE); Eleni Fotopoulou (Erlangen, DE); Markus Multrus (Erlangen, DE)
Assignee: Fraunhofer-Gesellschaft zur Förderung der angewandten Forschung e.V.
G10L21/04G10L19/008G10L19/022H04S1/007H04S7/30H04S2400/03
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,456,478
App. No.
17/501,993
Granted
Oct 28, 2025
Kind
B2
Abstract

An apparatus for generating an output downmix representation from an input downmix representation, wherein at least a portion of the input downmix representation is in accordance with a first downmixing scheme, includes: an upmixer for upmixing at least the portion of the input downmix representation using an upmixing scheme corresponding to the first downmixing scheme to obtain at least one upmixed portion; and a downmixer for downmixing the at least one upmixed portion in accordance with a second downmixing scheme different from the first downmixing scheme.

Claims (39)

1. An apparatus for generating an output downmix representation from an input downmix representation, wherein a first portion of the input downmix representation is in accordance with a first downmixing scheme, and a second portion of the input downmix representation is in accordance with a second downmixing scheme, the apparatus comprising:

an upmixer for upmixing the first portion of the input downmix representation using an upmixing scheme corresponding to the first downmixing scheme to acquire at least one first upmixed portion,

wherein the first portion of the input downmix representation is a first frequency band of a spectrum and the second portion of the input downmix representation is a second frequency band of the spectrum;

a downmixer for downmixing the at least one first upmixed portion in accordance with the second downmixing scheme different from the first downmixing scheme to acquire a first downmixed portion; and

a combiner configured for combining the first downmixed portion and the second portion of the input downmix representation to obtain the output downmix representation comprising a first output representation for the first portion of the input downmix representation and a second output representation for the second portion of the input downmix representation,

wherein the first output representation for the first portion of the input downmix representation and the second output representation for the second portion of the input downmix representation represent the output downmix representation and are based on the second downmixing scheme.

2. A method for generating an output downmix representation from an input downmix representation-, wherein a first portion of the input downmix representation is in accordance with a first downmixing scheme, and wherein a second portion of the input downmix representation is in accordance with a second downmixing scheme, the method comprising:

upmixing the first portion of the input downmix representation using an upmixing scheme corresponding to the first downmixing scheme to acquire at least one first upmixed portion,

wherein the first portion of the input downmix representation is a first frequency band of a spectrum and the second portion of the input downmix representation is a second frequency band of the spectrum;

downmixing the at least one first upmixed portion in accordance with the second downmixing scheme different from the first downmixing scheme to acquire a first downmixed portion; and

combining the first downmixed portion and the second portion of the input downmix representation to obtain the output downmix representation comprising a first output representation for the first portion of the input downmix representation and a second output representation for the second portion of the input downmix representation,

wherein the first output representation for the first portion of the input downmix representation and the second output representation for the second portion of the input downmix representation represent the output downmix representation and are based on the second downmixing scheme.

3. A non-transitory digital storage medium having a computer program stored thereon to perform, when said computer program is run by a computer, a method for generating an output downmix representation from an input downmix representation, wherein a first portion of the input downmix representation is in accordance with a first downmixing scheme, and wherein a second portion of the input downmix representation is in accordance with a second downmixing scheme, the method comprising:

upmixing the first portion of the input downmix representation using an upmixing scheme corresponding to the first downmixing scheme to acquire at least one first upmixed portion,

wherein the first portion of the input downmix representation is a first frequency band of a spectrum and the second portion of the input downmix representation is a second frequency band of the spectrum;

downmixing the at least one first upmixed portion in accordance with the second downmixing scheme different from the first downmixing scheme to acquire a first downmixed portion; and

combining the first downmixed portion and the second portion of the input downmix representation to obtain the output downmix representation comprising a first output representation for the first portion of the input downmix representation and a second output representation for the second portion of the input downmix representation,

wherein the first output representation for the first portion of the input downmix representation and the second output representation for the second portion of the input downmix representation represent the output downmix representation and are based on the second downmixing scheme.

4. An apparatus for generating an output downmix representation, comprising:

an interface for receiving an input downmix representation, wherein a first portion of the input downmix representation is a first frequency band of a spectrum and is in accordance with a first downmixing scheme and a second portion of the input downmix representation is a second frequency band of the spectrum and is in accordance with a second downmixing scheme, wherein the second downmixing scheme is different from the first downmixing scheme;

an upmixer for upmixing the first portion of the input downmix representation using a first upmixing scheme corresponding to the first downmixing scheme to acquire at least one first upmixed portion and for upmixing the second portion of the input downmix representation using a second upmixing scheme corresponding to the second downmixing scheme to acquire at least one second upmixed portion; and

a downmixer for downmixing the at least one first upmixed portion in accordance with a third downmixing scheme to acquire a first portion of the output downmix representation and for downmixing the at least one second upmixed portion in accordance with the third downmixing scheme to acquire a second portion of the output downmix representation,

wherein the third downmixing scheme is different from the first downmixing scheme and the second downmixing scheme,

wherein the first portion of the output downmix representation is a first frequency band of a spectrum of the output downmix representation and wherein the second portion of the output downmix representation is a second frequency band of the spectrum of the output downmix representation, and

wherein the first portion of the output downmix representation and the second portion of the output downmix representation are based on the third downmixing scheme.

5. A method for generating an output downmix representation, comprising:

receiving an input downmix representation, wherein a first portion of the input downmix representation is a first frequency band of a spectrum and is in accordance with a first downmixing scheme and a second portion of the input downmix representation is a second frequency band of the spectrum and is in accordance with a second downmixing scheme, wherein the second downmixing scheme is different from the first downmixing scheme;

upmixing the first portion of the input downmix representation using a first upmixing scheme corresponding to the first downmixing scheme to obtain at least one first upmixed portion and upmixing the second portion of the input downmix representation using a second upmixing scheme corresponding to the second downmixing scheme to obtain at least one second upmixed portion; and

downmixing the at least one first upmixed portion in accordance with a third downmixing scheme to acquire a first portion of the output downmix representation and downmixing the at least one second upmixed portion in accordance with the third downmixing scheme to acquire a second portion of the output downmix representation,

wherein the third downmixing scheme is different from the first downmixing scheme and the second downmixing scheme,

wherein the first portion of the output downmix representation is a first frequency band of a spectrum of the output downmix representation and wherein the second portion of the output downmix representation is a second frequency band of the spectrum of the output downmix representation, and

wherein the first portion of the output downmix representation and the second portion of the output downmix representation are based on the third downmixing scheme.

6. A non-transitory digital storage medium having a computer program stored thereon to perform, when said computer program is run by a computer, a method of generating an output downmix representation, the method comprising:

receiving an input downmix representation, wherein a first portion of an input downmix representation is a first frequency band of a spectrum and is in accordance with a first downmixing scheme and a second portion of the input downmix representation is a second frequency band of the spectrum and is in accordance with a second downmixing scheme, wherein the second downmixing scheme is different from the first downmixing scheme;

upmixing the first portion of the input downmix representation using a first upmixing scheme corresponding to the first downmixing scheme to obtain at least one first upmixed portion and upmixing the second portion of the input downmix representation using a second upmixing scheme corresponding to the second downmixing scheme to obtain at least one second upmixed portion; and

downmixing the at least one first upmixed portion in accordance with a third downmixing scheme to acquire a first portion of the output downmix representation and downmixing the at least one second upmixed portion in accordance with the third downmixing scheme to acquire a second portion of the output downmix representation,

wherein the third downmixing scheme is different from the first downmixing scheme and the second downmixing scheme,

wherein the first portion of the output downmix representation is a first frequency band of a spectrum of the output downmix representation and wherein the second portion of the output downmix representation is a second frequency band of the spectrum of the output downmix representation, and

wherein the first portion of the output downmix representation and the second portion of the output downmix representation are based on the third downmixing scheme.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 3, 2022
From: REUTELHUBER, FRANZ; FOTOPOULOU, ELENI; MULTRUS, MARKUS
To: FRAUNHOFER-GESELLSCHAFT ZUR FÖRDERUNG DER ANGEWANDTEN FORSCHUNG E.V.
Reel/Frame 059158/0303 →
Priority Claims (2)
EP 19170621 · Apr 23, 2019 · regional
WO PCT/EP2019/070376 · Jul 29, 2019 · international
Continuity (2)
Continuation PCTEP2020061233 · Apr 22, 2020
Related Publication 20220036911A1 · Feb 3, 2022
References Cited (38)
US 6005948A · Maeda · 1999 [cited by applicant]
US 8515104B2 · Dickins et al. · 2013 [cited by applicant]
US 20060233379A1 · Villemoes · 2006 [cited by examiner]
US 20070140499A1 · Davis et al. · 2007 [cited by applicant]
US 20070194952A1 · Breebaart et al. · 2007 [cited by applicant]
US 20110170721A1 · Dickins · 2011 [cited by examiner]
US 20110224994A1 · Norvell et al. · 2011 [cited by applicant]
US 20120014526A1 · Stoll et al. · 2012 [cited by applicant]
US 20120263308A1 · Herre et al. · 2012 [cited by applicant]
US 20160157040A1 · Ertel · 2016 [cited by examiner]
US 20170309288A1 · Koppens · 2017 [cited by examiner]
US 20170365263A1 · Disch et al. · 2017 [cited by applicant]
US 20180124541A1 · Ertel et al. · 2018 [cited by applicant]
US 20180197552A1 · Fuchs · 2018 [cited by examiner]
US 20180293992A1 · Chebiyyam · 2018 [cited by examiner]
CN 105580391A · 2016 [cited by applicant]
CN 106796804A · 2017 [cited by applicant]
JP 2007531913A · 2007 [cited by applicant]
JP 2016527804A · 2016 [cited by applicant]
JP 2017017749A · 2017 [cited by applicant]
KR 20060121985A · 2006 [cited by applicant]
RU 2380766C2 · 2010 [cited by applicant]
TW 201103008A · 2011 [cited by applicant]
WO 2010097748A1 · 2010 [cited by applicant]
WO 2013186343A2 · 2013 [cited by applicant]
WO 2014161996A2 · 2014 [cited by applicant]
WO 2016050854A1 · 2016 [cited by applicant]
WO 2017125562A1 · 2017 [cited by applicant]
WO 2017125563A1 · 2017 [cited by applicant]
WO 2018086946A1 · 2018 [cited by applicant]
3GPP TS 26.445, Technical Specification Group Services and System Aspects; Codec for Enhanced Voice Services (EVS); Detailed Algorithmic Description. [cited by applicant]
A. Adami, E. Habets und J. Herre, “Down-mixing using coherence suppression,” in IEEE International Conference on Acoustics, Speech and Signal Processing, Florence, 2014. [cited by applicant]
F. Baumgarte, C. Faller und P. Kroon, “Audio Coder Enhancement using Scalable Binaural Cue Coding with Equalized Mixing,” in 116th Convention of the AES, Berlin, 2004. [cited by applicant]
ISO/IEC 23008-3:, Information technology—High efficiency coding and media delivery in heterogeneous environments—Part 3: 3D audio. [cited by applicant]
ITU-R BS.775-2, Multichannel Stereophonic Sound System With and Without Accompanying Picture, Jul. 2006. [cited by applicant]
Jeroen Breebaart, et al, “Background, Concept, and Architecture for the Recent MPEG Surround Standard on Multichannel Audio Compression”, AES, 60 East 42nd Street, Room 2520 New York 10165-2520, USA, May 1, 2007 (May 1,… [cited by applicant]
Lapierre Jimmy et al: “On Improving Parametric Stereo Audio Coding”, AES Convention 120; May 2006, AES, 60 East 42ND Street, Room 2520 New York 10165-2520, USA, May 1, 2006 (May 1, 2006), XP040507698. [cited by applicant]
M. Kim, E. Oh und H. Shim, “Stereo audio coding improved by phase parameters,” in 129th Convention of the AES, San Francisco, 2010. [cited by applicant]