IP Library › Granted Patent US 11,450,330
Granted Patent B2
US 11,450,330 · App. 16/842,212 · Granted Sep 20, 2022

Parametric reconstruction of audio signals

Inventors: Lars Villemoes (Järfälla, SE); Heidi-Maria Lehtonen (Upplands Väsby, SE); Heiko Purnhagen (Sundbyberg, SE); Toni Hirvonen (Helsinki, FI)
Assignee: DOLBY INTERNATIONAL AB
G10L19/167G10L19/008G10L19/265H04S5/005H04S2420/03
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,450,330
App. No.
16/842,212
Granted
Sep 20, 2022
Kind
B2
Abstract

An encoding system encodes an N-channel audio signal (X), wherein N≥3, as a single-channel downmix signal (Y) together with dry and wet upmix parameters ({tilde over (C)}, {tilde over (P)}). In a decoding system, a decorrelating section outputs, based on the downmix signal, an (N−1)-channel decorrelated signal (Z); a dry upmix section maps the downmix signal linearly in accordance with dry upmix coefficients (C) determined based on the dry upmix parameters; a wet upmix section populates an intermediate matrix based on the wet upmix parameters and knowing that the intermediate matrix belongs to a predefined matrix class, obtains wet upmix coefficients (P) by multiplying the intermediate matrix by a predefined matrix, and maps the decorrelated signal linearly in accordance with the wet upmix coefficients; and a combining section combines outputs from the upmix sections to obtain a reconstructed signal ({circumflex over (X)}) corresponding to the signal to be reconstructed.

Claims (39)

1. A method of reconstructing an N-channel audio signal (X) based on a single-channel downmix signal (Y), the method comprising:

receiving, by a decorrelating section of a parametric reconstruction system, the single-channel downmix signal (Y);

processing the single-channel downmix signal (Y) to output an (N−1)-channel decorrelated signal (Z), the processing including applying respective filters to the single-channel downmix signal (Y);

receiving, by a dry upmix section of the parametric reconstruction system, the single-channel downmix signal (Y) and dry upmix parameters ({tilde over (C)}), the dry upmix parameters ({tilde over (C)}) coinciding with a first portion of a set of dry upmix coefficients (C);

determining a remaining dry upmix coefficient of the set of dry upmix coefficients (C) based on a predefined relation between the set of dry upmix coefficients (C);

outputting, by the dry upmix section, a dry upmix signal (CY) computed by mapping the single-channel downmix signal (Y) linearly in accordance with the set of dry upmix coefficients (C);

receiving, by a wet upmix section of the parametric reconstruction system, the (N−1)-channel decorrelated signal (Z) and a set of wet upmix parameters ({tilde over (P)});

deriving, from the set of wet upmix parameters ({tilde over (P)}), a set of wet upmix coefficients (P);

outputting, by the wet upmix section, a wet upmix signal (PZ) computed by mapping the (N−1)-channel decorrelated signal (Z) in accordance with the set of wet upmix coefficients (P); and

combining, by a combining section of the parametric reconstruction system, the dry upmix signal (CY) and the wet upmix signal (PZ) to obtain a multidimensional reconstructed signal ({circumflex over (X)}) corresponding to the N-channel audio signal (X) to be reconstructed,

wherein at least one channel of the N-channel audio signal is associated with a predefined spatial orientation, and wherein the parametric reconstruction system includes one or more processors.

2. The method of claim 1 , comprising:

populating an intermediate matrix having more elements than the number of received wet upmix parameters, based on the received wet upmix parameters, the intermediate matrix belonging to a predefined matrix class.

3. An audio decoding system comprising:

one or more processors; and

a non-transitory computer-readable medium storing instructions that, upon execution by the one or more processors, cause the one or more processors to perform operations of reconstructing an N-channel audio signal (X) based on a single-channel downmix signal (Y), the operations comprising:

receiving, by a decorrelating section of a parametric reconstruction system, the single-channel downmix signal (Y);

processing the single-channel downmix signal (Y) to output an (N−1)-channel decorrelated signal (Z), the processing including applying respective filters to the single-channel downmix signal (Y);

receiving, by a dry upmix section of the parametric reconstruction system, the single-channel downmix signal (Y) and dry upmix parameters ({tilde over (C)}), the dry upmix parameters ({tilde over (C)}) coinciding with a first portion of a set of dry upmix coefficients (C);

determining a remaining dry upmix coefficient of the set of dry upmix coefficients (C) based on a predefined relation between the set of dry upmix coefficients (C);

outputting, by the dry upmix section, a dry upmix signal (CY) computed by mapping the single-channel downmix signal (Y) linearly in accordance with the set of dry upmix coefficients (C);

receiving, by a wet upmix section of the parametric reconstruction system, the (N−1)-channel decorrelated signal (Z) and a set of wet upmix parameters ({tilde over (P)});

deriving, from the set of wet upmix parameters ({tilde over (P)}), a set of wet upmix coefficients (P);

outputting, by the wet upmix section, a wet upmix signal (PZ) computed by mapping the (N−1)-channel decorrelated signal (Z) in accordance with the set of wet upmix coefficients (P);

combining, by a combining section of the parametric reconstruction system, the dry upmix signal (CY) and the wet upmix signal (PZ) to obtain a multidimensional reconstructed signal ({circumflex over (X)}) corresponding to the N-channel audio signal (X) to be reconstructed,

wherein at least one channel of the N-channel audio signal is associated with a predefined spatial orientation.

4. The system of claim 3 , the operations comprising:

populating an intermediate matrix having more elements than the number of received wet upmix parameters, based on the received wet upmix parameters, the intermediate matrix belonging to a predefined matrix class.

5. A non-transitory computer-readable medium storing instructions that, when executed by one or more processors, cause the one or more processors to perform operations of reconstructing an N-channel audio signal (X) based on a single-channel downmix signal (Y), the operations comprising:

receiving, by a decorrelating section of a parametric reconstruction system, the single-channel downmix signal (Y);

processing the single-channel downmix signal (Y) to output an (N−1)-channel decorrelated signal (Z), the processing including applying respective filters to the single-channel downmix signal (Y);

receiving, by a dry upmix section of the parametric reconstruction system, the single-channel downmix signal (Y) and dry upmix parameters ({tilde over (C)}), the dry upmix parameters ({tilde over (C)}) coinciding with a first portion of a set of dry upmix coefficients (C);

determining a remaining dry upmix coefficient of the set of dry upmix coefficients (C) based on a predefined relation between the set of dry upmix coefficients (C);

outputting, by the dry upmix section, a dry upmix signal (CY) computed by mapping the single-channel downmix signal (Y) linearly in accordance with the set of dry upmix coefficients (C);

receiving, by a wet upmix section of the parametric reconstruction system, the (N−1)-channel decorrelated signal (Z) and a set of wet upmix parameters ({tilde over (P)});

deriving, from the set of wet upmix parameters ({tilde over (P)}), a set of wet upmix coefficients (P);

outputting, by the wet upmix section, a wet upmix signal (PZ) computed by mapping the (N−1)-channel decorrelated signal (Z) in accordance with the set of wet upmix coefficients (P); and

combining, by a combining section of the parametric reconstruction system, the dry upmix signal (CY) and the wet upmix signal (PZ) to obtain a multidimensional reconstructed signal ({circumflex over (X)}) corresponding to the N-channel audio signal (X) to be reconstructed,

wherein at least one channel of the N-channel audio signal is associated with a predefined spatial orientation.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 9, 2020
From: VILLEMOES, LARS; LEHTONEN, HEIDI-MARIA; PURNHAGEN, HEIKO; HIRVONEN, TONI
To: DOLBY INTERNATIONAL AB
Reel/Frame 052352/0894 →
Continuity (7)
Continuation 16363099 · Mar 25, 2019
Continuation 15985635 · May 21, 2018
Division 15031130
Provisional Application 62037693 · Aug 15, 2014
Provisional Application 61974544 · Apr 3, 2014
Provisional Application 61893770 · Oct 21, 2013
Related Publication 20200302943A1 · Sep 24, 2020