IP Library › Granted Patent US 10,362,426
Granted Patent B2
US 10,362,426 · App. 15/538,892 · Granted Jul 23, 2019

Upmixing of audio signals

Inventors: Jun Wang (Beijing, CN); Lie Lu (San Francisco, CA); Lianwu Chen (Beijing, CN); Mingqing Hu (Beijing, CN)
Assignee: Dolby Laboratories Licensing Corporation
H04S5/005H04R1/323H04S7/308H04S2400/11
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,362,426
App. No.
15/538,892
Granted
Jul 23, 2019
Kind
B2
Abstract

Example embodiments disclosed herein relates to upmixing of audio signals. A method of upmixing an audio signal is described. The method includes decomposing the audio signal into a diffuse signal and a direct signal, generating an audio bed at least in part based on the diffuse signal, the audio bed including a height channel, extracting an audio object from the direct signal, estimating metadata of the audio object, the metadata including height information of the audio object; and rendering the audio bed and the audio object as an upmixed audio signal, wherein the audio bed is rendered to a predefined position and the audio object is rendered according to the metadata. Corresponding system and computer program product are described as well.

Claims (38)

1. A method of upmixing an audio signal, comprising:

decomposing the audio signal into a diffuse signal and a direct signal;

determining complexity of the audio signal;

generating an audio bed at least in part based on the diffuse signal, the audio bed including a height channel;

extracting an audio object from the direct signal;

estimating metadata of the audio object, the metadata including height information of the audio object;

determining at least one of a diffuse gain for the audio signal based on the complexity, an object gain for the audio signal based on the complexity, or a height gain for the audio signal based on the complexity; and

rendering the audio bed and the audio object as an upmixed audio signal, wherein the audio bed is rendered to at least one predefined position and the audio object is rendered according to the metadata.

2. The method of claim 1 , wherein the generating the audio bed comprises:

upmixing the diffuse signal to create the height channel; and

including a residual signal into the audio bed, the residual signal obtained from the extracting of the audio object.

3. The method of claim 1 , wherein the decomposing the audio signal comprises:

determining a diffuse gain for the audio signal based on the complexity, the diffuse gain indicating a proportion of the diffuse signal in the audio signal; and

decomposing the audio signal based on the diffuse gain.

4. The method of claim 1 , wherein the extracting the audio object comprises:

determining an object gain for the audio signal based on the complexity, the object gain indicating a probability that the audio signal contains an audio object; and

extracting the audio object based on the object gain.

5. The method of claim 1 , wherein the extracting the metadata comprises:

determining a height gain for the audio object based on the complexity; and

modifying the height information of the audio object based on the height gain.

6. The method of claim 1 , wherein the rendering the audio object comprises:

determining an object-rendering gain based on at least one of the following:

the number of active channels of the audio object,

a position of the audio object with respect to a user, and

energy distribution among channels for the audio object; and

controlling, based on the object-rendering gain, a level of rendering related to the audio object in the rendering.

7. A method of upmixing an audio signal, comprising:

decomposing the audio signal into a diffuse signal and a direct signal;

generating an audio bed at least in part based on the diffuse signal, the audio bed including a height channel;

extracting an audio object from the direct signal;

estimating metadata of the audio object, the metadata including height information of the audio object and

rendering the audio bed and the audio object as an upmixed audio signal, wherein the audio bed is rendered to at least one predefined position and the audio object is rendered according to the metadata, wherein the decomposing the audio signal comprises:

applying a first decomposition process to obtain the diffuse signal; and

applying a second decomposition process to obtain the direct signal, the first decomposition process having less diffuse-to-direct leakage than the second decomposition process.

8. The method of claim 7 , further comprising:

pre-upmixing the audio signal,

wherein the first and second decomposition processes are separately applied to the pre-upmixed audio signal.

9. A computer program product of upmixing an audio signal, the computer program product being tangibly stored on a non-transient computer-readable medium and comprising machine executable instructions which, when executed, cause the machine to perform steps of the method according to claim 1 .

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 11, 2017
From: WANG, JUN; LU, LIE; CHEN, LIANWU; HU, MINGQING
To: DOLBY LABORATORIES LICENSING CORPORATION
Reel/Frame 043549/0457 →
Priority Claims (1)
CN 2015 1 0066647 · Feb 9, 2015 · national
Continuity (3)
Provisional Application 62117229 · Feb 17, 2015
Related Publication 20180262856A1 · Sep 13, 2018
Related Publication 20190052991A9 · Feb 14, 2019
Cited By (1)
US 12,581,266