IP Library › Granted Patent US 11,470,437
Granted Patent B2
US 11,470,437 · App. 16/825,776 · Granted Oct 11, 2022

Processing object-based audio signals

Inventors: Alan J. Seefeldt (Alameda, CA); Lie Lu (Dublin, CA); Chen Zhang (Beijing, CN)
Assignee: Dolby Laboratories Licensing Corporation
H04S7/302H04S3/008H04S7/30G10L19/008H04S2400/11
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,470,437
App. No.
16/825,776
Granted
Oct 11, 2022
Kind
B2
Abstract

An audio processing system and method which calculates, based on spatial metadata of the audio object, a panning coefficient for each of the audio objects in relation to each of a plurality of predefined channel coverage zones. Converts the audio signal into submixes in relation to the predefined channel coverage zones based on the calculated panning coefficients and the audio objects. Each of the submixes indicating a sum of components of the plurality of the audio objects in relation to one of the predefined channel coverage zones. Generating a submix gain by applying an audio processing to each of the submix and controls an object gain applied to each of the audio objects. The object gain being as a function of the panning coefficients for each of the audio objects and the submix gains in relation to each of the predefined channel coverage zones.

Claims (23)

1. A method of processing an audio signal, the audio signal having a plurality of audio objects, the method comprising: receiving spatial metadata corresponding to the plurality of audio objects; converting the audio signal into at least one submix corresponding to a subset of the plurality of audio objects, wherein the at least one submix includes rendering constraints regarding locations of the subset of the plurality of audio objects; determining a corresponding submix gain for the at least one submix; and rendering the at least one submix based on the rendering constraints, the spatial metadata, and the submix gain corresponding to the submix of the corresponding audio objects.

2. The method according to claim 1 , further comprising determining a weighted average of the plurality of audio objects for the at least one submix.

3. The method according to claim 2 , further comprising determining a weight corresponding to the at least one submix based on the weighted average, wherein the weight relates to a panning coefficient for each of the corresponding audio objects of the at least one submix.

4. The method according to claim 1 , wherein converting the audio signal into the at least one submix further comprises:

converting the audio signal into a front submix in relation to a front zone based on the panning coefficients for the audio objects;

converting the audio signal into a center submix in relation to a center zone based on the panning coefficients for the audio objects;

converting the audio signal into a surround submix in relation to a surround zone based on the panning coefficients for the audio objects; and

converting the audio signal into a height submix in relation to a height zone based on the panning coefficients for the audio objects.

5. The method according to claim 1 , further comprising:

for each of the audio objects, identifying a type of the audio object; and

generating the submix gain by applying an audio processing to the at least one submix based on the identified type of the audio object.

6. A computer program product for rendering an audio signal, the computer program product being tangibly stored on a non-transient computer-readable medium and comprising machine executable instructions which, when executed, cause the machine to perform steps of the method according to claim 1 .

7. A system for processing an audio signal, the audio signal having a plurality of audio objects, the system comprising: a receiver for receiving spatial metadata corresponding to the plurality of audio objects; a converter for converting the audio signal into at least one submix corresponding to a subset of the plurality of audio objects, wherein the at least one submix includes rendering constraints regarding locations of the subset of the plurality of audio objects; a processor for determining a corresponding submix gain for the at least one submix; and a renderer for rendering the at least one submix based on the rendering constraints, the spatial metadata, and the submix gain corresponding to the submix of the corresponding audio objects.

8. The system according to claim 7 , wherein the processor is further configured to determine a weighted average of the plurality of audio objects for the at least one submix.

9. The system according to claim 8 , wherein the processor is further configured to determine a weight corresponding to the at least one submix based on the weighted average, wherein the weight relates to a panning coefficient for each of the corresponding audio objects of the at least one submix.

10. The system according to claim 7 , wherein the converter is further configured to convert the audio signal into submixes by:

converting the audio signal into a front submix in relation to a front zone based on the panning coefficients for the audio objects;

converting the audio signal into a center submix in relation to a center zone based on the panning coefficients for the audio objects;

converting the audio signal into a surround submix in relation to a surround zone based on the panning coefficients for the audio objects; and

converting the audio signal into a height submix in relation to a height zone based on the panning coefficients for the audio objects.

11. The system according to claim 7 , wherein the processor is further configured to:

for each of the audio objects, identify a type of the audio object; and

generate the submix gain by applying an audio processing to the at least one submix based on the identified type of the audio object.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 17, 2020
From: SEEFELDT, ALAN J.; LU, LIE; ZHANG, CHEN
To: DOLBY LABORATORIES LICENSING CORPORATION
Reel/Frame 053802/0750 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 31, 2020
From: SEEFELDT, ALAN J.; LU, LIE; ZHANG, CHEN
To: DOLBY LABORATORIES LICENSING CORPORATION
Reel/Frame 052795/0392 →
Priority Claims (1)
CN 201510294063.7 · Jun 1, 2015 · national
Continuity (5)
Division 16368574 · Mar 28, 2019
Division 16143351 · Sep 26, 2018
Division 15577510
Provisional Application 62183491 · Jun 23, 2015
Related Publication 20200288260A1 · Sep 10, 2020