IP Library › Granted Patent US 9,646,622
Granted Patent B2
US 9,646,622 · App. 14/525,536 · Granted May 9, 2017

System and method for non-destructively normalizing loudness of audio signals within portable devices

Inventors: Jeffrey Riedmiller (Penngrove, CA); Harald Mundt (Fürth, DE); Michael Schug (Erlangen, DE); Martin Wolters (Nürnberg, DE)
Assignees: Dolby Laboratories Licensing Corporation; Dolby International AB
G10L19/02G10L19/167H03G3/32H03G7/007
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,646,622
App. No.
14/525,536
Granted
May 9, 2017
Kind
B2
Abstract

Many portable playback devices cannot decode and playback encoded audio content having wide bandwidth and wide dynamic range with consistent loudness and intelligibility unless the encoded audio content has been prepared specially for these devices. This problem can be overcome by including with the encoded content some metadata that specifies a suitable dynamic range compression profile by either absolute values or differential values relative to another known compression profile. A playback device may also adaptively apply gain and limiting to the playback audio. Implementations in encoders, in transcoders and in decoders are disclosed.

Claims (25)

1. A method for decoding an encoded input signal to generate an audio output signal, wherein the method comprises:

receiving the encoded input signal, wherein the encoded input signal includes encoded audio information and associated metadata including one or more decoding-control parameters, one or more first parameters specifying dynamic range compression suitable for use by a first mode of decoding that uses a first reference reproduction level, and one or more second parameters specifying dynamic range compression suitable for use by a second mode of decoding that uses a second reference reproduction level;

applying a decoding process to the encoded audio information to obtain subband signals, wherein the decoding process is adapted in response to the one or more decoding-control parameters;

modifying the subband signals to obtain modified subband signals with changed dynamic range characteristics, wherein the modifying is adapted in response to the one or more second parameters;

applying a synthesis filter bank to the modified subband signals to obtain a time-domain audio signal; and

applying a fixed gain and a limiter to the time-domain audio signal, wherein the application of the fixed gain raises an effective reference reproduction level of the audio output signal above the second reference reproduction level, and wherein the application of the limiter prevents amplitudes of the audio output signal from exceeding a clipping level.

2. The method of claim 1 , wherein the first reference reproduction level corresponds to an amplitude 31 dB below the clipping level, the second reference reproduction level corresponds to an amplitude 20 dB below the clipping level, and the effective reference reproduction level corresponds to an amplitude from 14 dB to 8 dB below clipping level.

3. The method of claim 2 , wherein the effective reference reproduction level corresponds to an amplitude 11 dB below clipping level.

4. An apparatus for decoding an encoded input signal to generate an audio output signal, wherein the apparatus comprises:

a receiver configured to receive the encoded input signal, wherein the encoded input signal includes encoded audio information and associated metadata including one or more decoding-control parameters, one or more first parameters specifying dynamic range compression suitable for use by a first mode of decoding that uses a first reference reproduction level, and one or more second parameters specifying dynamic range compression suitable for use by a second mode of decoding that uses a second reference reproduction level; and

a processor configured to:

apply a decoding process to the encoded audio information to obtain subband signals, wherein the decoding process is adapted in response to the one or more decoding-control parameters;

modify the subband signals to obtain modified subband signals with changed dynamic range characteristics, wherein the modifying is adapted in response to the one or more second parameters;

apply a synthesis filter bank to the modified subband signals to obtain a time-domain audio signal; and

apply a fixed gain and a limiter to the time-domain audio signal, wherein the application of the fixed gain raises an effective reference reproduction level of the audio output signal above the second reference reproduction level, and wherein the application of the limiter prevents amplitudes of the audio output signal from exceeding a clipping level.

5. The apparatus of claim 4 , wherein the first reference reproduction level corresponds to an amplitude 31 dB below the clipping level, the second reference reproduction level corresponds to an amplitude 20 dB below the clipping level, and the effective reference reproduction level corresponds to an amplitude from 14 dB to 8 dB below clipping level.

6. The apparatus of claim 5 , wherein the effective reference reproduction level corresponds to an amplitude 11 dB below clipping level.

7. A non-transitory medium recording a program of instructions that is executable by a device to perform a method for decoding an encoded input signal to generate an audio output signal, wherein the method comprises:

receiving the encoded input signal, wherein the encoded input signal includes encoded audio information and associated metadata including one or more decoding-control parameters, one or more first parameters specifying dynamic range compression suitable for use by a first mode of decoding that uses a first reference reproduction level, and one or more second parameters specifying dynamic range compression suitable for use by a second mode of decoding that uses a second reference reproduction level;

applying a decoding process to the encoded audio information to obtain subband signals, wherein the decoding process is adapted in response to the one or more decoding-control parameters;

modifying the subband signals to obtain modified subband signals with changed dynamic range characteristics, wherein the modifying is adapted in response to the one or more second parameters;

applying a synthesis filter bank to the modified subband signals to obtain a time-domain audio signal; and

applying a fixed gain and a limiter to the time-domain audio signal, wherein the application of the fixed gain raises an effective reference reproduction level of the audio output signal above the second reference reproduction level, and wherein the application of the limiter prevents amplitudes of the audio output signal from exceeding a clipping level.

8. The medium of claim 7 , wherein the first reference reproduction level corresponds to an amplitude 31 dB below the clipping level, the second reference reproduction level corresponds to an amplitude 20 dB below the clipping level, and the effective reference reproduction level corresponds to an amplitude from 14 dB to 8 dB below clipping level.

9. The medium of claim 8 , wherein the effective reference reproduction level corresponds to an amplitude 11 dB below clipping level.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 28, 2014
From: RIEDMILLER, JEFFREY; MUNDT, HARALD; SCHUG, MICHAEL; WOLTERS, MARTIN
To: DOLBY LABORATORIES LICENSING CORPORATION; DOLBY INTERNATIONAL AB
Reel/Frame 034054/0183 →
Continuity (3)
Continuation 13576386
Provisional Application 61303643 · Feb 11, 2010
Related Publication 20150043754A1 · Feb 12, 2015