IP Library Granted Patent US 12,567,430
Granted Patent B2
US 12,567,430 · App. 17/925,261 · Granted Mar 3, 2026

Method and device for improving dialogue intelligibility during playback of audio data

Inventors: Christian Schindler (Nuremberg, DE); Malte Schmidt (Feucht, DE)
Assignee: Dolby International AB
G10L21/0364G11B27/031H04N21/439
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,567,430
App. No.
17/925,261
Granted
Mar 3, 2026
Kind
B2
Abstract

Described herein is a method for improving dialogue intelligibility during playback of audio data on a playback device, wherein the audio data comprise dialogue audio data, and at least one of music and effects audio data, the method including the steps of: determining a volume mixing ratio based on a volume value for playback; mixing the dialogue audio data and the at least one of music and effects audio data based on said volume mixing ratio; and outputting the mixed audio data for playback. Described are further a respective playback device and a respective computer program product.

Claims (28)

1 . A method comprising:

determining a volume mixing ratio as a function of a sound pressure level based on a volume value for playback on a playback device by mapping the volume value for playback to the sound pressure level, wherein the volume mixing ratio refers to a ratio of the volume of the dialogue audio data over the volume of the at least one of music and effects audio data;

mixing the dialogue audio data and the at least one of music and effects audio data based on said volume mixing ratio; and

outputting the mixed audio data for playback.

2 . The method according to claim 1 , wherein mixing the dialogue audio data, and the at least one of music and effects audio data based on said volume mixing ratio includes applying a gain at least to the dialogue audio data.

3 . The method according to claim 1 , wherein the volume value for playback is based on a volume value setting of the playback device.

4 . The method according to claim 3 , wherein the volume value setting is a user-defined value.

5 . The method according to claim 1 , wherein the volume mixing ratio is further determined based on an ambient sound pressure level.

6 . The method according to claim 5 , wherein the ambient sound pressure level is determined based on a measurement by one or more microphones.

7 . The method according to claim 1 , wherein, prior to determining the volume mixing ratio, the method further includes:

receiving a bitstream including compressed audio data; and

core decoding, by a core decoder, the compressed audio data and providing the dialogue audio data, and the at least one of music and effects audio data.

8 . The method according to claim 7 , wherein the received bitstream further includes information on an audio content type, and wherein the volume mixing ratio is further determined based on said audio content type.

9 . The method according to claim 8 , wherein the audio content type includes one or more of audio content of a movie, audio content of a news program, audio content of a sports broadcast, and episodical audio content.

10 . The method according to claim 7 , wherein the method further includes analyzing the compressed audio data to provide the dialogue audio data, and the at least one of music and effects audio data.

11 . A media playback system, comprising including:

an audio processor for determining a volume mixing ratio as a function of a sound pressure level based on a volume value for playback by mapping the volume value for playback to the sound pressure level, wherein the volume mixing ratio refers to a ratio of the volume of the dialogue audio data over the volume of the at least one of music and effects audio data;

a mixer for mixing the dialogue audio data and the at least one of music and effects audio data based on said volume mixing ratio; and

a controller for outputting the mixed audio data for playback.

12 . The media playback system according to claim 11 , wherein the mixer is further configured to apply a gain at least to the dialogue audio data.

13 . The media playback system according to claim 11 , wherein the audio processor is configured to map the volume value for playback to a sound pressure level for determining the volume mixing ratio as a function of said sound pressure level.

14 . The media playback system according to claim 11 , wherein the playback device further includes a user interface for receiving a volume value setting of a user, and wherein the volume value for playback is based on said volume value setting.

15 . The media playback system according to claim 11 , wherein the playback device further includes one or more microphones for determining an ambient sound pressure level, and wherein the audio processor is configured to determine the volume mixing ratio further based on said ambient sound pressure level.

16 . The media playback system according to claim 11 , wherein the playback device further includes:

a receiver for receiving a bitstream including compressed audio data; and

a core decoder for core decoding the compressed audio data and for providing the dialogue audio data, and the at least one of music and effects audio data.

17 . The media playback system according to claim 16 , wherein the core decoder is further configured to analyze the compressed audio data to provide the dialogue audio data, and the at least one of music and effects audio data.

18 . A non-transitory, computer-readable medium storing instructions adapted to cause a device having processing capability to carry out the method according to claim 1 .

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 23, 2023
From: SCHINDLER, CHRISTIAN; SCHMIDT, MALTE
To: DOLBY INTERNATIONAL AB
Reel/Frame 063733/0919 →
Priority Claims (1)
EP 20174974 · May 15, 2020 · regional
Continuity (2)
Provisional Application 63025479 · May 15, 2020
Related Publication 20230238016A1 · Jul 27, 2023
References Cited (18)
US 9615170B2 · Kirsch · 2017 [cited by applicant]
US 20030125933A1 · Saunders · 2003 [cited by examiner]
US 20040213420A1 · Gundry · 2004 [cited by applicant]
US 20090245539A1 · Vaudrey · 2009 [cited by applicant]
US 20150237454A1 · Scheirer · 2015 [cited by applicant]
US 20160078879A1 · Lu · 2016 [cited by applicant]
US 20160315722A1 · Holman · 2016 [cited by examiner]
US 20190014435A1 · Baijal · 2019 [cited by examiner]
US 20200058317A1 · Gaalaas · 2020 [cited by examiner]
CN 1949795A · 2007 [cited by applicant]
CN 101518098A · 2009 [cited by applicant]
CN 103915103A · 2014 [cited by applicant]
CN 108337606A · 2018 [cited by applicant]
CN 108432130A · 2018 [cited by applicant]
JP 2012034295A · 2012 [cited by applicant]
WO 2008032209A2 · 2008 [cited by applicant]
Li Shengfei, Research on Digital Audio Processing Technology Based on FPGA+DSP Architecture, Microcontrollers & Embedded Systems, Jan. 1, 2018, 5 pages. [cited by applicant]
Ma Lin and Fu Rong, Common Problems in Broadcast Audio Mixing and How to Correct Them, Guangxi People's Broadcasting Station, Music Magazine, 181-182, 2 pages. [cited by applicant]