IP Library Granted Patent US 12,439,129
Granted Patent B2
US 12,439,129 · App. 17/963,500 · Granted Oct 7, 2025

Systems and methods for determining whether to adjust volumes of individual audio components in a media asset based on a type of a segment of the media asset

Inventor: Alexis Yelton (Somerville, MA)
Assignee: Adeia Guides Inc.
H04N21/4852G06F3/165H04N21/4394H04N21/4398H04N21/4532H04N21/8456
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,439,129
App. No.
17/963,500
Granted
Oct 7, 2025
Kind
B2
Abstract

Systems and methods are provided herein for determining whether to adjust volumes of individual audio components in a media asset based on a type of segment of the media asset that is playing back. A media guidance application may determine that a user is playing back a segment of a media asset. The media guidance application may determine a type corresponding to the segment. The media guidance application may parse a plurality of audio components of the media asset that are playing back during the segment. The media guidance application may determine, for each audio component, whether to adjust the volume playing back during the segment based on the type. For each audio component of the plurality of audio components, in response to determining to adjust the volume, the media guidance application may adjust the volume of the audio component playing back during the segment.

Claims (88)

1. A computer-implemented method, comprising:

determining that a particular scene of a plurality of scenes of a video is being presented for display in a single physical viewing environment to a plurality of users, wherein each user of the plurality of users has a respective profile that stores scene viewing history data of a respective user of the plurality of users;

retrieving metadata corresponding to the particular scene from a database;

determining, based on the metadata, a type of the particular scene;

identifying, based on the type of the particular scene, a preferred user of the plurality of users, wherein the preferred user is associated with a preferred user profile that comprises scene viewing history data indicating the preferred user has a highest affinity for the type of the particular scene among the plurality of users;

parsing audio components of an audio track of the particular scene, wherein each of the audio components comprise a specific type of sound;

isolating the audio components from the audio track of the particular scene;

for each respective portion of a plurality of portions of the particular scene, determining, based on the type of the particular scene and volume parameters stored via the preferred user profile, whether to adjust volume of audio associated with the isolated audio components of each respective portion;

in response to determining, based on the type of the particular scene and the volume parameters stored via the preferred user profile, to adjust the volume of audio associated with the isolated audio components of at least one respective portion of the particular scene, adjusting the volume of audio associated with the isolated audio components of the at least one respective portion of the particular scene;

playing the at least one respective portion of the particular scene with the adjusted volume;

determining a type of a subsequent scene is different from the type of the particular scene;

identifying an updated preferred user of the plurality of users, wherein the updated preferred user is associated with an updated preferred user profile that comprises scene viewing history data indicating the updated preferred user has a highest affinity for the type of the subsequent scene among the plurality of users; and

playing the subsequent scene of the video with volume adjusted according to updated volume parameters stored via the updated preferred user profile.

2. The method of claim 1 , further comprising:

determining a respective category corresponding to the at least one respective portion of the particular scene;

wherein determining, for the at least one respective portion of the plurality of portions of the particular scene, based on the type of the particular scene, whether to adjust the volume of audio associated with the at least one respective portion is based at least in part on the category of the at least one respective portion.

3. The method of claim 1 , further comprising:

determining, based on the type of the particular scene and for each respective portion of the plurality of portions of the particular scene, whether to adjust volume of audio associated with each respective portion;

adjusting the volume of audio for each respective portion of the plurality of portions for which it is determined that adjustment is to be performed; and

declining to adjust the volume of audio for each respective portion of the plurality of portions for which it is determined that adjustment is not to be performed.

4. The method of claim 3 , further comprising:

determining a respective category corresponding to each respective portion of the plurality of portions of the particular scene;

determining, based on the type of the particular scene, whether to adjust volume of audio associated with each respective portion further comprises:

retrieving, from a database, for each respective category, volume parameters for the user profile that correspond to the type; and

determining, for each respective category, whether audio of the corresponding portion is set to a volume that is within the volume parameters.

5. The method of claim 4 , wherein determining a respective category corresponding to each respective portion of the plurality of portions comprises:

retrieving, from the database, a data structure, wherein the data structure comprises a plurality of categories;

comparing each respective portion with one or more categories in the data structure;

determining, from the comparison, whether a match between a respective portion and the one or more categories in the data structure; and

in response to determining that the match exists, determining, from the match, the respective category corresponding to the respective portion.

6. The method of claim 1 , further comprising:

generating for display an option regarding whether to perform adjustment of audio of the at least one respective portion of the particular scene based on preferences of the user profile; and

performing the adjusting of the volume of audio associated with the at least one respective portion of the particular scene further based on receiving selection of the option.

7. The method of claim 1 , wherein:

each scene of the plurality of scenes of the video corresponds to a respective type; and

the method further comprises adjusting volume of audio for at least one portion of each of the plurality of scenes based on the respective type of each scene of the plurality of scenes.

8. The method of claim 7 , wherein the plurality of scenes correspond to a plurality of genres, respectively.

9. The method of claim 1 , further comprising:

playing the adjusted volume of audio components at a personal hearing device of the preferred user.

10. A computer-implemented system, comprising:

input/output circuitry;

control circuitry configured to:

determine that a particular scene of a plurality of scenes of a video is being presented for display in a single physical viewing environment to a plurality of users, wherein each user of the plurality of users has a respective profile that stores scene viewing history data of a respective user of the plurality of users;

retrieve metadata corresponding to the particular scene from a database;

determine, based on the metadata, a type of the particular scene;

rein the preferred user is associated with a preferred user profile that comprises scene viewing history data indicating the preferred user has a highest affinity for the type of the particular scene among the plurality of users;

parsing audio components of an audio track of the particular scene, wherein each of the audio components comprise a specific type of sound;

isolating the audio components from the audio track of the particular scene;

for each respective portion of a plurality of portions of the particular scene, determine, based on the type of the particular scene and volume parameters stored via the preferred user profile, whether to adjust volume of audio associated with the isolated audio components of each respective portion; and

in response to determining, based on the type of the particular scene and the volume parameters stored via the preferred user profile, to adjust the volume of audio associated with the isolated audio components of at least one respective portion of the particular scene, adjust the volume of audio associated with the isolated audio components of the at least one respective portion of the particular scene;

the input/output circuitry is configured to:

play the at least one respective portion of the particular scene with the adjusted volume;

wherein the control circuitry is further configured to:

determine a type of a subsequent scene is different from the type of the particular scene;

identify an updated preferred user of the plurality of users, wherein the updated preferred user is associated with an updated preferred user profile that comprises scene viewing history data indicating the updated preferred user has a highest affinity for the type of the subsequent scene among the plurality of users; and

wherein the input/output circuitry is further configured to:

play the subsequent scene of the video with volume adjusted according to updated volume parameters stored via the updated preferred user profile.

11. The system of claim 10 , wherein the control circuitry is further configured to:

determine a respective category corresponding to the at least one respective portion of the particular scene;

wherein determining, for the at least one respective portion of the plurality of portions of the particular scene, based on the type of the particular scene, whether to adjust the volume of audio associated with the at least one respective portion is based at least in part on the category of the at least one respective portion.

12. The system of claim 10 , wherein the control circuitry is further configured to:

determine, based on the type of the particular scene and for each respective portion of the plurality of portions of the particular scene, whether to adjust volume of audio associated with each respective portion;

adjust the volume of audio for each respective portion of the plurality of portions for which it is determined that adjustment is to be performed; and

decline to adjust the volume of audio for each respective portion of the plurality of portions for which it is determined that adjustment is not to be performed.

13. The system of claim 12 , wherein the control circuitry is further configured to:

determine a respective category corresponding to each respective portion of the plurality of portions of the particular scene;

determine, based on the type of the particular scene, whether to adjust volume of audio associated with each respective portion by:

retrieving, from a database, for each respective category, volume parameters for the user profile that correspond to the type; and

determining, for each respective category, whether audio of the corresponding portion is set to a volume that is within the volume parameters.

14. The system of claim 13 , wherein the control circuitry is configured to determine a respective category corresponding to each respective portion of the plurality of portions by:

retrieving, from the database, a data structure, wherein the data structure comprises a plurality of categories;

comparing each respective portion with one or more categories in the data structure;

determining, from the comparison, whether a match between a respective portion and the one or more categories in the data structure; and

in response to determining that the match exists, determining, from the match, the respective category corresponding to the respective portion.

15. The system of claim 10 , wherein the control circuitry is further configured to:

generate for display an option regarding whether to perform adjustment of audio of the at least one respective portion of the particular scene based on preferences of the user profile; and

perform the adjusting of the volume of audio associated with the at least one respective portion of the particular scene further based on receiving selection of the option.

16. The system of claim 10 , wherein:

each scene of the plurality of scenes of the video corresponds to a respective type; and

the control circuitry is configured to adjust volume of audio for at least one portion of each of the plurality of scenes based on the respective type of each scene of the plurality of scenes.

17. The system of claim 16 , wherein the plurality of scenes correspond to a plurality of genres, respectively.

18. The system of claim 10 , wherein the control circuitry is further configured to:

cause the adjusted volume of audio components to be played at a personal hearing device of the preferred user.

19. The computer-implemented method of claim 1 , wherein:

the determining the type of the subsequent scene is different from the type of the particular scene further comprises retrieving metadata corresponding to the subsequent scene of the video; and

the determining is based on the retrieved metadata.

20. The computer-implemented system of claim 10 , wherein the control circuitry configured to

determine the type of the subsequent scene is different from the type of the particular scene is further configured to retrieve metadata corresponding to a subsequent scene of the video, wherein the determining is based on the retrieved metadata.

Assignments (3)
CHANGE OF NAME Recorded Oct 2, 2024
From: ROVI GUIDES, INC.
To: ADEIA GUIDES INC.
Reel/Frame 069086/0215 →
SECURITY INTEREST Recorded May 3, 2023
From: ADEIA GUIDES INC.; ADEIA IMAGING LLC; ADEIA MEDIA HOLDINGS LLC; ADEIA MEDIA SOLUTIONS INC.; ADEIA SEMICONDUCTOR ADVANCED TECHNOLOGIES INC.; ADEIA SEMICONDUCTOR BONDING TECHNOLOGIES INC.; ADEIA SEMICONDUCTOR INC.; ADEIA SEMICONDUCTOR SOLUTIONS LLC; ADEIA SEMICONDUCTOR TECHNOLOGIES LLC; ADEIA SOLUTIONS LLC
To: BANK OF AMERICA, N.A., AS COLLATERAL AGENT
Reel/Frame 063529/0272 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 11, 2022
From: YELTON, ALEXIS
To: ROVI GUIDES, INC.
Reel/Frame 061379/0170 →
Continuity (2)
Continuation 16644039
Related Publication 20230156290A1 · May 18, 2023
References Cited (22)
US 6239794B1 · Yuen et al. · 2001 [cited by applicant]
US 6359661B1 · Nickum · 2002 [cited by examiner]
US 6564378B1 · Satterfield et al. · 2003 [cited by applicant]
US 7165098B1 · Boyer et al. · 2007 [cited by applicant]
US 7761892B2 · Ellis et al. · 2010 [cited by applicant]
US 8046801B2 · Ellis et al. · 2011 [cited by applicant]
US 10085072B2 · Shimy · 2018 [cited by examiner]
US 20020174430A1 · Ellis et al. · 2002 [cited by applicant]
US 20050251827A1 · Ellis et al. · 2005 [cited by applicant]
US 20100017003A1 · Oh · 2010 [cited by examiner]
US 20100153885A1 · Yates · 2010 [cited by applicant]
US 20100322592A1 · Casagrande · 2010 [cited by applicant]
US 20130294755A1 · Arme et al. · 2013 [cited by applicant]
US 20130311575A1 · Woods et al. · 2013 [cited by applicant]
US 20140240595A1 · DiNunzio · 2014 [cited by examiner]
US 20150277850A1 · Wheatley · 2015 [cited by applicant]
US 20160364397A1 · Lindner et al. · 2016 [cited by applicant]
US 20170038700A1 · Takahashi et al. · 2017 [cited by applicant]
US 20170155369A1 · Wang · 2017 [cited by examiner]
US 20170188106A1 · Harb · 2017 [cited by applicant]
WO 2011087460A1 · 2011 [cited by applicant]
ID3 Draft Specification (2003) (41 pages). [cited by applicant]