IP Library Granted Patent US 12,167,075
Granted Patent B2
US 12,167,075 · App. 17/850,470 · Granted Dec 10, 2024

Methods for conforming audio and short-form video

Inventors: Serhad Doken (Bryn Mawr, PA); V. Michael Bove, Jr. (Wrentham, MA)
Assignee: Adeia Guides Inc.
H04N21/4394G06F16/683H04N21/8456
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,167,075
App. No.
17/850,470
Granted
Dec 10, 2024
Kind
B2
Abstract

Systems and methods are provided herein for conforming audio to a video to avoid discordance. This may be accomplished by a system receiving a video and selection of an audio asset. The system may identify a plurality of break points in the audio asset based on one or more characteristics of the audio asset. A first portion of the audio asset may be generated based on one or more characteristics of the received video (e.g., length of the video), wherein the first portion of the audio asset begins and/or ends at a break point of the plurality of break points. The system may then generate a media item comprising the video and the first portion of the audio asset.

Claims (65)

1. A method comprising:

receiving a video comprising a plurality of segments;

determining a first type associated with a first segment of the plurality of segments;

determining a second type associated with a second segment of the plurality of segments, wherein the second type is different than the first type;

determining a length of the first segment;

determining a length of the second segment;

receiving a selection of an audio asset;

determining a plurality of break points in the audio asset based on a characteristic of the audio asset;

generating a first portion of the audio asset based on the length of the first segment and the first type associated with the first segment, wherein the first portion of the audio asset spans the length of the first segment;

generating a second portion of the audio asset based on the length of the second segment and the second type associated with the second segment, wherein the second portion of the audio asset ends at a break point of the plurality of break points; and

generating a media item comprising the first segment of the plurality of segments, the second segment of the plurality of segments, the first portion of the audio asset, and the second portion of the audio asset.

2. The method of claim 1 , wherein the characteristic corresponds to a decrease in an amplitude of sound of the audio asset.

3. The method of claim 1 , wherein the characteristic corresponds to an end of lyric.

4. The method of claim 1 , wherein the characteristic corresponds to an end of a chord pattern.

5. The method of claim 1 , wherein the characteristic corresponds to an end of a harmonic progression.

6. The method of claim 1 , wherein the audio comprises metadata corresponding the characteristic.

7. The method of claim 1 , further comprising:

generating a third portion of the audio asset based on the length of the first segment and the length of second segment, wherein the third portion of the audio asset ends at a second break point of the plurality of break points; and

generating a second media item comprising the first segment, the second segment, and the third portion of audio.

8. The method of claim 7 , further comprising:

ranking the first media item based on an attribute of the first media item;

ranking the second media item based on an attribute of the second media item; and

displaying the first media item and the second media item according to the ranking of the first media item and the ranking of the second media item.

9. The method of claim 1 , further comprising:

receiving an indication that the media item will be looped; and

determining a subset of the plurality of break points based on the indication.

10. An apparatus, comprising:

control circuitry; and

at least one memory including computer program code for one or more programs, the at least one memory and the computer program code configured to, with the control circuitry, cause the apparatus to perform at least the following:

receiving a video comprising a plurality of segments;

determine a first type associated with a first segment of the plurality of segments;

determine a second type associated with a second segment of the plurality of segments, wherein the second type is different than the first type;

determine a length of the first segment;

determine a length of the second segment;

receive a selection of an audio asset;

determine a plurality of break points in the audio asset based on a characteristic of the audio asset;

generate a first portion of the audio asset based on the length of the first segment and the first type associated with the first segment, wherein the first portion of the audio asset spans the length of the first segment;

generate a second portion of the audio asset based on the length of the second segment and the second type associated with the second segment, wherein the second portion of the audio asset ends at a break point of the plurality of break points; and

generate a media item comprising the first segment of the plurality of segments, the second segment of the plurality of segments, the first portion of the audio asset, and the second portion of the audio asset.

11. The apparatus of claim 10 , wherein the characteristic corresponds to a decrease in an amplitude of sound of the audio asset.

12. The apparatus of claim 10 , wherein the characteristic corresponds to an end of lyric.

13. The apparatus of claim 10 , wherein the characteristic corresponds to an end of a chord pattern.

14. The apparatus of claim 10 , wherein the characteristic corresponds to an end of a harmonic progression.

15. The apparatus of claim 10 , wherein the audio asset comprises metadata corresponding the characteristic.

16. The apparatus of claim 10 , wherein the apparatus is further caused to:

generate a third portion of the audio asset based on the length of the first segment and the length of second segment, wherein the third portion of the audio asset ends at a second break point of the plurality of break points; and

generate a second media item comprising the first segment, the second segment, and the third portion of audio.

17. The method of claim 16 , further comprising:

ranking the first media item based on an attribute of the first media item;

ranking the second media item based on an attribute of the second media item; and

displaying the first media item and the second media item according to the ranking of the first media item and the ranking of the second media item.

18. The apparatus of claim 10 , wherein the apparatus is further caused to:

receive an indication that the media item will be looped; and

determine a subset of the plurality of break points based on the indication.

19. A non-transitory computer-readable medium having instructions encoded thereon that, when executed by control circuitry, cause the control circuitry to:

receiving a video comprising a plurality of segments;

determine a first type associated with a first segment of the plurality of segments;

determine a second type associated with a second segment of the plurality of segments, wherein the second type is different than the first type;

determine a length of the first segment;

determine a length of the second segment;

receive a selection of an audio asset;

determine a plurality of break points in the audio asset based on a characteristic of the audio asset;

generate a first portion of the audio asset based on the length of the first segment and the first type associated with the first segment, wherein the first portion of the audio asset spans the length of the first segment;

generate a second portion of the audio asset based on the length of the second segment and the second type associated with the second segment, wherein the second portion of the audio asset ends at a break point of the plurality of break points; and

generate a media item comprising the first segment of the plurality of segments, the second segment of the plurality of segments, the first portion of the audio asset, and the second portion of the audio asset.

Assignments (3)
CHANGE OF NAME Recorded Oct 4, 2024
From: ROVI GUIDES, INC.
To: ADEIA GUIDES INC.
Reel/Frame 069113/0413 →
SECURITY INTEREST Recorded May 3, 2023
From: ADEIA GUIDES INC.; ADEIA IMAGING LLC; ADEIA MEDIA HOLDINGS LLC; ADEIA MEDIA SOLUTIONS INC.; ADEIA SEMICONDUCTOR ADVANCED TECHNOLOGIES INC.; ADEIA SEMICONDUCTOR BONDING TECHNOLOGIES INC.; ADEIA SEMICONDUCTOR INC.; ADEIA SEMICONDUCTOR SOLUTIONS LLC; ADEIA SEMICONDUCTOR TECHNOLOGIES LLC; ADEIA SOLUTIONS LLC
To: BANK OF AMERICA, N.A., AS COLLATERAL AGENT
Reel/Frame 063529/0272 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 10, 2023
From: DOKEN, SERHAD; BOVE JR., V. MICHEAL
To: ROVI GUIDES, INC.
Reel/Frame 062328/0813 →