IP Library › Granted Patent US 11,062,736
Granted Patent B2
US 11,062,736 · App. 16/854,062 · Granted Jul 13, 2021

Automated audio-video content generation

Inventor: Christophe Vaucher (Bandol, FR)
Assignee: SOCLIP!
G11B27/031G06F3/0482G06F3/04845
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,062,736
App. No.
16/854,062
Filed
Apr 21, 2020
Granted
Jul 13, 2021
Kind
B2
Art Unit
2481
USPC
386/280
Abstract

Systems and methods for generating media content include processing an audio file to determine one or more parameters of the audio file. Based on a skin associated with the audio file, one or more media effects are determined corresponding to the one or more parameters of the audio file. The one or more media effects are applied to a media file to generate a modified media file, wherein the media file excludes the audio file. An audio-video file is generated to include a combination of the audio file and the modified media file.

Claims (45)

1. A method of generating media content, the method comprising:

processing an audio file to determine one or more parameters of the audio file, wherein processing the audio file comprises:

determining one or more sections of the audio file, wherein each of the one or more sections comprises substantially uniform energy levels, with transitions in energy levels between adjacent sections corresponding to beats;

determining, one or more characteristics for each of the one or more sections, the one or more of characteristics comprising one or more of: occurrence of drum beats, types of drum beats, or distance between drum beats sequences; and

determining a section type unique number (STUN) for each of the one or more sections based on the one or more characteristics for each of the one or more sections, wherein the one or more parameters of the audio file are based on the STUN for each of the one or more sections of the audio file;

determining, based on a skin associated with the audio file, one or more media effects corresponding to the one or more parameters of the audio file;

applying the one or more media effects to a media file to generate a modified media file; and

generating an audio-video file comprising a combination of the audio file and the modified media file.

2. The method of claim 1 , further comprising:

obtaining the media file from one or more images, photographs, or videos, based on receiving a first user input from a first user interface.

3. The method of claim 2 , wherein the first user interface comprises a first spinner for displaying visual representations of the one or more images, photographs, or videos and receiving a selection of the one or more images, photographs, or videos corresponding to the first user input.

4. The method of claim 1 , further comprising:

obtaining the audio file from one or more music, song, or other audio data, based on receiving a second user input from a second user interface.

5. The method of claim 4 , wherein the second user interface comprises a second spinner for displaying visual representations of the one or more music, song, or other audio data and receiving a selection of the one or more music, song, or other audio data corresponding to the second user input.

6. The method of claim 1 , further comprising:

obtaining the skin from one or more skins comprising media effects, based on receiving a third user input from a third user interface.

7. The method of claim 6 , wherein the third user interface comprises a third spinner for displaying visual representations of the one or more skins and receiving a selection of the one or more skins corresponding to the third user input.

8. The method of claim 1 , wherein the one or more media effects comprise one or more of: edits comprising changing aspects of the media file to incorporate one or more of zoom, color translations, or brightness adjustments; transitions comprising one or more crossfade, or dissolve effects; or adjusting one or more levels of intensity, speed, or duration associated with displaying the media content.

9. A system, comprising:

one or more processors; and

a non-transitory computer-readable storage medium containing instructions which, when executed on the one or more processors, cause the one or more processors to perform operations for generating media content, the operations including:

processing an audio file to determine one or more parameters of the audio file; determining, based on a skin associated with the audio file, one or more media effects corresponding to the one or more parameters of the audio file; applying the one or more media effects to a media file to generate a modified media file, wherein processing the audio file further comprises:

determining one or more sections of the audio file, wherein each of the one or more sections comprises substantially uniform energy levels, with transitions in energy levels between adjacent sections corresponding to beats;

determining, one or more characteristics for each of the one or more sections, the one or more of characteristics comprising one or more of: occurrence of drum beats, types of drum beats, or distance between drum beats sequences; and

determining a section type unique number (STUN) for each of the one or more sections based on the one or more characteristics for each of the one or more sections, wherein the one or more parameters of the audio file are based on the STUN for each of the one or more sections of the audio file; and

generating an audio-video file comprising a combination of the audio file and the modified media file.

10. The system of claim 9 , wherein the operations further comprise:

obtaining the media file from one or more images, photographs, or videos, based on receiving a first user input from a first user interface.

11. The system of claim 10 , wherein the first user interface comprises a first spinner for displaying visual representations of the one or more images, photographs, or videos and receiving a selection of the one or more images, photographs, or videos corresponding to the first user input.

12. The system of claim 9 , wherein the operations further comprise:

obtaining the audio file from one or more music, song, or other audio data, based on receiving a second user input from a second user interface.

13. The system of claim 12 , wherein the second user interface comprises a second spinner for displaying visual representations of the one or more music, song, or other audio data and receiving a selection of the one or more music, song, or other audio data corresponding to the second user input.

14. The system of claim 9 , wherein the operations further comprise:

obtaining the skin from one or more skins comprising media effects, based on receiving a third user input from a third user interface.

15. The system of claim 14 , wherein the third user interface comprises a third spinner for displaying visual representations of the one or more skins and receiving a selection of the one or more skins corresponding to the third user input.

16. The system of claim 9 , wherein the one or more media effects comprise one or more of: edits comprising changing aspects of the media file to incorporate one or more of zoom, color translations, or brightness adjustments;

transitions comprising one or more crossfade, or dissolve effects; or adjusting one or more levels of intensity, speed, or duration associated with displaying the media content.

17. A non-transitory computer-readable medium having stored thereon instructions that, when executed by one or more processors, cause the one or more processors to:

process an audio file to determine one or more parameters of the audio file, wherein processing the audio file comprises:

determining one or more sections of the audio file, wherein each of the one or more sections comprises substantially uniform energy levels, with transitions in energy levels between adjacent sections corresponding to beats;

determining, one or more characteristics for each of the one or more sections, the one or more of characteristics comprising one or more of: occurrence of drum beats, types of drum beats, or distance between drum beats sequences; and

determining a section type unique number (STUN) for each of the one or more sections based on the one or more characteristics for each of the one or more sections, wherein the one or more parameters of the audio file are based on the STUN for each of the one or more sections of the audio file;

determine, based on a skin associated with the audio file, one or more media effects corresponding to the one or more parameters of the audio file; and

apply the one or more media effects to a media file to generate a modified media file; and generate an audio-video file comprising a combination of the audio file and the modified media file.

18. The non-transitory computer-readable medium of claim 17 , wherein the one or more media effects comprise one or more of: edits comprising changing aspects of the media file to incorporate one or more of zoom, color translations, or brightness adjustments; transitions comprising one or more crossfade, or dissolve effects; or adjusting one or more levels of intensity, speed, or duration associated with displaying the media content.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 21, 2020
From: VAUCHER, CHRISTOPHE
To: SOCLIP!
Reel/Frame 052452/0233 →
Continuity (2)
Provisional Application 62837122 · Apr 22, 2019
Related Publication 20200335133A1 · Oct 22, 2020