IP Library › Granted Patent US 11,004,434
Granted Patent B2
US 11,004,434 · App. 16/725,736 · Granted May 11, 2021

Systems and methods for visual image audio composition based on user input

Inventor: Roy Elkins (Oregon, WI)
G10H1/0025G10H1/18G10H2210/105G10H2210/125G10H2220/091G10H2220/355G10H2220/441G10H2240/145
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,004,434
App. No.
16/725,736
Filed
Dec 23, 2019
Granted
May 11, 2021
Kind
B2
Art Unit
2837
USPC
84/615
Abstract

The present invention relates to systems and methods for visual image audio composition. In particular, the present invention provides systems and methods for audio composition from a diversity of visual images and user determined sound database sources.

Claims (17)

1. A system for generating an audio composition, the system comprising:

a processor configured to:

access a visual image;

receive a user input including at least one audio parameter;

assign a plurality of instruments to the visual image, wherein each of the plurality of instruments is associated with one of a plurality of regions in the visual image;

select and assemble, for each of the plurality of instruments, a plurality of notes from a sound database based on the one of the plurality of regions in the visual image associated with each instrument and the user input; and

generate the audio composition by mixing the plurality of notes selected and assembled for each of the plurality of instruments.

2. The system of claim 1 , wherein the user input is a first user input and wherein the processor is further configured to add an additional instrument to the plurality of instruments or remove one of the plurality of instruments in response to second user input.

3. The system of claim 1 , wherein the user input is a first user input and wherein the processor is further configured to generate the plurality of regions in the visual image based on a second user input.

4. The system of claim 1 , wherein the processor is further configured to apply at least one user-defined filter to at least one of the plurality of notes.

5. The system of claim 1 , wherein each of the plurality of notes in the sound database is assigned a tag.

6. The system of claim 5 , wherein the plurality of notes is selected from the sound database based on the tag assigned to each of the plurality of notes and the at least one audio parameter.

7. The system of claim 6 , wherein the at least one audio parameter includes a song format selected from a plurality of song formats generated based on the visual image.

8. The system of claim 1 , wherein the at least one audio parameter includes at least one selected from a group consisting of a genre, a method of generating the audio composition, and feedback on compatibility of a previously-generated audio composition.

9. The system of claim 1 , wherein the processor is further configured to receive a geographical coordinate associated with a user and generate the audio composition based on the geographical coordinate.

10. The system of claim 1 , wherein the processor is configured to generate the audio composition in real-time as the visual image changes.

11. The system of claim 1 , wherein the processor is configured to access the visual image by converting text to the visual image.

Continuity (3)
Continuation 15753393
Provisional Application 62204805 · Aug 20, 2015
Related Publication 20200265817A1 · Aug 20, 2020
Cited By (1)
US 12,368,820