IP Library › Granted Patent US 10,515,615
Granted Patent B2
US 10,515,615 · App. 15/753,393 · Granted Dec 24, 2019

Systems and methods for visual image audio composition based on user input

Inventor: Roy Elkins (Oregon, WI)
G10H1/0025G10H1/18G10H2210/105G10H2210/125G10H2220/091G10H2220/355G10H2220/441G10H2240/145
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,515,615
App. No.
15/753,393
Granted
Dec 24, 2019
Kind
B2
Abstract

The present invention relates to systems and methods for visual image audio composition. In particular, the present invention provides systems and methods for audio composition from a diversity of visual images and user determined sound database sources.

Claims (27)

1. A method for audio composition generation using at least one computer system, said computer system comprising a processor, a user interface, a visual image display screen, a sound database, an audio composition computer program on a computer readable medium configured to receive input from said visual image display screen and from said sound database to generate an audio composition, and an audio system, said method comprising:

a) displaying a visual image on said visual image display screen;

b) receiving a user's input relating to one or more audio parameters;

c) scanning said visual image to identify a plurality of visual image regions;

d) selecting and assembling a plurality of blocks in said sound database based on said visual image regions and said user's input of said one or more audio parameters; and

e) generating said audio composition based on said selecting and assembling said plurality of blocks in said sound database using said audio system,

wherein each of said plurality of blocks in said sound database is assigned a tag, and

wherein said selecting and assembling a plurality of blocks from said sound database based on said visual image regions and said user's input of said one or more audio parameters comprises at least one user-determined filter to pass and to not pass said tags of said blocks to said auditory composition.

2. The method of claim 1 , wherein said generating said audio composition based on said selecting and assembling said plurality of blocks in said sound database using said audio system is in real time.

3. The method of claim 1 , wherein said visual image on said visual image display screen comprises a digital photograph, a digital photograph selected from a digital photograph database, a digital photograph selected from a web-based database, a captured digital image, a digital video image, a film image, a visual hologram or a user-edited visual image.

4. The method of claim 1 , wherein said plurality of blocks in said sound database comprise a note comprising a duration, pitch, volume or tonal content, or a plurality of notes comprising melody, harmony, rhythm, tempo, voice, key signature, key change, intonation, temper, repetition, tracks, samples, loops, counterpoint, dissonance, sound effects, reverberation, delay, chorus, flange, dynamics, instrumentation, artist or artists sources, musical genre or style, monophonic and stereophonic reproduction, equalization, compression and mute/unmute blocks.

5. The method of claim 1 , wherein said plurality of blocks in said sound database comprise a recorded analog or digital sound, an analog or digital sound selected from a recording database, a digital sound selected from a web-based database, or a user-edited analog or digital sound.

6. The method of claim 1 , further comprising receiving a user's input relating to compatibility, alignment and transposition of said plurality of said passed blocks from said sound database to said audio composition.

7. The method of claim 1 , wherein said user interface is a screen interface, a keyboard interface, or a voice recognition interface.

8. The method of claim 1 , wherein said visual image display screen comprises at least one cursor configured to scan said visual image for input relating to said plurality of visual image regions of said displayed visual image.

9. The method of claim 8 , further comprising receiving said user's input relating to said plurality of said visual image regions.

10. The method of claim 9 , wherein said receiving said user's input relating to said plurality of said visual image regions of said displayed visual image comprises said user's input relating to the number, orientation, dimensions, rate of travel, direction of travel, steerability and image resolution of said at least one cursor.

11. The method of claim 9 , wherein said user's input relating to said plurality of said visual image regions of said displayed visual image comprises input relating to one or more pixel dimensions, coordinates, brightness, grey scale, or RGB scale, visual image regions comprising a plurality of pixels with user input relating to dimensions, color, tone, composition, content, feature resolution, a combination thereof, or a user-edited visual image region.

12. The method of claim 1 , further comprising receiving a user's input relating to said audio composition computer program on said computer readable medium.

13. The method of claim 1 , further comprising receiving a user's input relating to generating said audio composition using said audio system.

14. The method of claim 1 , wherein said audio composition is stored on a computer readable medium, stored in said sound database, stored on a web-based medium, edited, privately shared, publicly shared, available for sale, licensed, or converted to a visual image.

15. The method of claim 1 , wherein said audio composition computer program on a computer readable medium configured to receive input from said visual image display screen and from said sound database to generate an audio composition is downloadable to a mobile device, a phone, a tablet, a computer, a device configured to receive Mp3 files, or other digital source.

16. The method of claim 1 , further comprising receiving auditory input in real time.

17. The method of claim 1 , further comprising receiving GPS coordinates to filter said blocks of said sound database specific to a location, or to the distance and rate of travel between a plurality of locations.

18. The method of claim 1 , wherein said visual image on said visual image display screen is a visual image text.

19. A system comprising a computer configured to practice the method of claim 1 .

20. The system of claim 19 , further comprising software that directs said computer to practice the method of claim 1 .

Continuity (2)
Provisional Application 62207805 · Aug 20, 2015
Related Publication 20180247624A1 · Aug 30, 2018