IP Library Granted Patent US 12682915
Granted Patent B2
US 12682915 · App. 18/622,993 · Granted Jul 14, 2026

System and method for generating brand standards from voice input

Inventors: Dimitrios Spirou (Oakville, CA); David Ormonde (Oakville, CA); Christina Spirou (Oakville, CA)
G10L25/18G06T11/10G10L15/1815
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12682915
App. No.
18/622,993
Granted
Jul 14, 2026
Kind
B2
Abstract

A system and method for identifying key notes from voice input, generating color representations and cymatic shape representations. The system includes an electronic device with a computational engine implemented using Python code. The method involves processing an audio file containing voice input, to identify the key notes that prevail across noted sentiments and converting the key notes into their representative frequencies. The resulting frequencies are further processed to generate representative colors, and representative cymatic shapes. Together, these generated representations serve as foundational design elements referred to as Brand Standards, for the purposes of branding, promotion, and trademarking.

Claims (40)

1 . A method for generating foundational design elements, or Brand Standards, from voice input, comprising:

processing an audio file containing original voice input, to identify the key notes, expressed as root notes;

wherein the audio file comprises multiple key notes;

processing each key note to compute a representative frequency in hertz;

processing each representative frequency to compute a wavelength in nanometers, corresponding to each identified key note;

processing each wavelength corresponding to each identified key note to generate representative color values in RBG, CMYK and HEX, for each identified key note;

processing each frequency representing an identified key note, to generate a cymatic shape, vectorized in SVG, WMF, EPS, PDF, CDR, and AI formats;

processing each vectorized cymatic shape, generating independent, vectorized components that make up the corresponding cymatic shape, in SVG, WMF, EPS, PDF, CDR, and AI formats;

generating a Brand Standard package including the key notes identified from voice input, color values associated with the key notes, and vectorized and componentized shapes associated with the key notes.

2 . The method of claim 1 , wherein the audio file is a .wav file.

3 . The method of claim 2 , wherein the frequency is computed via Python Script.

4 . The method of claim 3 , wherein the wavelength is computed via Python Script.

5 . The method of claim 4 , wherein the color values are computed via Python Script.

6 . The method of claim 5 , wherein the cymatic shapes are generated via Python Script.

7 . The method of claim 6 , wherein the cymatic shapes componentization is executed via Python Script.

8 . The method of claim 1 , further comprising:

identification of sentiments within the voice input, and further identifying their corresponding key notes, via analysis of an audio file containing voice input.

9 . A system for generating foundational design elements, or Brand Standards, from voice input, the system comprising:

an electronic device having a processor and a memory, wherein the electronic device is adapted to read an audio file from an input device;

wherein the audio file contains original voice input;

a computer program having logic, that when executed by the processor causes the electronic device to:

process the audio file to identify key notes expressed as root notes;

wherein the audio file comprises one or more key notes;

process each key note to compute a representative frequency for each key note in hertz;

process each representative frequency to compute a wavelength in nanometers for each key note;

process each wavelength corresponding to each identified key note to generate representative color values in RBG, CMYK and HEX, for each identified key note;

process each frequency representing an identified key note, to generate a cymatic shape, vectorized in SVG, WMF, EPS, PDF, CDR, and AI formats;

process each vectorized cymatic shape, generating independent, vectorized components that make up the corresponding cymatic shape, in SVG, WMF, EPS, PDF, CDR, and AI formats;

generating a Brand Standard package including the key notes identified from voice input, color values associated with the key notes, and vectorized and componentized shapes associated with the key notes.

10 . The system of claim 9 , wherein the input device is selected from the following: DVDs, CDs, hard disk drives, magnetic tape, cloud storage, digital audio files, and servers for streaming media over networks.

11 . The system of claim 9 , wherein a sample comprising of the voice input and the generated components derived from the voice, including the root notes, colors, and vectorized shapes, is stored on non-transitory medium.

12 . The system of claim 9 , wherein the electronic device is adapted to correspond to any remote electronic device that is configured to communicate with one or more other wireless communication devices over a wireless communications system.

13 . The system of claim 9 , wherein the audio file is a “.wave” file.

14 . The system of claim 13 , wherein the frequency is computed via Python Script.

15 . The system of claim 14 , wherein the wavelength is computed via Python Script.

16 . The system of claim 15 , wherein the color values are computed via Python Script.

17 . The method of claim 16 , wherein the cymatic shapes are generated via Python Script.

18 . The method of claim 17 , wherein the cymatic shapes componentization is executed via Python Script.

19 . The system of claim 1 , wherein the computer program is further adapted to:

identify of sentiments within the voice input, and further identifying their corresponding key notes, via analysis of an audio file containing voice input.