IP Library Granted Patent US 11,264,025
Granted Patent B2
US 11,264,025 · App. 16/519,628 · Granted Mar 1, 2022

Automated graphical user interface control methods and systems using voice commands

Inventors: Joseph Kessler (Grayslake, IL); Suresh Bellam (Vernon Hills, IL); Andre Coetzee (Cary, IL); Dan Verdeyen (Glenview, IL)
Assignee: CDW LLC
G10L15/22G06F3/0484G06F3/167G06F40/295G06F40/30G06N3/08G06Q10/0635G10L15/1815G10L15/30G10L2015/223
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,264,025
App. No.
16/519,628
Granted
Mar 1, 2022
Kind
B2
Abstract

A system includes a processor and a memory storing instructions that, when executed by the processor, cause the system to receive an utterance, transmit the utterance to a cloud to generate an intent and an entity, receive the intent and the entity, and perform an action with respect to a graphical user interface. A method includes receiving an utterance, transmitting the utterance to a cloud to generate an intent and an entity, receiving the intent and the entity, and performing an action with respect to a graphical user interface. A non-transitory computer readable medium includes program instructions that when executed, cause a computer to receive an utterance, transmit the utterance to a cloud to generate an intent and an entity, receive the intent and the entity, and perform an action with respect to a graphical user interface.

Claims (53)

1. A computing system for automated graphical user interface control of legacy applications of an enterprise using voice commands, comprising:

one or more processors, and

a memory storing instructions that, when executed by the one or more processors, cause the computing system to:

receive an utterance of a user with respect to a graphical user interface of an application of an enterprise,

transmit the utterance to a cloud,

wherein the cloud analyzes the utterance using a speech-to-text application programming interface to generate a string output, and

wherein the cloud analyzes the string output using a command interpreter to generate an intent and an entity,

wherein the entity is associated with the intent,

receive the intent and the entity, and

perform an action with respect to an element of the graphical user interface of the application of the enterprise, wherein the intent corresponds to the intent and the element corresponds to the entity.

2. The computing system of claim 1 , wherein the graphical user interface is included in one or both of (i) a customer relationship management system, and (ii) a quantitative risk management system.

3. The computing system of claim 1 , wherein the intent corresponds to a navigation intent, the graphical user interface displays a first web page, and the action corresponds to displaying a second web page.

4. The computing system of claim 1 , wherein the entity includes a parameter, wherein the intent corresponds to an update intent, and wherein the action corresponds to updating the element of the graphical user interface using the parameter.

5. The computing system of claim 1 , wherein the memory includes further instructions that, when executed by the one or more processors, cause the computing system to:

determine that further information is needed from the user, and

receive the further information from the user via a multi-turn interaction.

6. The computing system of claim 1 , wherein the memory includes further instructions that, when executed by the one or more processors, cause the computing system to:

train a convolutional neural network using a plurality of labeled digital images to output an identifier corresponding to the element of the graphical user interface, and

manipulate the element of the graphical user interface using the identifier.

7. A computer-implemented method for automated graphical user interface control of legacy applications of an enterprise using voice commands, comprising:

receiving an utterance of a user with respect to a graphical user interface of an application of the enterprise,

transmitting the utterance to a cloud,

wherein the cloud analyzes the utterance using a speech-to-text application programming interface to generate a string output, and

wherein the cloud analyzes the string output using a command interpreter to generate an intent and an entity,

wherein the entity is associated with the intent,

receiving the intent and the entity, and

performing an action with respect to an element of the graphical user interface of the application of the enterprise, wherein the intent corresponds to the intent and the element corresponds to the entity.

8. The method of claim 7 , wherein the graphical user interface is included in one or both of (i) a customer relationship management system, and (ii) a quantitative risk management system.

9. The method of claim 7 , wherein the intent corresponds to a navigation intent, the graphical user interface displays a first web page, and the action corresponds to displaying a second web page.

10. The method of claim 7 , wherein the entity includes a parameter, wherein the intent corresponds to an update intent, and wherein the action corresponds to updating the element of the graphical user interface using the parameter.

11. The method of claim 7 , further comprising:

determining that further information is needed from the user, and

receiving the further information from the user via a multi-turn interaction.

12. The method of claim 7 , further comprising:

training a convolutional neural network using a plurality of labeled digital images to output an identifier corresponding to the element of the graphical user interface, and

manipulating the element of the graphical user interface using the identifier.

13. A non-transitory computer readable medium containing program instructions that when executed, cause a computer to:

receive an utterance of a user with respect to a graphical user interface of an application of an enterprise,

transmit the utterance to a cloud, wherein the cloud analyzes the utterance using a speech-to-text application programming interface to generate a string output, and wherein the cloud analyzes the string output using a command interpreter to generate an intent and an entity, wherein the entity is associated with the intent,

receive the intent and the entity, and

perform an action with respect to an element of the graphical user interface of the application of the enterprise, wherein the intent corresponds to the intent and the element corresponds to the entity.

14. The non-transitory computer readable medium of claim 13 , wherein the graphical user interface is included in one or both of (i) a customer relationship management system, and (ii) a quantitative risk management system.

15. The non-transitory computer readable medium of claim 13 , wherein the intent corresponds to a navigation intent, the graphical user interface displays a first web page, and the action corresponds to displaying a second web page.

16. The non-transitory computer readable medium of claim 13 , wherein the entity includes a parameter, wherein the intent corresponds to an update intent, and wherein the action corresponds to updating the element of the graphical user interface using the parameter.

17. The non-transitory computer readable medium of claim 13 , wherein the computer readable medium includes further instructions that, when executed by the one or more processors, cause a computer to:

determine that further information is needed from the user, and

receive the further information from the user via a multi-turn interaction.

18. The non-transitory computer readable medium of claim 13 , wherein the computer readable medium contains further program instructions that when executed, cause a computer to:

train a convolutional neural network using a plurality of labeled digital images to output an identifier corresponding to the element of the graphical user interface, and

manipulate the element of the graphical user interface using the identifier.

19. The non-transitory computer readable medium of claim 13 , wherein the utterance of the user is received from a mobile computing device of the user.

20. The non-transitory computer readable medium of claim 13 , wherein the computer readable medium contains further program instructions that when executed, cause a computer to:

locate the element using a fuzzy matching algorithm.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 24, 2019
From: KESSLER, JOSEPH; BELLAM, SURESH; COETZEE, ANDRE; VERDEYEN, DAN
To: CDW LLC
Reel/Frame 049846/0310 →
Continuity (1)
Related Publication 20210027774A1 · Jan 28, 2021
Cited By (1)
US 12,217,003