IP Library Granted Patent US 10,262,195
Granted Patent B2
US 10,262,195 · App. 15/244,794 · Granted Apr 16, 2019

Predictive and responsive video analytics system and methods

Inventors: Kelly Conway (Lake Bluff, IL); Christopher Danson (Austin, TX)
Assignee: MATTERSIGHT CORPORATION
G06K9/00342G06K9/00302G06K9/00335G06K9/00597G06T7/20G10L15/26G10L17/22G10L17/26G10L25/57G10L25/63G06K2009/00328G06T2207/10016G06T2207/30201
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,262,195
App. No.
15/244,794
Granted
Apr 16, 2019
Kind
B2
Abstract

Methods and systems to predict user behavior based on analysis of a video communication by one or more processors, which methods include receiving a user video communication, extracting video analysis data optionally including facial analysis data for the user from the video communication, extracting, by the one or more processors, voice analysis data from the user video communication, generating an outcome prediction score based on the video analysis data and voice analysis data that predicts a likelihood that a user will take an action leading to an outcome.

Claims (53)

1. A video analytics system adapted to predict user behavior based on analysis of a video communication, which comprises:

a node comprising a processor and a non-transitory computer readable medium operably coupled thereto, the non-transitory computer readable medium comprising a plurality of instructions stored in association therewith that are accessible to, and executable by, the processor, where the plurality of instructions comprises:

instructions that, when executed, receive a video communication from a user, wherein the video communication comprises an audio component and a video component;

instructions that, when executed, analyze the video component to provide timecoded video behavioral data from the user and identify an attire of the user;

instructions that, when executed, generate an avatar that includes one or more elements of the user's attire, wherein the avatar depicts a customer service agent wearing the one or more elements of the user's attire;

instructions that, when executed, generate a responsive communication from the avatar to the user;

instructions that, when executed, analyze the audio component to provide timecoded spoken words from the user and from the responsive communication; and

instructions that, when executed, generate an outcome prediction score, based on the time-coded video behavioral data and the time-coded spoken words, that predicts a likelihood that the user will take an action leading to an outcome.

2. The system of claim 1 , which further comprises instructions that, when executed, associate the time-coded spoken words with the video behavioral data to determine an emotional state of the user, wherein the generated outcome prediction score is further based on the emotional state of the user.

3. The system of claim 1 , which further comprises instructions that, when executed, determine a personality type of the user by applying a linguistic-based algorithm to a text of the spoken words, wherein the generated outcome prediction score is further based on the personality type of the user.

4. The system of claim 1 , wherein the user's attire is analyzed for elements comprising clothing, jewelry, or accessories.

5. The system of claim 1 , wherein the outcome comprises one or more of whether the user will terminate his or her account, whether the user will purchase a product, whether the user is a fraudster, and whether the user will initiate additional subsequent interaction sessions regarding an issue.

6. The system of claim 1 , wherein the video component includes a non-verbal, non-textual element comprising one or more of eye movement, facial expressions, gestures, activities, body postures, behaviors, attire, and actions.

7. The system of claim 1 , which further comprises instructions that, when executed, determine user value data based on estimated cost of the user's attire.

8. The system of claim 1 , which further comprises instructions that, when executed, recommend items based on the attire of the user and display one or more of the items in the responsive communication.

9. The system of claim 1 , which further comprises instructions that, when executed, determine distress and engagement of the user.

10. The system of claim 9 , wherein the outcome prediction score is further based on the user's distress and engagement.

11. The system of claim 1 , which further comprises instructions that, when executed:

determine, based on the outcome prediction score for a selected outcome of the video communication, a plurality of specific words to include in the responsive communication to the user; and

further develop the responsive communication including the specific words.

12. The system of claim 1 , which further comprises instructions that, when executed, generate an avatar that is displayed to the user and provides the responsive communication to the user.

13. A method to predict user behavior based on analysis of a video communication, which comprises:

receiving, by one or more processors, a user video communication;

extracting, by the one or more processors, video analysis data from the user video communication including facial analysis data and identification of an attire of the user;

generating an avatar that includes one or more elements of the user's attire, wherein the avatar depicts a customer service agent wearing the one or more elements of the user's attire;

generating a responsive communication from the avatar to the user;

extracting, by the one or more processors, voice analysis data from the user in the user video communication and from the responsive communication; and

generating an outcome prediction score, based on the video analysis data including facial analysis data and identification of the attire of the user and based on the voice analysis data, that predicts a likelihood that the user will take an action leading to an outcome.

14. The method of claim 13 , which further comprises associating the time-coded spoken words with the video analysis data to determine an emotional state of the user, wherein the generated outcome prediction score is further based on the emotional state of the user.

15. The method of claim 13 , which further comprises determining a personality type of the user by applying a linguistic-based algorithm to a text of the spoken words, wherein the generated outcome prediction score is further based on the personality type of the user.

16. The method of claim 13 , wherein the user's attire is analyzed for elements selected to comprise clothing, jewelry, or accessories.

17. The method of claim 13 , wherein the outcome comprises one or more of whether the user will terminate his or her account, whether the user will purchase a product, whether the user is a fraudster, and whether the user will initiate additional subsequent interaction sessions regarding an issue.

18. The method of claim 13 , wherein the video analysis data includes a non-verbal, non-textual element comprising one or more of eye movement, facial expressions, gestures, activities, body postures, behaviors, attire, and actions.

19. The method of claim 13 , which further comprises determining user value data based on estimated cost of the user's attire.

20. The method of claim 13 , which further comprises recommending items based on the attire of the user and displaying one or more of the items in the responsive communication.

21. The method of claim 13 , which further comprises determining distress and engagement of the user.

22. The method of claim 21 , wherein the outcome prediction score is further based on the user's distress and engagement.

23. The method of claim 13 , which further comprises determining, based on the outcome prediction score for a selected outcome of the video communication, a plurality of specific words to include in the responsive communication to the user; and further developing the responsive communication including the specific words.

24. A non-transitory machine-readable medium comprising a plurality of instructions which, in response to a computer system, cause the computer system to perform a method which comprises:

receiving a user video communication from a user;

separating an audio component from a video component of the video communication;

analyzing facial expressions of the user in the video component;

identifying an attire of the user in the video component;

generating an avatar that includes one or more elements of the user's attire, wherein the avatar depicts a customer service agent wearing the one or more elements of the user's attire;

generating a responsive communication from the avatar to the user;

transcribing words spoken of the user in the audio component and from the responsive communication; and

generating an outcome prediction score, based on the analyzed video component and transcribed words, that predicts a likelihood that the user will take an action leading to an outcome.

25. The non-transitory machine-readable medium of claim 24 , which further comprises associating the transcribed words with the analyzed facial expressions to determine an emotional state of the user, wherein the generated outcome prediction score is further based on the emotional state of the user.

26. The non-transitory machine-readable medium of claim 24 , which further comprises determining a personality type of the user by applying a linguistic-based algorithm to the transcribed words, wherein the generated outcome prediction score is further based on the personality type of the user.

27. The non-transitory machine-readable medium of claim 24 , wherein the user's attire is analyzed for elements comprising clothing, jewelry, or accessories.

28. The non-transitory machine-readable medium of claim 24 , which further comprises determining user value data based on estimated cost of the user's attire.

29. The non-transitory machine-readable medium of claim 24 , which further comprises recommending items based on the attire of the user and displaying one or more of the items in the responsive communication.

30. The non-transitory machine-readable medium of claim 24 , which further comprises determining, based on the outcome prediction score for a selected outcome of the video communication, a plurality of specific words to include in the responsive communication to the user; and further developing the responsive communication including the specific words.

Assignments (2)
SECURITY INTEREST Recorded Jul 14, 2017
From: MATTERSIGHT CORPORATION
To: THE PRIVATEBANK AND TRUST COMPANY
Reel/Frame 043200/0001 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 23, 2016
From: CONWAY, KELLY; DANSON, CHRISTOPHER
To: MATTERSIGHT CORPORATION
Reel/Frame 039512/0764 →
Continuity (3)
Continuation 14996913 · Jan 15, 2016
Continuation 14525002 · Oct 27, 2014
Related Publication 20160364606A1 · Dec 15, 2016
Cited By (2)
US 12,598,270 US 12,641,193