IP Library Granted Patent US 10,289,900
Granted Patent B2
US 10,289,900 · App. 15/267,732 · Granted May 14, 2019

System and method for body language analysis

Inventor: Jonathan M. Keller (Lafayette, IN)
G06K9/00335G06T7/20G08B21/18G09B5/065G09B5/14
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,289,900
App. No.
15/267,732
Granted
May 14, 2019
Kind
B2
Abstract

A system and method are presented for body language analysis of a video interaction. In a contact center system, the video interaction between an agent and a customer may be monitored and used to determine automatic actions when threshold are met and/or matches are made. Training videos comprising determined metrics may be used for comparison to real-time interactions. Scoring and/or matches may be utilized to determine a threshold to trigger pre-determined actions based on comparison to the training videos.

Claims (56)

1. A method for analyzing gestures of one or more parties to a video interaction in a contact center system, wherein the contact center system comprises at least a video stream analyzer, and performing actions based on the gestures, the method comprising the steps of:

a. receiving the video interaction in the contact center system from a first user to a second user, wherein the video interaction occurs through a video stream comprising a plurality of video frames comprising a plurality of pixels;

b. determining metrics, for the first user, for a previous time interval of the on-going video stream by the video stream analyzer, the determining the metrics comprising:

identifying pixels that moved in the video frames of the video stream; and

computing at least one of a motion energy image metric or a motion history image metric based on the identified pixels;

c. referencing stored metrics of system training videos, by the video stream analyzer, and comparing the stored metrics with the determined metrics of step (b) of the on-going video stream; and

d. determining if a match is made between the determined metrics and the stored metrics from the comparing of step (c), wherein if the match is made, pre-classifying the video interaction and performing an action, wherein if the match is not made, repeating the process from step (b);

wherein steps (b)-(d) cyclically occur continuously in real-time throughout a duration of the video interaction.

2. The method of claim 1 , wherein the video interaction comprises a video chat.

3. The method of claim 1 , wherein the system training videos comprise gestures indicating emotion and a plurality of levels of said emotion based on gestures.

4. The method of claim 1 , wherein the system training videos comprise gestures used in relation to at least one of: speech amplitude and keyword spotting.

5. The method of claim 1 , wherein the determining the metrics comprises determining the metrics on a cumulative basis with each passing time interval.

6. The method of claim 1 , wherein the pre-classifying comprises assigning a match to the video interaction and the action performed is based on the pre-classifying.

7. The method of claim 1 , wherein the action comprises at least one of: alerting a supervisor, alerting an authoritative figure, altering script used by the second user, recording the video interaction, generating an email to the first user, impacting the second user's performance evaluation, terminating the video interaction, placing the video interaction on hold, altering the second user's utilization score, and altering a number of interactions routed to the second user for a time interval.

8. The method of claim 1 , wherein the first user comprises a customer.

9. The method of claim 1 , wherein the second user comprises a customer service representative.

10. The method of claim 1 , wherein the previous time interval of the on-going video stream is a fraction of a second in length, and wherein the motion energy image metric or the motion history image metric is determined at least once during each such time interval.

11. A method for analyzing gestures of one or more parties to a video interaction in a contact center system, wherein the contact center system comprises at least a video stream analyzer, and performing actions based on the gestures, the method comprising the steps of:

a. receiving the video interaction in the contact center system from a first user to a second user, wherein the video interaction occurs through a video stream comprising a plurality of video frames comprising a plurality of pixels;

b. determining metrics, for at least one of the first user and the second user, for a previous time interval of the on-going video stream by the video stream analyzer, the determining the metrics comprising:

identifying pixels that moved in the video frames of the video stream; and;

computing at least one of a motion energy image metric or a motion history image metric based on the identified pixels;

c. referencing stored metrics of system training videos, by the video stream analyzer, and comparing the stored metrics with the determined metrics of step (b) of the on-going video stream;

d. updating a cumulative score in response to determining a match from the comparing the stored metrics with the determined metrics of step (c); and

e. determining if a threshold is met by the cumulative score, wherein if the threshold is met, performing an action, wherein if the threshold is not met, repeating the process from step (b), wherein

steps (b)-(d) cyclically occur continuously in real-time throughout the duration of the video interaction.

12. The method of claim 11 , wherein the threshold is determined by cumulative points from negative gestures.

13. The method of claim 11 , wherein the threshold is determined by cumulative points from negative gestures and positive gestures, wherein points from positive gestures negate points from negative gestures.

14. The method of claim 11 , wherein the video interaction comprises a video chat.

15. The method of claim 11 , wherein the system training videos comprise gestures indicating emotion and a plurality of levels of said emotion based on gestures.

16. The method of claim 11 , wherein the system training videos comprise gestures used in relation to at least one of: speech amplitude and keyword spotting.

17. The method of claim 11 , wherein the determining the metrics comprises determining the metrics on a cumulative basis with each passing time interval.

18. The method of claim 11 , wherein the action comprises at least one of: alerting a supervisor, alerting an authoritative figure, altering script used by the second user, recording the video interaction, generating an email to the first user, impacting the second user's performance evaluation, terminating the video interaction, placing the video interaction on hold, altering the second user's utilization score, and altering a number of interactions routed to the second user for a time interval.

19. The method of claim 11 , wherein the first user comprises a customer.

20. The method of claim 11 , wherein the second user comprises a customer service representative.

21. The method of claim 11 , wherein the previous time interval of the on-going video stream is a fraction of a second in length, and wherein the motion energy image metric or the motion history image metric is determined at least once during each such time interval.

22. A method for analyzing gestures of one or more parties to a video interaction in a contact center system, wherein the contact center system comprises at least a video stream analyzer, and performing actions based on the gestures, the method comprising the steps of:

a. receiving the video interaction in the contact center system from a first user to a second user, wherein the video interaction occurs through a video stream comprising a plurality of video frames comprising a plurality of pixels;

b. determining metrics, for at least one of the first user and the second user, for a previous time interval of the video stream by the video stream analyzer, the determining the metrics comprising:

identifying pixels that moved in the video frames of the video stream; and

computing at least one of a motion energy metric or a motion history metric based on the identified pixels;

c. referencing stored metrics of system training videos, by the video stream analyzer, and comparing the stored metrics with the determined metrics of step (b) of the on-going video stream; and

d. determining if one or more conditions are met from the comparing of step (c), wherein if the one or more conditions are met, pre-classifying the video interaction and performing an action, wherein if the one or more conditions are not met, repeating the process from step (b), wherein

steps (b)-(d) cyclically occur continuously in real-time throughout the duration of the video interaction.

23. The method of claim 22 , wherein the conditions comprise a match and a threshold.

24. The method of claim 23 , wherein the threshold is determined by cumulative points from negative gestures.

25. The method of claim 23 , wherein the threshold is determined by cumulative points from negative gestures and positive gestures, wherein points from positive gestures negate points from negative gestures.

26. The method of claim 22 , wherein the video interaction comprises a video chat.

27. The method of claim 22 , wherein the system training videos comprise gestures indicating emotion and a plurality of levels of said emotion based on gestures.

28. The method of claim 22 , wherein the system training videos comprise gestures used in relation to at least one of: speech amplitude and keyword spotting.

29. The method of claim 22 , wherein the determining the metrics comprises determining the metrics on a cumulative basis with each passing time interval.

30. The method of claim 22 , wherein the pre-classifying comprises assigning a match to the video interaction and the action performed is based on the pre-classifying.

31. The method of claim 22 , wherein the action comprises at least one of: alerting a supervisor, alerting an authoritative figure, altering script used by the second user, recording the video interaction, generating an email to the first user, impacting the second user's performance evaluation, terminating the video interaction, placing the video interaction on hold, altering the second user's utilization score, and altering a number of interactions routed to the second user for a time interval.

32. The method of claim 22 , wherein the first user comprises a customer.

33. The method of claim 22 , wherein the second user comprises a customer service representative.

34. The method of claim 22 , wherein the previous time interval of the on-going video stream is a fraction of a second in length, and wherein the motion energy image metric or the motion history image metric is determined at least once during each such time interval.

Assignments (5)
NOTICE OF SUCCESSION OF SECURITY INTERESTS AT REEL/FRAME 040815/0001 Recorded Feb 3, 2025
From: BANK OF AMERICA, N.A., AS RESIGNING AGENT
To: GOLDMAN SACHS BANK USA, AS SUCCESSOR AGENT
Reel/Frame 070498/0001 →
CHANGE OF NAME Recorded Jun 7, 2024
From: GENESYS TELECOMMUNICATIONS LABORATORIES, INC.
To: GENESYS CLOUD SERVICES, INC.
Reel/Frame 067651/0814 →
MERGER Recorded Jul 1, 2018
From: INTERACTIVE INTELLIGENCE GROUP, INC.
To: GENESYS TELECOMMUNICATIONS LABORATORIES, INC.
Reel/Frame 046463/0839 →
SECURITY AGREEMENT Recorded Dec 5, 2016
From: GENESYS TELECOMMUNICATIONS LABORATORIES, INC., AS GRANTOR; ECHOPASS CORPORATION; INTERACTIVE INTELLIGENCE GROUP, INC.; BAY BRIDGE DECISION TECHNOLOGIES, INC.
To: BANK OF AMERICA, N.A., AS COLLATERAL AGENT
Reel/Frame 040815/0001 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 19, 2016
From: KELLER, JONATHAN M.
To: INTERACTIVE INTELLIGENCE GROUP, INC.
Reel/Frame 039779/0732 →
Continuity (1)
Related Publication 20180082112A1 · Mar 22, 2018