IP Library Granted Patent US 10,043,065
Granted Patent B2
US 10,043,065 · App. 14/813,500 · Granted Aug 7, 2018

Systems and methods for determining meaning of cultural gestures based on voice detection

Inventors: Alejandro S. Pulido (Chatsworth, CA); Michael R. Nichols (La Canada Flintridge, CA)
Assignee: Rovi Guides, Inc.
G06K9/00355G10L15/005G10L2015/027
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,043,065
App. No.
14/813,500
Granted
Aug 7, 2018
Kind
B2
Abstract

In some embodiments, control circuitry may detect a voice communication from a human using voice detection circuitry during playback of a media asset being consumed by the user. Control circuitry may then identify an accent characteristic of the voice communication. Control circuitry may cross-reference the accent characteristic against listings of an accent database and then determine, based on the cross-referencing of the accent characteristic a country of origin of the human. Control circuitry may detect, using imaging circuitry, a gesture made by the human. Control circuitry may then cross-reference the gesture against listings of a gesture database associated with the country of origin and then determine based on the cross-referencing of the gesture, a meaning of the gesture in relation to the media asset.

Claims (53)

1. A method for providing recommendations in relation to media assets in gesture recognition computer systems by determining a meaning of cultural gestures based on voice detection, the method comprising:

generating for display a media asset for playback;

detecting a voice communication from a human using voice detection circuitry during playback of the media asset being consumed by the human;

identifying an accent characteristic of the voice communication;

cross-referencing the accent characteristic against listings of an accent database;

determining, based on the cross-referencing of the accent characteristic, a country of origin of the human;

detecting, using imaging circuitry, a gesture made by the human;

cross-referencing the gesture against listings of a gesture database associated with the country of origin to determine whether the gesture corresponds to a meaning in the country of origin;

determining, based on the cross-referencing of the gesture, a meaning of the gesture in relation to the media asset; and

providing a recommendation to the human based on the meaning of the gesture in relation to the media asset.

2. The method of claim 1 , wherein the recommendation is tailored to the country of origin.

3. The method of claim 1 , wherein the meaning comprises at least one of: an indication of enjoyment of the media asset by the human; an indication of distaste for the media asset by the human; an indication that the human wishes to suspend viewing the media asset; and an indication that the human wishes to alert information about the media asset to another human.

4. The method of claim 1 , wherein the identifying of the accent characteristic comprise:

detecting, using voice processing circuitry, a respective manner of annunciating each syllable of a plurality of syllables of the voice communication;

determining whether a threshold amount of syllables correspond to a single respective manner; and

in response to determining that the threshold amount of syllables corresponds to the single respective manner, identifying the accent characteristics by determining that the single respective manner corresponds to the accent characteristic.

5. The method of claim 4 , wherein detecting the respective manner of annunciating each syllable comprises comparing the respective manner of annunciating each syllable to a known universe of potential manners of annunciating each syllable, and identifying a match between the respective manner and a manner in the known universe of potential manners.

6. The method of claim 1 , wherein the meaning of the gesture varies based on the country of origin.

7. The method of claim 1 , wherein multiple countries of origin are determined based on the cross-referencing of the accent characteristic, and wherein a single country of origin of the multiple countries of origin is identified by:

determining, using imaging circuitry, body characteristics of the human;

comparing the body characteristics of the human to listings of a body characteristic database that correspond to each country of the multiple countries of origin; and

determining, based on the comparing of the body characteristics of the human to the listings of the body characteristic database, the country of origin.

8. The method of claim 1 , wherein detecting the voice communication comprises:

detecting, using a microphone, audio comprising communication from the human and ambient noise comprising audio of the media asset; and

isolating the audio comprising communication from the human from the ambient noise to detect the voice communication.

9. The method of claim 1 , wherein the gesture comprises at least one of a hand movement, a leg movement, a body movement, and a collision between a body part of the human and an inanimate object.

10. A system for providing recommendations in relation to media assets in gesture recognition computer systems by determining a meaning of cultural gestures based on voice detection, the system comprising:

voice detection circuitry;

imaging circuitry configured to generate for display a media asset for playback; and

control circuitry configured to:

detect, using the voice detection circuitry, a voice communication from a human during playback of the media asset being consumed by the human;

identify an accent characteristic of the voice communication;

cross-reference the accent characteristic against listings of an accent database;

determine, based on the cross-referencing of the accent characteristic, a country of origin of the human;

detect, using the imaging circuitry, a gesture made by the human;

cross-reference the gesture against listings of a gesture database associated with the country of origin to determine whether the gesture corresponds to a meaning in the country of origin; and

determine, based on the cross-referencing of the gesture, a meaning of the gesture in relation to the media asset, wherein a recommendation is provided, by the imaging circuitry, to the human based on the meaning of the gesture in relation to the media asset.

11. The system of claim 10 , wherein the recommendation is tailored to the country of origin.

12. The system of claim 10 , wherein the meaning comprises at least one of: an indication of enjoyment of the media asset by the human; an indication of distaste for the media asset by the human; an indication that the human wishes to suspend viewing the media asset; and an indication that the human wishes to alert information about the media asset to another human.

13. The system of claim 10 , wherein the system further comprises voice processing circuitry, and wherein the control circuitry is further configured, when identifying of the accent characteristic, to:

detect, using the voice processing circuitry, a respective manner of annunciating each syllable of a plurality of syllables of the voice communication;

determine whether a threshold amount of syllables correspond to a single respective manner; and

in response to determining that the threshold amount of syllables corresponds to the single respective manner, identify the accent characteristics by determining that the single respective manner corresponds to the accent characteristic.

14. The system of claim 13 , wherein the control circuitry is further configured, when detecting the respective manner of annunciating each syllable, to compare the respective manner of annunciating each syllable to a known universe of potential manners of annunciating each syllable, and to identify a match between the respective manner and a manner in the known universe of potential manners.

15. The system of claim 10 , wherein the meaning of the gesture varies based on the country of origin.

16. The system of claim 10 , wherein multiple countries of origin are determined based on the cross-referencing of the accent characteristic, and wherein control circuitry is further configured to identify a single country of origin of the multiple countries of origin by:

determining, using the imaging circuitry, body characteristics of the human;

comparing the body characteristics of the human to listings of a body characteristic database that correspond to each country of the multiple countries of origin; and

determining, based on the comparing of the body characteristics of the human to the listings of the body characteristic database, the country of origin.

17. The system of claim 10 , wherein the system further comprises a microphone, and wherein the control circuitry is further configured, when detecting the voice communication, to:

detect, using the microphone, audio comprising communication from the human and ambient noise comprising audio of the media asset; and

isolate the audio comprising communication from the human from the ambient noise to detect the voice communication.

18. The system of claim 10 , wherein the gesture comprises at least one of a hand movement, a leg movement, a body movement, and a collision between a body part of the human and an inanimate object.

Assignments (7)
CHANGE OF NAME Recorded Oct 2, 2024
From: ROVI GUIDES, INC.
To: ADEIA GUIDES INC.
Reel/Frame 069085/0715 →
RELEASE OF SECURITY INTEREST Recorded Jun 5, 2020
From: HPS INVESTMENT PARTNERS, LLC
To: ROVI SOLUTIONS CORPORATION; ROVI TECHNOLOGIES CORPORATION; ROVI GUIDES, INC.; TIVO SOLUTIONS, INC.; VEVEO, INC.
Reel/Frame 053458/0749 →
RELEASE OF SECURITY INTEREST Recorded Jun 5, 2020
From: MORGAN STANLEY SENIOR FUNDING, INC.
To: ROVI SOLUTIONS CORPORATION; ROVI TECHNOLOGIES CORPORATION; ROVI GUIDES, INC.; TIVO SOLUTIONS, INC.; VEVEO, INC.
Reel/Frame 053481/0790 →
SECURITY INTEREST Recorded Jun 1, 2020
From: ROVI SOLUTIONS CORPORATION; ROVI TECHNOLOGIES CORPORATION; ROVI GUIDES, INC.; TIVO SOLUTIONS INC.; VEVEO, INC.; INVENSAS CORPORATION; INVENSAS BONDING TECHNOLOGIES, INC.; TESSERA, INC.; TESSERA ADVANCED TECHNOLOGIES, INC.; DTS, INC.; PHORUS, INC.; IBIQUITY DIGITAL CORPORATION
To: BANK OF AMERICA, N.A.
Reel/Frame 053468/0001 →
PATENT SECURITY AGREEMENT Recorded Nov 25, 2019
From: ROVI SOLUTIONS CORPORATION; ROVI TECHNOLOGIES CORPORATION; ROVI GUIDES, INC.; TIVO SOLUTIONS, INC.; VEVEO, INC.
To: MORGAN STANLEY SENIOR FUNDING, INC., AS COLLATERAL AGENT
Reel/Frame 051110/0006 →
SECURITY INTEREST Recorded Nov 22, 2019
From: ROVI SOLUTIONS CORPORATION; ROVI TECHNOLOGIES CORPORATION; ROVI GUIDES, INC.; TIVO SOLUTIONS, INC.; VEVEO, INC.
To: HPS INVESTMENT PARTNERS, LLC, AS COLLATERAL AGENT
Reel/Frame 051143/0468 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 30, 2015
From: PULIDO, ALEJANDRO S.; NICHOLS, MICHAEL R.
To: ROVI GUIDES, INC.
Reel/Frame 036218/0495 →
Continuity (1)
Related Publication 20170032792A1 · Feb 2, 2017