IP Library Granted Patent US 11,604,830
Granted Patent B2
US 11,604,830 · App. 16/736,065 · Granted Mar 14, 2023

Systems and methods for performing a search based on selection of on-screen entities and real-world entities

Inventors: Susanto Sen (Karnataka, IN); Charishma Chundi (Andhra Pradesh, IN)
Assignee: Rovi Guides, Inc.
G06F16/90332G06V40/28G10L15/26H04N21/47202
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,604,830
App. No.
16/736,065
Granted
Mar 14, 2023
Kind
B2
Abstract

A search is performed based on a voice input combined with user selection of entities displayed on a display screen as well as real-world entities. A voice input is received from the user by a media device, as well as a selection of a first entity being displayed on the media device. A gesture made by the user is also identified, and a second, real-world entity corresponding to the gesture is determined. A search query is constructed based on a search operator in the voice input, the first entity, and the second entity. The search query is transmitted to a database and, in response, the media device receives at least one identifier of a least one content item. The at least one identifier is then generated for display to the user.

Claims (95)

1. A method for searching for media content, the method comprising:

receiving, at a media device, a voice input from a user;

processing the voice input to identify a pronoun;

receiving, at the media device, a selection of a first entity currently being displayed on a display of the media device;

identifying a gesture made by the user, wherein the selection of the first entity is distinct from the gesture;

identifying a plurality of entities in an image associated with the gesture;

determining, based on the identity of each respective entity of the plurality of entities, a respective pronoun corresponding to each respective entity of the plurality of entities;

determining a second entity associated with the gesture, wherein the second entity is not being displayed on the display of the media device, and determining the second entity comprises selecting, as the second entity, an entity of the plurality of entities having a respective pronoun that matches the identified pronoun;

processing the voice input to identify a search operator;

constructing a search query based on the identified search operator, the first entity, and the second entity;

transmitting the search query to a database;

receiving, in response to the search query, at least one identifier of at least one content item; and

generating for presentation, using the media device, the at least one identifier.

2. The method of claim 1 , further comprising:

determining an identifier of the first entity; and

determining an identifier of the second entity;

wherein constructing the search query based on the identified search operator, the first entity, and the second entity comprises constructing the search query based on the identified search operator, the identifier of the first entity, and the identifier of the second entity.

3. The method of claim 1 , wherein the media device is a virtual reality display.

4. The method of claim 1 , wherein identifying the gesture made by the user comprises capturing, using a camera, a motion of the user.

5. The method of claim 1 , wherein determining the second entity associated with the gesture comprises:

determining a direction of the gesture;

capturing, using a camera, the image representing an area corresponding to the direction of the gesture; and

identifying the entity in the image associated with the gesture.

6. The method of claim 5 , wherein identifying the entity in the image associated with the gesture comprises:

performing image processing to identify the plurality of entities in the image;

extrapolating a path from the direction of the gesture; and

determining an entity of the plurality of entities that the path intersects.

7. The method of claim 5 , further comprising:

performing image processing to identify the plurality of entities in the image;

wherein identifying the entity in the image associated with the gesture comprises:

selecting, as the second entity, an entity of the plurality of entities having a respective pronoun that matches the identified pronoun by comparing the respective pronoun of each respective entity of the plurality of entities with the identified pronoun.

8. The method of claim 5 , wherein the camera is a first camera and wherein the image captured by the first camera is a first image, the method further comprising:

capturing, using a second camera, a second image representing the area corresponding to the direction of the gesture, wherein the first image and the second image each represent a different perspective of the area corresponding to the direction of the gesture, and wherein the gesture is captured in the first image and in the second image;

extrapolating a first path from the direction of the gesture in the first image;

extrapolating a second path from the direction of the gesture in the second image; and

identifying a point at which the first path crosses the second path.

9. The method of claim 8 , further comprising identifying at least one entity at the point at which the first path crosses the second path.

10. The method of claim 5 , wherein the camera is a first camera that is facing a first direction, the method further comprising:

capturing, using a second camera that is facing a second direction, a second image representing an area in which the user made the gesture;

performing image processing of the second image to identify the gesture;

extrapolating a first path from the direction of the gesture in the second image;

calculating, based on a position and an angle of the first camera and a position and an angle of the second camera, a second path in the first image corresponding to the first path;

performing image processing of the first image to identify the plurality of entities in the first image; and

determining an entity of the plurality of entities that the second path intersects.

11. A system for searching for media content, the system comprising:

a display;

memory; and

control circuitry configured to:

receive a voice input from a user;

process the voice input to identify a pronoun;

receive a selection of a first entity currently being displayed on the display, wherein the selection of the first entity is distinct from the gesture;

identify a gesture made by the user;

identify a plurality of entities in an image associated with the gesture;

determine, based on the identity of each respective entity of the plurality of entities, a respective pronoun corresponding to each respective entity of the plurality of entities;

determine a second entity associated with the gesture, wherein the second entity is not being displayed on the display, and the control circuitry is configured to determine the second entity by selecting, as the second entity, an entity of the plurality of entities having a respective pronoun that matches the identified pronoun;

process the voice input to identify a search operator;

construct a search query based on the identified search operator, the first entity, and the second entity;

transmit the search query to a database stored in the memory;

receive, in response to the search query, at least one identifier of at least one content item; and

generate for presentation, using the media device, the at least one identifier.

12. The system of claim 11 , wherein the control circuitry is further configured to:

determine an identifier of the first entity; and

determine an identifier of the second entity;

wherein the control circuitry configured to construct the search query based on the identified search operator, the first entity, and the second entity is configured to construct the search query based on the identified search operator, the identifier of the first entity, and the identifier of the second entity.

13. The system of claim 11 , wherein the display is a virtual reality display.

14. The system of claim 11 , further comprising:

a camera; and

wherein the control circuitry configured to identify the gesture made by the user is further configured to capture, using the camera, a motion of the user.

15. The system of claim 11 , further comprising:

a camera;

wherein the control circuitry configured to determine the second entity associated with the gesture is further configured to:

determine a direction of the gesture;

capture, using the camera, the image representing an area corresponding to the direction of the gesture; and

identify the entity in the image associated with the gesture.

16. The system of claim 15 , wherein the control circuitry configured to identify the entity in the image associated with the gesture is configured to:

perform image processing to identify the plurality of entities in the image;

extrapolate a path from the direction of the gesture; and

determine an entity of the plurality of entities that the path intersects.

17. The system of claim 15 , wherein the control circuitry is further configured to:

perform image processing to identify the plurality of entities in the image;

wherein the control circuitry configured to identify the entity in the image associated with the gesture is configured to:

select, as the second entity, an entity of the plurality of entities having a respective pronoun that matches the identified pronoun by comparing the respective pronoun of each respective entity of the plurality of entities with the identified pronoun.

18. The system of claim 15 , wherein the camera is a first camera and wherein the image captured by the first camera is a first image, the system further comprising a second camera, wherein the control circuitry is further configured to:

capture, using the second camera, a second image representing the area corresponding to the direction of the gesture, wherein the first image and the second image each represent a different perspective of the area corresponding to the direction of the gesture, and wherein the gesture is captured in the first image and in the second image;

extrapolate a first path from the direction of the gesture in the first image;

extrapolate a second path from the direction of the gesture in the second image; and

identify a point at which the first path crosses the second path.

19. The system of claim 18 , wherein the control circuitry is further configured to identify at least one entity at the point at which the first path crosses the second path.

20. The system of claim 15 , wherein the camera is a first camera that is facing a first direction, and wherein the control circuitry is further configured to:

capture, using a second camera that is facing a second direction, a second image representing an area in which the user made the gesture;

perform image processing of the second image to identify the gesture;

extrapolate a first path from the direction of the gesture in the second image;

calculate, based on a position and an angle of the first camera and a position and an angle of the second camera, a second path in the first image corresponding to the first path;

perform image processing of the first image to identify the plurality of entities in the first image; and

determine an entity of the plurality of entities that the second path intersects.

Assignments (3)
CHANGE OF NAME Recorded Oct 3, 2024
From: ROVI GUIDES, INC.
To: ADEIA GUIDES INC.
Reel/Frame 069106/0238 →
SECURITY INTEREST Recorded Jun 1, 2020
From: ROVI SOLUTIONS CORPORATION; ROVI TECHNOLOGIES CORPORATION; ROVI GUIDES, INC.; TIVO SOLUTIONS INC.; VEVEO, INC.; INVENSAS CORPORATION; INVENSAS BONDING TECHNOLOGIES, INC.; TESSERA, INC.; TESSERA ADVANCED TECHNOLOGIES, INC.; DTS, INC.; PHORUS, INC.; IBIQUITY DIGITAL CORPORATION
To: BANK OF AMERICA, N.A.
Reel/Frame 053468/0001 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 8, 2020
From: SEN, SUSANTO; CHUNDI, CHARISHMA
To: ROVI GUIDES, INC.
Reel/Frame 051447/0561 →