IP Library Granted Patent US 12,339,894
Granted Patent B2
US 12,339,894 · App. 17/481,831 · Granted Jun 24, 2025

Method and system for voice based media search

Inventors: Mukesh Patel (Fremont, CA); Lu Silverstein (Alviso, CA); Srinivas Jandhyala (Alviso, CA)
Assignee: Adeia Media Solutions Inc.
G06F16/48G06F16/40
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,339,894
App. No.
17/481,831
Granted
Jun 24, 2025
Kind
B2
Abstract

Voice-based input is used to operate a media device and/or to search for media content. Voice input is received by a media device via one or more audio input devices and is translated into a textual representation of the voice input. The textual representation of the voice input is used to search one or more cache mappings between input commands and one or more associated device actions and/or media content queries. One or more natural language processing techniques may be applied to the translated text and the resulting text may be transmitted as a query to a media search service. A media search service returns results comprising one or more content item listings and the results may be presented on a display to a user.

Claims (56)

1. A method comprising:

configuring a first user voice profile associated with a first genre, wherein the first user voice profile is configured to process speech related to the first genre;

configuring a second user voice profile associated with a second genre, wherein the second user voice profile is configured to process speech related to the second genre;

receiving, via a user interface, a selection of one of the first user voice profile or the second user voice profile as a selected user profile;

receiving a voice input data at a media device;

based at least in part on receiving the voice input data:

accessing the selected user profile;

processing the voice input data using the selected user profile to generate a signature based at least in part on a textual representation of the voice input data;

identifying a particular data entry of a set of data entries, wherein each data entry of the set of data entries specifies a mapping between a given signature and one or more device actions;

updating the set of data entries by storing the mapping between the signature and the textual representation of the voice input data;

based at least in part on identifying the particular data entry, sending a media content query to a media search service; and

based at least in part on the media content query, generating for output media assets identifiers of media assets related to the particular data entry and associated with a genre of the selected user profile.

2. The method of claim 1 , wherein the first genre is distinct from the second genre.

3. The method of claim 1 , wherein the first user voice profile is associated with a sports genre.

4. The method of claim 1 , wherein the first user voice profile is associated with a movie genre.

5. The method of claim 1 , wherein the first and the second user voice profile is associated with a same first user.

6. The method of claim 1 , wherein each user profile, from a plurality of user profiles, is customized to a genre and a sub-genre.

7. The method of claim 1 , wherein configuring the first user voice profile associated with the first genre comprises:

capturing characteristics of voice input data while a first user is searching for a media asset associated with the first genre; and

creating the first user voice profile based at least in part on captured characteristics.

8. The method of claim 1 , wherein configuring the second user voice profile associated with the second genre comprises:

capturing characteristics of voice input data while a first user is searching for a media asset associated with the second genre; and

creating the second user voice profile based at least in part on captured characteristics.

9. The method of claim 1 , further comprising:

searching for media assets based at least in part on the selected user profile; and

generating for display identifiers relating to the searched media assets.

10. The method of claim 1 , wherein, configuring the first user voice profile associated with the first genre is performed via a Graphical User Interface (GUI).

11. A system comprising:

communications circuitry configured to access a first and a second profile associated with a first user; and

control circuitry configured to:

configure a first user voice profile associated with a first genre, wherein the first user voice profile is configured to process speech related to the first genre;

configure a second user voice profile associated with a second genre, wherein the second user voice profile is configured to process speech related to the second genre;

receive, via a user interface, a selection of one of the first user voice profile or the second user voice profile as a selected user profile;

receive a voice input data at a media device;

based at least in part on receiving the voice input data:

access, using the communications circuitry, the selected user profile;

process the voice input data using the selected user profile to generate a signature based at least in part on a textual representation of the voice input data;

identify a particular data entry of a set of data entries, wherein each data entry of the set of data entries specifies a mapping between a given signature and one or more device actions;

update the set of data entries by storing the mapping between the signature and the textual representation of the voice input data;

based at least in part on identifying the particular data entry, send a media content query to a media search service; and

based at least in part on the media content query, generating for output media assets identifiers of media assets related to the particular data entry and associated with a genre of the selected user profile.

12. The system of claim 11 , wherein the first genre is distinct from the second genre.

13. The system of claim 11 , wherein the first user voice profile is associated with a sports genre.

14. The system of claim 11 , wherein the first user voice profile is associated with a movie genre.

15. The system of claim 11 , wherein the first and the second user voice profile is associated with a same first user.

16. The system of claim 11 , wherein each user profile, from a plurality of user profiles, is customized to a genre and a sub-genre.

17. The system of claim 11 , wherein configuring the first user voice profile associated with the first genre comprises, the control circuitry configured to:

capture characteristics of voice input data while the first user is searching for a media asset associated with the first genre; and

create the first user voice profile based at least in part on captured characteristics.

18. The system of claim 11 , wherein configuring the second user voice profile associated with the second genre comprises, the control circuitry configured to:

capture characteristics of voice input data while the first user is searching for a media asset associated with the second genre; and

create the second user voice profile based at least in part on captured characteristics.

19. The system of claim 11 , further comprising, the control circuitry configured to:

search for media assets based at least in part on the selected user profile; and

generate for display identifiers relating to the searched media assets.

20. The system of claim 11 , wherein, configuring the first user voice profile associated with the first genre is performed by the control circuitry via a Graphical User Interface (GUI).

Assignments (4)
CHANGE OF NAME Recorded Sep 27, 2024
From: TIVO SOLUTIONS INC.
To: ADEIA MEDIA SOLUTIONS INC.
Reel/Frame 069067/0504 →
SECURITY INTEREST Recorded May 19, 2023
From: ADEIA GUIDES INC.; ADEIA MEDIA HOLDINGS LLC; ADEIA MEDIA SOLUTIONS INC.; ADEIA SEMICONDUCTOR BONDING TECHNOLOGIES INC.; ADEIA SEMICONDUCTOR SOLUTIONS LLC; ADEIA SEMICONDUCTOR TECHNOLOGIES LLC
To: BANK OF AMERICA, N.A., AS COLLATERAL AGENT
Reel/Frame 063707/0884 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 23, 2021
From: PATEL, MUKESH; SILVERSTEIN, LU; JANDHYALA, SRINIVAS
To: TIVO INC.
Reel/Frame 057572/0165 →
CHANGE OF NAME Recorded Sep 23, 2021
From: TIVO INC.
To: TIVO SOLUTIONS INC.
Reel/Frame 057572/0175 →
Continuity (5)
Continuation 16265932 · Feb 1, 2019
Continuation 15949754 · Apr 10, 2018
Continuation 15645526 · Jul 10, 2017
Continuation 13665735 · Oct 31, 2012
Related Publication 20220012275A1 · Jan 13, 2022
References Cited (41)
US 6374214B1 · Friedland et al. · 2002 [cited by applicant]
US 6823493B2 · Baker · 2004 [cited by applicant]
US 8428227B2 · Angel et al. · 2013 [cited by applicant]
US 8564544B2 · Jobs · 2013 [cited by examiner]
US 8875021B2 · Lee · 2014 [cited by examiner]
US 9734151B2 · Patel et al. · 2017 [cited by applicant]
US 9971772B2 · Patel et al. · 2018 [cited by applicant]
US 10242005B2 · Patel et al. · 2019 [cited by applicant]
US 11151184B2 · Patel et al. · 2021 [cited by applicant]
US 20040102959A1 · Estrin · 2004 [cited by applicant]
US 20040220926A1 · Lamkin et al. · 2004 [cited by applicant]
US 20070150275A1 · Garner et al. · 2007 [cited by applicant]
US 20080005688A1 · Najdenovski · 2008 [cited by examiner]
US 20090228277A1 · Bonforte et al. · 2009 [cited by applicant]
US 20090299752A1 · Rodriguez et al. · 2009 [cited by applicant]
US 20090326938A1 · Marila et al. · 2009 [cited by applicant]
US 20110098917A1 · Lebeau et al. · 2011 [cited by applicant]
US 20110173539A1 · Rottler · 2011 [cited by examiner]
US 20110208524A1 · Haughay · 2011 [cited by examiner]
US 20110286584A1 · Angel et al. · 2011 [cited by applicant]
US 20120016678A1 · Gruber et al. · 2012 [cited by applicant]
US 20120035924A1 · Jitkoff et al. · 2012 [cited by applicant]
US 20120201362A1 · Crossan et al. · 2012 [cited by applicant]
US 20120245936A1 · Treglia · 2012 [cited by applicant]
US 20120259924A1 · Patil et al. · 2012 [cited by applicant]
US 20130332168A1 · Kim et al. · 2013 [cited by applicant]
US 20140115465A1 · Lee · 2014 [cited by examiner]
US 20140122059A1 · Patel et al. · 2014 [cited by applicant]
US 20180025001A1 · Patel et al. · 2018 [cited by applicant]
US 20180053507A1 · Wang et al. · 2018 [cited by applicant]
US 20180232368A1 · Patel et al. · 2018 [cited by applicant]
US 20190027131A1 · Zajac · 2019 [cited by applicant]
US 20190080685A1 · Johnson · 2019 [cited by applicant]
US 20190206405A1 · Gillespie et al. · 2019 [cited by applicant]
US 20190236089A1 · Patel et al. · 2019 [cited by applicant]
US 20230012940A1 · Patel et al. · 2023 [cited by applicant]
US 20230016510A1 · Patel et al. · 2023 [cited by applicant]
US 20230017928A1 · Patel et al. · 2023 [cited by applicant]
WO 2010119288A1 · 2010 [cited by applicant]
Takuo Henmi, Shengyang Huang and Fuji Ren, “Wisdom media “CAIWA Channel” based on natural language interface agent,” Proceedings of the 6th International Conference on Natural Language Processing and Knowledge Engineeri… [cited by examiner]
International Search Report and Written Opinion received for PCT Patent Application No. PCT/US13/67602, mailed on Jan. 30, 2014, 7 pages. [cited by applicant]