IP Library Granted Patent US 11,966,442
Granted Patent B2
US 11,966,442 · App. 16/926,830 · Granted Apr 23, 2024

Recommending language models for search queries based on user profile

Inventor: Arun Sreedhara (Karnataka, IN)
Assignee: Rovi Product Corporation
G06F16/90332G10L15/02G10L15/187G10L15/22G10L25/51G06F16/9038G10L2015/025G10L15/1822G10L2015/223
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,966,442
App. No.
16/926,830
Granted
Apr 23, 2024
Kind
B2
Abstract

Systems and methods for a media guidance application that generates results in multiple languages for search queries. In particular, the media guidance application ranks search results according the language model associated with the search result.

Claims (93)

1. A method for generating results in multiple languages for search queries, comprising:

receiving, by a media guidance application on a display device, a first voice query;

accessing, by the media guidance application on the display device, a first language model and a second language model;

identifying, by the media guidance application on the display device, a first audio segment and a second audio segment based on audible breaks in the first voice query;

determining, by the media guidance application on the display device, a first resolved word in a first language and a second resolved word in the first language by combining a first set of phonemes to the first audio segment and second audio segment, respectively, the first set of phonemes selected from phonemes of the first language model;

generating, by the media guidance application on the display device, a first text query including the first resolved word in the first language and the second resolved word in the first language;

determining, by the media guidance application on the display device, a first numerical ranking for the first resolved word in the first language and a second numerical ranking for the second resolved word in the first language based on a likelihood of usage of the first resolved word in the first language and the second resolved word in the first language;

determining, by the media guidance application on the display device, a first composite score for the first text query based on summing the first numerical ranking and the second numerical ranking;

determining, by the media guidance application on the display device, a first resolved word in a second language and a second resolved word in the second language by combining a second set of phonemes to the first audio segment and second audio segment, respectively, the second set of phonemes selected from phonemes of the second language model;

generating, by the media guidance application on the display device, a second text query including the first resolved word in the second language and the second resolved word in the second language;

determining, by the media guidance application on the display device, a third numerical ranking for the first resolved word in the second language and a fourth numerical ranking for the second resolved word in the second language based on a likelihood of usage of the first resolved word in the second language and the second resolved word in the second language;

determining, by the media guidance application on the display device, a second composite score for the second text query based on summing the third numerical ranking and the fourth numerical ranking;

generating, by the media guidance application on the display device, a first search result for the first text query and a second search result for the second text query;

ranking, by the media guidance application on the display device, the first search result and the second search result based on a highest rank combination of the first composite score and the second composite score; and

displaying the first search result and second search result according to the ranking on the display device.

2. The method of claim 1 , further comprising:

determining, by the media guidance application on the display device, that metadata for a third search result indicates that the third search result corresponds to a third language; and

in response to determining that metadata for a third search result indicates that the third search result corresponds to the third language, generating, by the media guidance application on the display device, a third language model for the third language for determining search results.

3. The method of claim 2 , wherein the third language model is generated, by the media guidance application on the display device, after receiving confirmation apply the third language model for determining search results, and wherein a prompt for receiving the confirmation is generated in response to determining that the metadata for the third search result indicates that the third search result corresponds to the third language.

4. The method of claim 3 , wherein generating the third language model further comprises:

accessing, by the media guidance application on the display device, a language setting, the language setting including a setting language;

comparing, by the media guidance application on the display device, the setting language to the third language; and

determining, by the media guidance application on the display device, that the third language does not correspond to the setting language.

5. The method of claim 4 , further comprising:

receiving, by the media guidance application on the display device, a second voice query;

identifying, by the media guidance application on the display device, a third audio segment and a fourth audio segment based on audible breaks in the second voice query;

determining, by the media guidance application on the display device, a first resolved word in the third language and a second resolved word in the third language by combining a third set of phonemes to the third audio segment and fourth audio segment, respectively, the third set of phonemes selected from phonemes of the third language model; and

generating, by the media guidance application on the display device, a third text query including the first resolved word in the third language and the second resolved word in the third language.

6. The method of claim 1 , further comprising:

determining, by the media guidance application on the display device, a plurality of potential text queries, using the first language model, by applying different combinations of the first set of phonemes to the first audio segment and second audio segment, respectively;

ranking, by the media guidance application on the display device, the potential text queries according to their respective composite scores; and

identifying, by the media guidance application on the display device, a highest ranked text query from the potential text queries based on the first composite score having a highest ranking of the respective composite scores.

7. The method of claim 6 , further comprising generating, by the media guidance application on the display device, a prompt to select the highest ranked text query if multiple text queries of the potential text queries share the highest ranking of the respective composite scores.

8. The method of claim 7 , further comprising:

accessing, by the media guidance application on the display device, a composite score threshold; and

providing, by the media guidance application on the display device, the highest ranked search result for the first text query in response to determining that a corresponding composite score is above the composite score threshold score.

9. The method of claim 8 , further comprising:

accessing, by the media guidance application on the display device, first metadata for the highest ranked search result;

automatically determining, by the media guidance application on the display device, that the first metadata for the highest ranked search result indicates that the highest ranked search result does not correspond to the first language; and

not provide the highest ranked search result for the first text query in response to determining that the first metadata for the highest ranked search result indicates that the highest ranked search result does not correspond to a first language.

10. The method of claim 1 , further comprising:

determining, by the media guidance application on the display device, for a third text query, the first resolved word in the first language and the second resolved word in the second language by combining the first set of phonemes to the first audio segment and the second set of phonemes to the second audio segment;

determining, by the media guidance application on the display device, a numerical ranking for the first resolved word in the first language and a numerical ranking for the second resolved word in the second language based on a likelihood of usage of the first resolved word in the first language and the second resolved word in the second language; and

determining, the media guidance application on the display device, a composite score for the third text query based on summing the numerical ranking for the first resolved word in the first language and the numerical ranking for the second resolved word in the second language.

11. A system for generating results in multiple languages for search queries, comprising:

memory configured to store a plurality of language models;

first input/output circuitry configured to:

receive, by a media guidance application on the display device, a first voice query;

access, by the media guidance application on the display device, a first language model and a second language model and

processing circuitry configured to:

identify, by the media guidance application on a display device, a first audio segment and a second audio segment based on audible breaks in the first voice query;

determine, by the media guidance application on the display device, a first resolved word in a first language and a second resolved word in the first language by combining a first set of phonemes to the first audio segment and second audio segment, respectively, the first set of phonemes selected from phonemes of the first language model;

generate, by the media guidance application on the display device, a first text query including the first resolved word in the first language and the second resolved word in the first language;

determine, by the media guidance application on the display device, a first numerical ranking for the first resolved word in the first language and a second numerical ranking for the second resolved word in the first language based on a likelihood of usage of the first resolved word in the first language and the second resolved word in the first language;

determine, by the media guidance application on the display device, a first composite score for the first text query based on summing the first numerical ranking and the second numerical ranking;

determine, by the media guidance application on the display device, a first resolved word in a second language and a second resolved word in the second language by combining a second set of phonemes to the first audio segment and second audio segment, respectively, the second set of phonemes selected from phonemes of the second language model;

generate, by the media guidance application on the display device, a second text query including the first resolved word in the second language and the second resolved word in the second language;

determine, by the media guidance application on the display device, a third numerical ranking for the first resolved word in the second language and a fourth numerical ranking for the second resolved word in the second language based on a likelihood of usage of the first resolved word in the second language and the second resolved word in the second language;

determine, by the media guidance application on the display device, a second composite score for the second text query based on summing the third numerical ranking and the fourth numerical ranking;

generate, by the media guidance application on the display device, a first search result for the first text query and a second search result for the second text query;

rank, by the media guidance application on the display device, the first search result and the second search result based on a highest rank combination of the first composite score and the second composite score; and

second input/output circuitry configured to display the first search result and second search result according to the ranking on the display device.

12. The system of claim 11 , wherein the processing circuitry is further configured to:

determine, by the media guidance application on the display device, that metadata for a third search result indicates that the third search result corresponds to a third language; and

in response to determining that metadata for a third search result indicates that the third search result corresponds to the third language, generate, by the media guidance application on the display device, a third language model for the third language for determining search results.

13. The system of claim 12 , wherein the third language model is generated, by the media guidance application on the display device, after receiving confirmation to apply the third language model for determining search results, and wherein a prompt for receiving the confirmation is generated in response to determining that the metadata for the third search result indicates that the third search result corresponds to the third language.

14. The system of claim 13 , wherein the processing circuitry is further configured to:

access, by the media guidance application on the display device, a language setting, the language setting including a setting language;

compare, by the media guidance application on the display device, the setting language to the third language; and

determine, by the media guidance application on the display device, that the third language does not correspond to the setting language.

15. The system of claim 14 , wherein

the first input/output circuitry is further configured to:

receive, by the media guidance application on the display device, a second voice query; and

processing circuitry is further configured to:

identify, by the media guidance application on the display device, a third audio segment and a fourth audio segment based on audible breaks in the second voice query;

determine, by the media guidance application on the display device, a first resolved word in the third language and a second resolved word in the third language by combining a third set of phonemes to the third audio segment and fourth audio segment, respectively, the third set of phonemes selected from phonemes of the third language model; and

generating, by the media guidance application on the display device, a third text query including the first resolved word in the third language and the second resolved word in the third language.

16. The system of claim 11 , wherein the processing circuitry is further configured to:

determine, by the media guidance application on the display device, a plurality of potential text queries, using the first language model, by applying different combinations of the first set of phonemes to the first audio segment and second audio segment, respectively;

rank, by the media guidance application on the display device, the potential text queries according to their respective composite scores; and

identify, by the media guidance application on the display device, a highest ranked text query from the potential text queries based on the first composite score having a highest ranking of the respective composite scores.

17. The system of claim 16 , wherein the processing circuitry is further configured to generating, by the media guidance application on the display device, a prompt to select the highest ranked text query if multiple text queries of the potential text queries share the highest ranking of the respective composite scores.

18. The system of claim 17 , wherein the processing circuitry is further configured to access, by the media guidance application on the display device, a composite score threshold; and

the second input/output circuitry is further configured to provide, by the media guidance application on the display device, the highest ranked search result for the first text query in response to determining that a corresponding composite score is above the composite score threshold score.

19. The system of claim 18 , wherein

the processing circuitry is further configured to:

access, by the media guidance application on the display device, first metadata for the highest ranked search result;

automatically determine, by the media guidance application on the display device, that the first metadata for the highest ranked search result indicates that the highest ranked search result does not correspond to the first language; and

the second input/output circuitry is further configured to not provide the highest ranked search result for the first text query in response to determining that the first metadata for the highest ranked search result indicates that the highest ranked search result does not correspond to a first language.

20. The system of claim 11 , wherein the processing circuitry is further configured to:

determine, by the media guidance application on the display device, for a third text query, the first resolved word in the first language and the second resolved word in the second language by combining the first set of phonemes to the first audio segment and the second set of phonemes to the second audio segment;

determine, by the media guidance application on the display device, a numerical ranking for the first resolved word in the first language and a numerical ranking for the second resolved word in the second language based on a likelihood of usage of the first resolved word in the first language and the second resolved word in the second language; and

determine, by the media guidance application on the display device, a composite score for the third text query based on summing the numerical ranking for the first resolved word in the first language and the numerical ranking for the second resolved word in the second language.

Assignments (3)
CHANGE OF NAME Recorded Oct 23, 2022
From: ROVI GUIDES, INC.
To: ADEIA GUIDES INC.
Reel/Frame 061746/0981 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 23, 2022
From: ADEIA GUIDES INC.
To: ROVI PRODUCT CORPORATION
Reel/Frame 061747/0001 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 13, 2020
From: SREEDHARA, ARUN
To: ROVI GUIDES, INC.
Reel/Frame 053187/0315 →
Continuity (2)
Continuation 15720975 · Sep 29, 2017
Related Publication 20200342034A1 · Oct 29, 2020