IP Library Granted Patent US 12,499,152
Granted Patent B1
US 12,499,152 · App. 18/744,027 · Granted Dec 16, 2025

Query modification based on non-textual resource context

Inventors: Gokhan H. Bakir (Zurich, CH); Behshad Behzadi (Freienbach, CH)
Assignee: GOOGLE LLC
G06F16/5846G06F16/248G06F16/3322G06F16/3325G06F16/951G06F16/9535
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,499,152
App. No.
18/744,027
Granted
Dec 16, 2025
Kind
B1
Abstract

Methods, systems, and apparatus, including computer programs encoded on a computer storage medium, modifying queries based on non-textual content. In one aspect, a method includes receiving, from a user device, a query including a plurality of terms; determining active non-textual data displayed in an application environment on the user device; determining, from the non-textual textual data, modification data for the query; generating a set of modified queries based on the query and the modification parameters; scoring the modified queries according to one or more scoring criteria; selecting one of the modified queries based on the scoring; and providing, to the user device, search results responsive to the selected modified query.

Claims (46)

1 . A method implemented by one or more processors, the method comprising:

receiving a plurality of terms spoken by a user of a user device, the user device displaying a plurality of non-textual data elements in an application environment when the user spoke the plurality of terms;

determining that the plurality of terms is an ambiguous instruction, and in response:

identifying, as an active non-textual data element, one of the non-textual data elements that has been selected by the user;

automatically determining, by the data processing apparatus, for the active non-textual data element, modification data from the active non-textual data element for the ambiguous instruction;

automatically generating, by the data processing apparatus, a modified instruction based on the ambiguous instruction and the modification data; and

providing, to the user device, the modified instruction, wherein the modified instruction causes the user device to perform an action defined by the modified instruction.

2 . The method of claim 1 , wherein the active non-textual data element comprises an image that is selected by the user.

3 . The method of claim 1 , wherein the active non-textual data element comprises a video that is selected by the user.

4 . The method of claim 1 , wherein the active non-textual data element comprises an image that is captured by a camera of the user device.

5 . The method of claim 1 , wherein determining modification data for the ambiguous instruction comprises:

determining, from the active non-textual data element, one or more labels that describe subject matter of the active non-textual data element; and

for each label, generating a modified instruction based on the ambiguous instruction and the label.

6 . The method of claim 5 , wherein generating a modified instruction based on the ambiguous instruction and the label comprises concatenating the ambiguous instruction with the label.

7 . The method of claim 1 , wherein determining modification data for the ambiguous instruction comprises:

determining, from the active non-textual data element, entity text describing one or more entities associated with the active non-textual data element; and

for each entity, generating a modified instruction based on the ambiguous instruction and the entity.

8 . The method of claim 7 , wherein generating a modified instruction based on the ambiguous instruction and the entity text describing the entity comprises revising one or more terms of the ambiguous instruction based on the entity text describing the entity.

9 . The method of claim 1 , wherein receiving a plurality of terms spoken by a user of a user device comprises receiving text generated by a speech analyzer.

10 . A system comprising one or more processors and memory storing instructions that, in response to execution by the one or more processors, cause the one or more processors to:

receive a plurality of terms spoken by a user of a user device, the user device displaying a plurality of non-textual data elements in an application environment when the user spoke the plurality of terms;

determine that the plurality of terms is an ambiguous instruction, and in response:

identify, as an active non-textual data element, one of the non-textual data elements that has been selected by the user;

automatically determine, by the data processing apparatus, for the active non-textual data element, modification data from the active non-textual data element for the ambiguous instruction;

automatically generate, by the data processing apparatus, a modified instruction based on the ambiguous instruction and the modification data; and

provide, to the user device, the modified instruction, wherein the modified instruction causes the user device to perform an action defined by the modified instruction.

11 . The system of claim 10 , wherein the active non-textual data element comprises an image that is selected by the user.

12 . The system of claim 10 , wherein the active non-textual data element comprises a video that is selected by the user.

13 . The system of claim 10 , wherein the active non-textual data element comprises an image that is captured by a camera of the user device.

14 . The system of claim 10 , wherein the instructions to determine modification data for the ambiguous instruction comprise instructions to:

determine, from the active non-textual data element, one or more labels that describe subject matter of the active non-textual data element; and

for each label, generate a modified instruction based on the ambiguous instruction and the label.

15 . The system of claim 14 , wherein the instructions to generate a modified instruction based on the ambiguous instruction and the label comprise instructions to concatenate the ambiguous instruction with the label.

16 . The system of claim 10 , wherein the instructions to determine modification data for the ambiguous instruction comprise instructions to:

determine, from the active non-textual data element, entity text describing one or more entities associated with the active non-textual data element; and

for each entity, generate a modified instruction based on the ambiguous instruction and the entity.

17 . The system of claim 16 , wherein the instructions to generate a modified instruction based on the ambiguous instruction and the entity text describing the entity comprise instructions to revise one or more terms of the ambiguous instruction based on the entity text describing the entity.

18 . The system of claim 10 , wherein the instructions to receive a plurality of terms spoken by a user of a user device comprise instructions to receive text generated by a speech analyzer.

19 . At least one non-transitory computer-readable medium comprising instructions that, in response to execution by one or more processors, cause the one or more processors to:

receive a plurality of terms spoken by a user of a user device, the user device displaying a plurality of non-textual data elements in an application environment when the user spoke the plurality of terms;

determine that the plurality of terms is an ambiguous instruction, and in response:

identify, as an active non-textual data element, one of the non-textual data elements that has been selected by the user;

automatically determine, by the data processing apparatus, for the active non-textual data element, modification data from the active non-textual data element for the ambiguous instruction;

automatically generate, by the data processing apparatus, a modified instruction based on the ambiguous instruction and the modification data; and

provide, to the user device, the modified instruction, wherein the modified instruction causes the user device to perform an action defined by the modified instruction.

20 . The at least one non-transitory computer-readable medium of claim 19 , wherein the active non-textual data element comprises an image that is selected by the user.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 14, 2024
From: BAKIR, GOKHAN H.; BEHZADI, BEHSHAD
To: GOOGLE INC.
Reel/Frame 068284/0526 →
CHANGE OF NAME Recorded Aug 14, 2024
From: GOOGLE INC.
To: GOOGLE LLC
Reel/Frame 068606/0030 →
Continuity (4)
Continuation 18109226 · Feb 13, 2023
Continuation 16785241 · Feb 7, 2020
Continuation 15791079 · Oct 23, 2017
Continuation 14313559 · Jun 24, 2014
References Cited (63)
US 7389181B2 · Meadow et al. · 2008 [cited by applicant]
US 7689613B2 · Candelore · 2010 [cited by applicant]
US 7788266B2 · Venkataraman et al. · 2010 [cited by applicant]
US 8126897B2 · Sznajder · 2012 [cited by applicant]
US 8316019B1 · Ainslie et al. · 2012 [cited by applicant]
US 8321406B2 · Garg et al. · 2012 [cited by applicant]
US 8391618B1 · Chuang et al. · 2013 [cited by applicant]
US 8392435B1 · Yamauchi · 2013 [cited by applicant]
US 8452794B2 · Yang · 2013 [cited by applicant]
US 8515185B2 · Lee et al. · 2013 [cited by applicant]
US 8521764B2 · Pfleger · 2013 [cited by applicant]
US 8606781B2 · Chi et al. · 2013 [cited by applicant]
US 8898095B2 · Agrawal et al. · 2014 [cited by applicant]
US 8977639B2 · Petrou et al. · 2015 [cited by applicant]
US 9116952B1 · Heiler et al. · 2015 [cited by applicant]
US 9135305B2 · Taubman et al. · 2015 [cited by applicant]
US 9830391B1 · Bakir · 2017 [cited by examiner]
US 10303719B1 · Tan · 2019 [cited by examiner]
US 10331330B2 · Petrov · 2019 [cited by examiner]
US 10489410B2 · Sharifi · 2019 [cited by examiner]
US 10592571B1 · Bakir · 2020 [cited by examiner]
US 11354358B1 · Tan · 2022 [cited by examiner]
US 11580181B1 · Bakir · 2023 [cited by examiner]
US 12026194B1 · Bakir · 2024 [cited by examiner]
US 12124606B2 · Gaddam · 2024 [cited by examiner]
US 12230261B2 · Samal · 2025 [cited by examiner]
US 20010053968A1 · Galitsky et al. · 2001 [cited by applicant]
US 20040267730A1 · Dumais et al. · 2004 [cited by applicant]
US 20070060114A1 · Ramer et al. · 2007 [cited by applicant]
US 20070071320A1 · Yada · 2007 [cited by applicant]
US 20070118357A1 · Kasravi · 2007 [cited by applicant]
US 20070140595A1 · Taylor et al. · 2007 [cited by applicant]
US 20070214131A1 · Cucerzan et al. · 2007 [cited by applicant]
US 20080046405A1 · Olds et al. · 2008 [cited by applicant]
US 20080270110A1 · Yurick et al. · 2008 [cited by applicant]
US 20100205202A1 · Yang · 2010 [cited by applicant]
US 20100306249A1 · Hill et al. · 2010 [cited by applicant]
US 20100318532A1 · Sznajder · 2010 [cited by applicant]
US 20110035406A1 · Petrou et al. · 2011 [cited by applicant]
US 20110038512A1 · Petrou et al. · 2011 [cited by applicant]
US 20110125735A1 · Petrou · 2011 [cited by applicant]
US 20110128288A1 · Petrou et al. · 2011 [cited by applicant]
US 20110137895A1 · Petrou et al. · 2011 [cited by applicant]
US 20120109858A1 · Makadia et al. · 2012 [cited by applicant]
US 20120191745A1 · Velipasaoglu et al. · 2012 [cited by applicant]
US 20120215533A1 · Aravamudan et al. · 2012 [cited by applicant]
US 20130132361A1 · Chen et al. · 2013 [cited by applicant]
US 20130346400A1 · Ramsey et al. · 2013 [cited by applicant]
US 20140046935A1 · Bengio et al. · 2014 [cited by applicant]
US 20140172881A1 · Petrou et al. · 2014 [cited by applicant]
US 20150058318A1 · Blackwell et al. · 2015 [cited by applicant]
US 20210019026A1 · Pagaime Da Silva · 2021 [cited by examiner]
US 20210064674A1 · Alexeev · 2021 [cited by examiner]
US 20210103337A1 · Jeppsson · 2021 [cited by examiner]
US 20210103348A1 · Jeppsson · 2021 [cited by examiner]
US 20220327234A1 · Gaddam · 2022 [cited by examiner]
US 20230020743A1 · Alexeev · 2023 [cited by examiner]
Combining textual and non-textual features for e-mail importance estimation (Year: 2013). [cited by examiner]
Searching and Classifying Non-Textual Information (Year: 2004). [cited by examiner]
10 Mobile Astronomy Apps for Stargazers, [online] [Retrieved on Apr. 29, 2014]; Retrieved from the Internet URL: http://mashable.com/2011/06/21/astronomy-mobile-apps/; 13 pages; dated 2013. [cited by applicant]
Google announces Search by Image and Voice Search for desktop, revamped mobile search, [online] Retrieved on Apr. 29, 2014]; Retrieved from the Internet URL: http://www.engadget.com/2011/06/14/google-announces-search-by… [cited by applicant]
Google's Impressive “Conversational Search” Goes Live On Chrome, [online] [Retrieved on May 5, 2014]; Retrieved from the Internet URL: http://searchengineland.com/googles-impressive-conversational-search-goes-line-on-ch… [cited by applicant]
Zhang et al., “Probabilistic Query Rewriting for Efficient and Effective Keyword Search on Graph Data,” Proceedings of the VLDB Endowment 6(14):1642-1653; 12 pages; dated 2013. [cited by applicant]