IP Library Granted Patent US 11,908,187
Granted Patent B2
US 11,908,187 · App. 18/117,798 · Granted Feb 20, 2024

Systems, methods, and apparatus for providing image shortcuts for an assistant application

Inventors: Marcin Nowak-Przygodzki (Bäch, CH); Gökhan Bakir (Zurich, CH)
Assignee: GOOGLE LLC
G06V20/20G06F3/005G06F3/017G06F3/0304G06F3/0481G06F3/167G06F9/453G06F16/5866G06F16/9032H04N23/63
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,908,187
App. No.
18/117,798
Granted
Feb 20, 2024
Kind
B2
Abstract

Methods, apparatus, systems, and computer-readable media are set forth for generating and/or utilizing image shortcuts that cause one or more corresponding computer actions to be performed in response to determining that one or more features are present in image(s) from a camera of a computing device of a user (e.g., present in a real-time image feed from the camera). An image shortcut can be generated in response to user interface input, such as a spoken command. For example, the user interface input can direct the automated assistant to perform one or more actions in response to object(s) having certain feature(s) being present in a field of view of the camera. Subsequently, when the user directs their camera at object(s) having such feature(s), the assistant application can cause the action(s) to be automatically performed. For example, the assistant application can cause data to be presented and/or can control a remote device in accordance with the image shortcut.

Claims (35)

1. A method implemented by one or more processors, the method comprising:

receiving audio data in the form of a verbal command corresponding to a request to create an image shortcut setting;

identifying, from the audio data, one or more computer actions to be performed, a condition for the image shortcut setting, and an object which is the subject of the condition;

generating the image shortcut setting based on the one or more actions, the condition, and the object identified from the audio data, wherein the image shortcut setting is configured to cause the one or more computer actions to be performed in response to identifying the object from image data from a real-time image feed from a camera of a computing device when the condition is satisfied; and

causing, according to the image shortcut setting, the one or more computer actions to be performed in response to identifying the object from the image data when the condition is satisfied, wherein identifying the object from the image data includes processing the image data using one or more image processing techniques.

2. The method of claim 1 , wherein the one or more computer actions comprise transmitting a command to at least one peripheral device, wherein the command causes a state of the at least one peripheral device to be altered.

3. The method of claim 1 , wherein identifying the object identifier corresponding to the object includes identifying multiple object identifiers corresponding to multiple different objects at which the camera of the computing device is directed, wherein the image shortcut setting is based on the multiple object identifiers.

4. The method of claim 1 , further comprising:

identifying, from the audio data, a context identifier for the request, wherein the image shortcut setting is generated further based on the context identifier.

5. The method of claim 4 , wherein the context identifier identifies a location at which the real-time image feed is provided by the camera.

6. The method of claim 4 , wherein the context identifier identifies at least one time or at least one location, and wherein causing, according to the image shortcut setting, the one or more computer actions to be performed is further in response to the image data being provided at a time that matches the at least one time, or at a location that matches the at least one location.

7. The method of claim 6 , wherein the context identifier identifies the at least one location.

8. A system, comprising:

a camera;

a microphone;

a display device;

a speaker;

one or more processors in communication with the camera, the microphone, the display device, and the speaker; and

memory configured to store instructions that, when executed by the one or more processors, cause the one or more processors to:

receive, via the microphone, audio data in the form of a verbal command corresponding to a request to create an image shortcut setting;

identify, from the audio data, one or more computer actions to be performed, a condition for the image shortcut setting, and an object which is the subject of the condition;

generate the image shortcut setting based on the one or more actions, the condition, and the object identified from the audio data, wherein the image shortcut setting is configured to cause the one or more computer actions to be performed in response to identifying the object from image data from a real-time image feed from the camera when the condition is satisfied; and

cause, according to the image shortcut setting, the one or more computer actions to be performed in response to identifying the object from the image data when the condition is satisfied, wherein identifying the object from the image data includes processing the image data using one or more image processing techniques.

9. The system of claim 8 , wherein the one or more computer actions comprise transmitting a command to at least one peripheral device, wherein the command causes a state of the at least one peripheral device to be altered.

10. The system of claim 8 , wherein in identifying the object identifier corresponding to the object one or more of the processors are to identify multiple object identifiers corresponding to multiple different objects at which the camera of the computing device is directed, wherein the image shortcut setting is based on the multiple object identifiers.

11. The system of claim 8 , wherein the instructions, when executed by the one or more processors, further cause one or more of the processors to:

identify, from the audio data, a context identifier for the request, wherein the image shortcut setting is generated further based on the context identifier.

12. The system of claim 11 , wherein the context identifier identifies a location at which the real-time image feed is provided by the camera.

13. The system of claim 11 , wherein the context identifier identifies at least one time or at least one location, and wherein causing, according to the image shortcut setting, the one or more computer actions to be performed is further in response to the image data being provided at a time that matches the at least one time, or at a location that matches the at least one location.

14. The system of claim 13 , wherein the context identifier identifies the at least one location.

15. At least one non-transitory computer readable medium configured to store instructions that, when executed by one or more processors, cause the one or more processors to perform a method comprising:

receiving audio data in the form of a verbal command corresponding to a request to create an image shortcut setting;

identifying, from the audio data, one or more computer actions to be performed, a condition for the image shortcut setting, and an object which is the subject of the condition;

generating the image shortcut setting based on the one or more actions, the condition, and the object identified from the audio data, wherein the image shortcut setting is configured to cause the one or more computer actions to be performed in response to identifying the object from image data from a real-time image feed from a camera of a computing device when the condition is satisfied; and

causing, according to the image shortcut setting, the one or more computer actions to be performed in response to identifying the object from the image data when the condition is satisfied, wherein identifying the object from the image data includes processing the image data using one or more image processing techniques.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 5, 2023
From: NOWAK-PRZYGODZKI, MARCIN; BAKIR, GÖKHAN
To: GOOGLE INC.
Reel/Frame 063546/0140 →
CHANGE OF NAME Recorded May 5, 2023
From: GOOGLE INC.
To: GOOGLE LLC
Reel/Frame 063557/0556 →
Continuity (5)
Continuation 17838914 · Jun 13, 2022
Continuation 16850294 · Apr 16, 2020
Continuation 16459869 · Jul 2, 2019
Continuation 15700104 · Sep 9, 2017
Related Publication 20230206628A1 · Jun 29, 2023