IP Library Granted Patent US 10,893,202
Granted Patent B2
US 10,893,202 · App. 16/585,085 · Granted Jan 12, 2021

Storing metadata related to captured images

Inventors: Ibrahim Badr (Zurich, CH); Gökhan Bakir (Zurich, CH); Daniel Kunkle (Boston, MA); Kavin Karthik Ilangovan (Zurich, CH); Denis Burakov (Zurich, CH)
Assignee: GOOGLE LLC
H04N5/232935G06F16/5846G06F16/5866G06K9/00671G06K9/00684H04N1/00244H04N1/00331H04N1/2187H04N1/2191H04N5/23222H04N5/232933G06K2209/01H04N2101/00H04N2201/0084H04N2201/0096H04N2201/3266
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,893,202
App. No.
16/585,085
Granted
Jan 12, 2021
Kind
B2
Abstract

The present disclosure relates to user-selected metadata related to images captured by a camera of a client device. User-selected metadata may include contextual information and/or information provided by a user when the images are captured. In various implementations, a free form input may be received at a first client device of one or more client devices operated by a user. A task request may be recognized from the free form input, and it may be determined that the task request includes a request to store metadata related to one or more images captured by a camera of the first client device. The metadata may be selected based on content of the task request. The metadata may then be stored, e.g., in association with one or more images captured by the camera, in computer-readable media. The computer-readable media may be searchable by the metadata.

Claims (40)

1. A method implemented using one or more processors, comprising:

streaming data captured by one or more cameras to a camera application active on a first client device of one or more client devices operated by a user;

invoking an automated assistant at least partially based on the camera application being active on the first client device;

performing image recognition analysis on the data captured by one or more of the cameras to detect a vehicle;

in response to detection of the vehicle, provide to the user, as output from the automated assistant, a suggested task request to remember a parking location associated with the depicted vehicle;

receiving, at the first client device while the data captured by the one or more cameras is streamed to the camera application, confirmation from the user to perform the suggested task request; and

storing metadata indicative of the parking location in one or more computer-readable mediums, wherein the one or more computer-readable mediums are searchable by the automated assistant using the metadata.

2. The method of claim 1 , wherein the method further comprises:

receiving, at the first client device or a second client device of the one or more client devices, a free form input;

recognizing another task request from the free form input;

determining that the metadata indicative of the parking location is responsive to the another task request; and

in response to determining that the metadata is responsive to the another task request, providing, as output via one or more output devices of the first or second client device, content indicative of the metadata.

3. The method of claim 1 , further comprising performing optical character recognition on a portion of the data captured by one or more of the cameras to determine textual content depicted in the data captured by one or more of the cameras.

4. The method of claim 3 , wherein the metadata further includes at least some of the textual content.

5. The method of claim 1 , wherein the metadata includes at least some of a content of the suggested task request.

6. The method of claim 1 , wherein the metadata includes a position coordinate obtained simultaneously with capture of the data captured by one or more of the cameras.

7. A system comprising:

one or more processors;

one or more cameras operably coupled with the one or more processors;

a microphone operably coupled with one or more of the processors; and

memory storing instructions that, in response to execution of the instructions by one or more of the processors, cause one or more of the processors to operate a camera application and at least a portion of an automated assistant, wherein the automated assistant is invoked at least in part based on the camera application, and the one or more processors are to:

perform image recognition analysis on data captured by one or more of the cameras to detect a vehicle;

in response to detection of the vehicle, cause the automated assistant to provide a suggested task request to remember a parking location associated with the depicted vehicle;

receive confirmation input from a user to perform the suggested task request; and

store metadata indicative of the parking location in one or more computer-readable mediums, wherein the one or more computer-readable mediums are searchable by the automated assistant using the metadata.

8. The system of claim 7 , wherein the automated assistant is further to:

receive a free form input;

recognize another task request from the free form input;

determine that the metadata related to the data captured by one or more of the cameras is responsive to the another task request; and

in response to determining that the metadata is responsive to the another task request, providing, as output via one or more output devices, content indicative of the metadata.

9. The system of claim 7 , wherein one or more of the processors are to perform optical character recognition on a portion of the data captured by one or more of the cameras to determine textual content depicted in the data captured by one or more of the cameras.

10. The system of claim 9 , wherein the metadata further includes at least some of the textual content.

11. The system of claim 7 , wherein the metadata includes at least some of the content of the suggested task request.

12. At least one non-transitory computer-readable medium comprising instructions that, in response to execution of the instructions by one or more processors, cause the one or more processors to perform the following operations:

streaming data captured by one or more cameras to a camera application active on a first client device of one or more client devices operated by a user;

invoking an automated assistant at least partially based on the camera application being active on the first client device;

performing image recognition analysis on data captured by one or more of the cameras to detect a vehicle;

in response detection of the vehicle, provide to the user, as output from the automated assistant, a suggested task request to remember a parking location associated with the depicted vehicle;

receiving, at the first client device, confirmation a free form input from the user to perform the suggested task request; and

storing metadata indicative of the parking location in one or more computer-readable mediums, wherein the one or more computer-readable mediums are searchable by the automated assistant using the metadata.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 6, 2019
From: BADR, IBRAHIM; BAKIR, GÖKHAN; KUNKLE, DANIEL; ILANGOVAN, KAVIN KARTHIK; BURAKOV, DENIS
To: GOOGLE INC.
Reel/Frame 050633/0963 →
CHANGE OF NAME Recorded Oct 6, 2019
From: GOOGLE INC.
To: GOOGLE LLC
Reel/Frame 050638/0914 →
Continuity (3)
Continuation 15602661 · May 23, 2017
Provisional Application 62507108 · May 16, 2017
Related Publication 20200021740A1 · Jan 16, 2020
Cited By (3)
US 12,192,426 US 12,513,254 US 12,621,557