IP Library Granted Patent US 10,469,755
Granted Patent B2
US 10,469,755 · App. 15/602,661 · Granted Nov 5, 2019

Storing metadata related to captured images

Inventors: Ibrahim Badr (Zurich, CH); Gökhan Bakir (Zurich, CH); Daniel Kunkle (Boston, MA); Kavin Karthik Ilangovan (Zurich, CH); Denis Burakov (Zurich, CH)
Assignee: GOOGLE LLC
H04N5/23293G06F16/5846G06F16/5866G06K9/00671G06K9/325H04N1/00H04N5/23222H04N5/232933H04N5/232935H04N5/772H04N9/8205G06K9/4628G06K9/6273G06K2209/01
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,469,755
App. No.
15/602,661
Filed
May 23, 2017
Granted
Nov 5, 2019
Kind
B2
Art Unit
2696
USPC
348/231.5
Abstract

The present disclosure relates to user-selected metadata related to images captured by a camera of a client device. User-selected metadata may include contextual information and/or information provided by a user when the images are captured. In various implementations, a free form input may be received at a first client device of one or more client devices operated by a user. A task request may be recognized from the free form input, and it may be determined that the task request includes a request to store metadata related to one or more images captured by a camera of the first client device. The metadata may be selected based on content of the task request. The metadata may then be stored, e.g., in association with one or more images captured by the camera, in computer-readable media. The computer-readable media may be searchable by the metadata.

Claims (51)

1. A method comprising:

streaming data captured by one or more cameras to an electronic viewfinder of a first client device of one or more client devices operated by a user;

invoking an automated assistant at least partially in response to the streaming;

receiving, at the first client device while the data captured by the camera is streamed to the electronic viewfinder, a free form input from the user directed at the automated assistant;

recognizing a task request from the free form input;

determining that the task request comprises a request for the automated assistant to store metadata related to one or more images captured by one or more of the cameras, wherein the metadata is selected based on content of the task request; and

storing the metadata in one or more computer-readable mediums, wherein the one or more computer-readable mediums are searchable by the automated assistant using the metadata.

2. The method of claim 1 , wherein the free-form input is a first input, and the method further comprises:

receiving, at the first client device or a second client device of the one or more client devices, a second free form input;

recognizing another task request from the second free form input;

determining that the metadata related to the one or more images captured by the camera is responsive to the another task request; and

in response to determining that the metadata is responsive to the another task request, providing, as output via one or more output devices of the first or second client device, content indicative of the metadata.

3. The method of claim 1 , further comprising providing, as output via one or more output devices of the first client device, the task request as a suggestion to the user, wherein the task request is selected based on one or more signals generated by one or more sensors of the first client device.

4. The method of claim 3 , wherein the one or more signals include data captured by the camera.

5. The method of claim 3 , wherein the one or more signals include position coordinate data from a position coordinate sensor.

6. The method of claim 1 , further comprising:

performing image processing on the one or more images;

based on the image processing, identifying an object depicted in the one or more images; and

storing the metadata in association with another stored image that depicts the same object or another object sharing one or more attributes with the object.

7. The method of claim 1 , further comprising performing optical character recognition on a portion of the one or more images to determine textual content depicted in the one or more images.

8. The method of claim 7 , wherein the metadata further includes at least some of the textual content.

9. The method of claim 1 , wherein the metadata includes at least some of the content of the task request.

10. The method of claim 1 , wherein the metadata includes a position coordinate obtained simultaneously with capture of the one or more images.

11. At least one non-transitory computer-readable medium comprising instructions that, in response to execution of the instructions by one or more processors, cause the one or more processors to perform the following operations:

streaming data captured by one or more cameras to an electronic viewfinder of a first client device of one or more client devices operated by a user;

invoking an automated assistant at least partially in response to the streaming;

receiving, at the first client device while the data captured by the camera is streamed to the electronic viewfinder, a free form input from the user directed at the automated assistant;

recognizing a task request from the free form input;

determining that the task request comprises a request for the automated assistant to store metadata related to one or more images captured by one or more of the cameras, wherein the metadata is selected based on content of the task request; and

storing the metadata in one or more computer-readable mediums, wherein the one or more computer-readable mediums are searchable by the automated assistant using the metadata.

12. The method of claim 1 , wherein invoking the automated assistant is further based on an invocation phrase that is activated to invoke the automated assistant during the streaming.

13. A system comprising one or more processors and memory storing instructions that, in response to execution of the instructions by the one or more processors, cause the one or more processors to:

stream data captured by one or more cameras to an electronic viewfinder of a first client device of one or more client devices operated by a user;

invoke an automated assistant at least partially in response to the streaming;

receive, at the first client device while the data captured by the camera is streamed to the electronic viewfinder, a free form input from the user directed at the automated assistant;

recognize a task request from the free form input;

determine that the task request comprises a request for the automated assistant to store metadata related to one or more images captured by one or more of the cameras, wherein the metadata is selected based on content of the task request; and

store the metadata in one or more computer-readable mediums, wherein the one or more computer-readable mediums are searchable by the automated assistant using the metadata.

14. The system of claim 13 , wherein the free-form input is a first input, and the system further comprises instructions to:

receive, at the first client device or a second client device of the one or more client devices, a second free form input;

recognize another task request from the second free form input;

determine that the metadata related to the one or more images captured by the camera is responsive to the another task request; and

in response to the determination that the metadata is responsive to the another task request, provide, as output via one or more output devices of the first or second client device, content indicative of the metadata.

15. The system of claim 13 , further comprising providing, as output via one or more output devices of the first client device, the task request as a suggestion to the user, wherein the task request is selected based on one or more signals generated by one or more sensors of the first client device.

16. The system of claim 15 , wherein the one or more signals include data captured by the camera.

17. The system of claim 15 , wherein the one or more signals include position coordinate data from a position coordinate sensor.

18. The system of claim 13 , further comprising instructions to:

perform image processing on the one or more images;

based on the image processing, identify an object depicted in the one or more images; and

store the metadata in association with another stored image that depicts the same object or another object sharing one or more attributes with the object.

19. The system of claim 13 , further comprising instructions to perform optical character recognition on a portion of the one or more images to determine textual content depicted in the one or more images.

Assignments (2)
CHANGE OF NAME Recorded Oct 20, 2017
From: GOOGLE INC.
To: GOOGLE LLC
Reel/Frame 044567/0001 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 23, 2017
From: BADR, IBRAHIM; BAKIR, GÖKHAN; KUNKLE, DANIEL; ILANGOVAN, KAVIN KARTHIK; BURAKOV, DENIS
To: GOOGLE INC.
Reel/Frame 042482/0521 →
Continuity (2)
Provisional Application 62507108 · May 16, 2017
Related Publication 20180338109A1 · Nov 22, 2018