IP Library Granted Patent US 12,380,713
Granted Patent B2
US 12,380,713 · App. 17/932,152 · Granted Aug 5, 2025

System and method for smart recipe generation

Inventor: Omar Estrada Diaz (Sacramento, CA)
Assignee: Google LLC
G06V20/68G06V10/82G06V20/50G06V20/70G06V40/20G06F3/011G06F3/14
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,380,713
App. No.
17/932,152
Granted
Aug 5, 2025
Kind
B2
Abstract

Methods and devices are provided where a wearable device may receive sensor data and activate a recipe building mode of the wearable device when the sensor data satisfies a commencement condition. An image sensor of the wearable device may capture images of a physical environment. A recognition engine of the wearable device may identify ingredients detected in the images, determine an amount of the ingredients, identify utensils detected in the images, track actions of a user based on the images, and determine the name of a recipe, in response to terminating the capture of the images. The wearable device may store the recipe, the recipe including the name of the recipe, the ingredients, the amount of ingredients, the utensils, the actions of the user, and the one or more images. The recipe may be annotated with captions and output on a display of the wearable device.

Claims (58)

1. A method, comprising:

capturing, by an image sensor of a wearable device, one or more images of a physical environment around the wearable device;

identifying, based on at least one recognition engine of the wearable device, ingredients detected in the one or more images;

determining an amount of each of the ingredients detected in the one or more images;

identifying, based on the at least one recognition engine, utensils detected in the one or more images;

identifying, based on the at least one recognition engine, actions of a user based on the one or more images;

determining a time that each action of the actions is performed based on the one or more images;

building a recipe based on the ingredients, the amount of each of the ingredients, the utensils, the actions of the user, and the time that each action is performed as captured in the one or more images and processed by the at least one recognition engine;

determining a name for the recipe by the at least one recognition engine;

annotating the recipe with captions on at least a portion of the one or more images including the name of the recipe, the ingredients, the amount of each of the ingredients, the utensils, the actions of the user and the time that each action is performed; and

storing the recipe in a memory.

2. The method of claim 1 , further comprising outputting the recipe on a display of the wearable device.

3. The method of claim 1 , further comprising detecting a time spent on each action and a time between each action.

4. The method of claim 1 , further comprising a smart phone in communication with the wearable device and the at least one recognition engine being disposed on the smart phone.

5. The method of claim 1 , wherein the wearable device comprises smart glasses and identifying the actions comprises identifying the actions based on the one or more images and data output from one or more inertial sensors installed in another wearable device.

6. The method of claim 5 , wherein the one or more inertial sensors comprises any one or any combination of an accelerometer, a gyroscope, and a magnetometer.

7. The method of claim 5 , further comprising projecting, on a display of the smart glasses, the recipe onto a field of view of the user.

8. The method of claim 1 , wherein the at least one recognition engine comprises a neural network trained to recognize any one or any combination of the ingredients, the utensils, and the actions of the user.

9. The method of claim 1 , further comprising:

receiving sensor data; and

activating a recipe building mode of the wearable device, in response to the sensor data satisfying a commencement condition,

wherein the commencement condition comprises at least one of a location indicating the user is proximate to a kitchen or a location indicating the user is proximate to a place of creation of a previous recipe, and the at least one recognition engine identifying one or more objects associated with the kitchen.

10. The method of claim 9 , wherein the commencement condition comprises data from a hunger sensor indicating the user being hungry.

11. The method of claim 9 , wherein activating of the recipe building mode comprises activating the recipe building mode, in response to the at least one recognition engine recognizing an input from the user.

12. The method of claim 1 , further comprising generating a textual recipe as an output based on the recipe with the captions.

13. A wearable device, comprising:

at least one processor; and

a memory storing instructions that, when executed by the at least one processor, configures the at least one processor to:

capture, by an image sensor of the wearable device, one or more images of a physical environment around the wearable device;

identify, based on at least one recognition engine of the wearable device, ingredients detected in the one or more images;

determine an amount of each of the ingredients detected in the one or more images;

identify, based on the at least one recognition engine, utensils detected in the one or more images;

identify, based on the at least one recognition engine, actions of a user based on the one or more images;

determine a time that each action of the actions is performed based on the one or more images;

build a recipe based on the ingredients, the amount of each of the ingredients, the utensils, the actions of the user, and the time that each action is performed as captured in the one or more images and processed by the at least one recognition engine;

determine a name for the recipe;

annotate the recipe with captions on at least a portion of the one or more images including the name of the recipe, the ingredients, the amount of each of the ingredients, the utensils, the actions of the user and the time that each action is performed; and

store the recipe in the memory.

14. The wearable device of claim 13 , wherein the at least one processor is further configured to output the recipe on a display of the wearable device.

15. The wearable device of claim 13 , further comprising a smart phone in communication with the wearable device and the at least one recognition engine being installed on the smart phone.

16. The wearable device of claim 13 , wherein the at least one recognition engine comprises a neural network trained to recognize at least one of the ingredients, the utensils, or the actions of the user.

17. The wearable device of claim 13 , wherein the at least one processor is further configured to:

receive sensor data; and

activate a recipe building mode of the wearable device, in response to the sensor data satisfying a commencement condition,

wherein the commencement condition comprises at least one of a location indicating the user is proximate to a kitchen or a location indicating the user is proximate to a place of creation of a previous recipe, and the at least one recognition engine identifying one or more objects associated with the kitchen.

18. The wearable device of claim 13 , wherein the wearable device includes smart glasses.

19. A computer program product tangibly embodied on a non-transitory computer-readable medium and including executable code that, when executed, causes a wearable device to:

capture, by an image sensor of the wearable device, one or more images of a physical environment around the wearable device;

identify, based on at least one recognition engine of the wearable device, ingredients detected in the one or more images;

determine an amount of each of the ingredients detected in the one or more images;

identify, based on the at least one recognition engine, utensils detected in the one or more images;

identify, based on the at least one recognition engine, actions of a user based on the one or more images;

determine a time that each action of the actions is performed based on the one or more images;

build a recipe based on the ingredients, the amount of each of the ingredients, the utensils, the actions of the user, and the time that each action is performed as captured in the one or more images and processed by the at least one recognition engine;

determine a name for the recipe;

annotate the recipe with captions on at least a portion of the one or more images including the name of the recipe, the ingredients, the amount of each of the ingredients, the utensils, the actions of the user and the time that each action is performed; and

store the recipe in the memory.

20. The computer program product of claim 19 , wherein the wearable device includes smart glasses.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 21, 2022
From: DIAZ, OMAR ESTRADA
To: GOOGLE LLC
Reel/Frame 061158/0798 →
Continuity (1)
Related Publication 20240087345A1 · Mar 14, 2024
References Cited (35)
US 11238664B1 · Tavakoli · 2022 [cited by examiner]
US 20090258331A1 · Do et al. · 2009 [cited by applicant]
US 20140147829A1 · Jerauld · 2014 [cited by applicant]
US 20170103676A1 · Allen · 2017 [cited by examiner]
US 20180000274A1 · Sun · 2018 [cited by examiner]
US 20180257219A1 · Oleynik · 2018 [cited by applicant]
US 20200211062A1 · Kossakovski · 2020 [cited by examiner]
US 20210118447A1 · Kim · 2021 [cited by examiner]
US 20220273139A1 · Mahapatra · 2022 [cited by examiner]
US 20220351509A1 · Kanemura · 2022 [cited by examiner]
US 20230252806A1 · DeSantola · 2023 [cited by examiner]
DE 102016109894A1 · 2017 [cited by examiner]
JP 2021140711A · 2021 [cited by examiner]
Malmaud, Jonathan, et al. “What's cookin'? interpreting cooking videos using text, speech and vision.” arXiv preprint arXiv: 1503.01558 (2015). (Year: 2015). [cited by examiner]
Bianco, Simone, et al. “Cooking Action Recognition with i VAT: An Interactive Video Annotation Tool.” Image Analysis and Processing—ICIAP 2013: 17th International Conference, Naples, Italy, Sep. 9-13, 2013, Proceedings,… [cited by examiner]
Salvador, Amaia, et al. “Learning cross-modal embeddings for cooking recipes and food images.” Proceedings of the IEEE conference on computer vision and pattern recognition. 2017. (Year: 2017). [cited by examiner]
Chen, Jing-Jing, et al. “Deep understanding of cooking procedure for cross-modal recipe retrieval.” Proceedings of the 26th ACM international conference on Multimedia. 2018. (Year: 2018). [cited by examiner]
Malmaud, Jonathan, et al. “What's cookin'? interpreting cooking videos using text, speech and vision.” arXiv preprint arXiv:1503.01558 (2015). (Year: 2015) (Year: 2015). [cited by examiner]
Bianco, Simone, et al. “Cooking Action Recognition with i VAT: An Interactive Video Annotation Tool.” Image Analysis and Processing—ICIAP 2013: 17th International Conference, Naples, Italy, Sep. 9-13, 2013, Proceedings,… [cited by examiner]
Salvador, Amaia, et al. “Learning cross-modal embeddings for cooking recipes and food images.” Proceedings of the IEEE conference on computer vision and pattern recognition. 2017. (Year: 2017) (Year: 2017). [cited by examiner]
Chen, Jing-Jing, et al. “Deep understanding of cooking procedure for cross-modal recipe retrieval.” Proceedings of the 26th ACM international conference on Multimedia. 2018. (Year: 2018) (Year: 2018). [cited by examiner]
“4.1 / Recipe Builder”, retrieved from: https://help.healthiapp.com/support/solutions/articles/13000060423-4-1-recipe-builder, Nov. 17, 2021, 10 pages. [cited by applicant]
“AR Recipe App Development”, Switchstance, retrieved on Sep. 14, 2022 from https://switchstance.agency/portfolio/ar-augmented reality-recipe-app-development/, 2022, 5 pages. [cited by applicant]
“Create Custom Recipes”, retrieved from: https://help.carbmanager.com/docs/create-custom-recipes, Oct. 27, 2021, 3 pages. [cited by applicant]
“Future Friday: Food Information Displayed on AR Smart Glasses”, Vuzix Corporation, retrieved from https://www.vuzix.com/blogs/vuzix-blog/future-friday-food-information-displayed-on-ar-smart-glasses, 2019, 3 pages. [cited by applicant]
“Future Friday: How Smart Glasses Could Change Kitchen Life”, Vuzix Corporation, retrieved from https://www.vuzix.com/blogs/vuzix-blog/future-friday-how-smart-glasses-could change-kitchen-life, 2020, 3 pages. [cited by applicant]
“How To Track A Homemade Recipe”, Cali Bett, retrieved on Sep. 14, 2022 from: https://www.calibett.com/blog/how-to-track-at-home, 24 pages. [cited by applicant]
“Speech-to-Text”, Google Cloud, retrieved on Sep. 14, 2022 from: https://cloud.google.com/speech-to-text, 14 pages. [cited by applicant]
Albright, “How Google Glass Will Change the Way You Cook”, National Geographic Society, retrieved from: https://www.nationalgeographic.com/culture/article/how-google-glass-will-change-the-way-you-cook, Aug. 4, 2014, 8 p… [cited by applicant]
Cheung, “YOLO for Real-Time Food Detection”, retrieved from: http://bennycheung.github.io/yolo-for-real-time-food-detection, Jun. 7, 2018, 22 pages. [cited by applicant]
Diete, et al., “Recognizing Grabbing Actions from Inertial and Video Sensor Data in a Warehouse Scenario”, ScienceDirect, Procedia Computer Science 110, 2017, pp. 16-23. [cited by applicant]
Farr, “Microsoft Has An Idea For Using Smart Glasses To Track Your Diet”, retrieved from: https://www.cnbc.com/2017/05/09/microsoft-patents-smart-glasses-for-diet-tracking-.html, May 9, 2017, 5 pages. [cited by applicant]
Gogate, et al., “Hunger and stress monitoring system using galvanic skin response”, Indonesian Journal of Electrical Engineering and Computer Science, vol. 13, No. 3, Mar. 2019, pp. 861-865. [cited by applicant]
Olynick, “AR Cooking (In Progress)”, Augmented Reality, retrieved from https://www.dianaolynick.com/blog/ar-cooking-in-progress, Feb. 7, 2022, 3 pages. [cited by applicant]
Scholl, “Extract And Tag Ingredients From A Website”, retrieved from: https://schollz.com/blog/ingredients/, Jul. 16, 2018, 8 pages. [cited by applicant]