IP Library › Granted Patent US 12,505,665
Granted Patent B2
US 12,505,665 · App. 17/594,209 · Granted Dec 23, 2025

Distributed sensor data processing using multiple classifiers on multiple devices

Inventors: Alex Olwal (Santa Cruz, CA); Kevin Balke (Sunnyvale, CA); Dmitrii Votintcev (Menlo Park, CA)
Assignee: GOOGLE LLC
G06V10/95G06V10/764G06V40/168H04N7/188
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,505,665
App. No.
17/594,209
Granted
Dec 23, 2025
Kind
B2
Abstract

According to an aspect, a method for distributed sound/image recognition using a wearable device includes receiving, via at least one sensor device, sensor data, and detecting, by a classifier of the wearable device, whether or not the sensor data includes an object of interest. The classifier configured to execute a first machine learning (ML) model. The method includes transmitting, via a wireless connection, the sensor data to a computing device in response to the object of interest being detected within the sensor data, where the sensor data is configured to be used by a second ML model on the computing device or a server computer for further sound/image classification.

Claims (66)

1 . A method comprising:

determining that a condition for motion of a wearable device achieves a threshold for stability based on sensor data from a motion sensor of the wearable device;

detecting, by a first model on the wearable device, whether or not an object of interest is included within a first image frame;

in response to the object of interest being detected in the first image frame and the condition achieving the threshold, transmitting the first image frame to a computing device;

receiving location data generated by a second model on the computing device, the location data identifying a location of the object of interest in the first image frame;

identifying an image region in a second image frame using the location data of the first image frame, the second image frame being subsequent to the first image frame; and

transmitting the image region to the computing device.

2 . The method of claim 1 , further comprising:

extracting the image region from the second image frame; and

generating a compressed image region from the image region, wherein the compressed image region is transmitted to the computing device.

3 . The method of claim 1 , wherein the object of interest includes facial features, wherein the method further comprises:

performing, by the second model, facial recognition using the image region of the second image frame.

4 . The method of claim 1 , further comprising:

receiving first image data from a first camera;

detecting whether the object of interest is included in the first image data; and

in response to the object of interest being detected in the first image data, activating a second camera to obtain second image data, the second image data having a resolution higher than the first image data, wherein the first image frame is selected from the second image data.

5 . The method of claim 4 , further comprising:

receiving, via a light condition sensor of the wearable device, light condition information, the light condition information indicating a level of light; and

activating the first camera based on the light condition information.

6 . The method of claim 1 , wherein the wearable device includes an extended reality device configured to be worn on a head of a user, and the computing device includes a mobile user device wirelessly connected to the extended reality device.

7 . The method of claim 1 , further comprising:

receiving classification data associated with the object of interest, the classification data being determined by the second model; and

rendering the classification data on the wearable device.

8 . The method of claim 7 , wherein the classification data is rendered at a location in a field of view of a user that corresponds to the object of interest.

9 . A non-transitory computer-readable medium storing executable instructions that when executed by at least one processor cause the at least one processor to execute operations, the operations comprising:

determining that a condition for motion of a wearable device achieves a threshold for stability based on sensor data from a motion sensor of the wearable device;

detecting, by a first model, whether or not an object of interest is included within a first image frame;

in response to the object of interest being detected in the first image frame and the condition achieving the threshold, transmitting the first image frame to a computing device;

receiving location data generated by a second model on the computing device, the location data identifying a location of the object of interest in the first image frame;

identifying an image region in a second image frame using the location data of the first image frame, the second image frame being subsequent to the first image frame; and

transmitting the image region to the computing device.

10 . The non-transitory computer-readable medium of claim 9 ,

wherein the operations further comprise:

selecting the image region from the second image frame; and

generating a compressed image region from the image region, wherein the compressed image region is transmitted to the computing device.

11 . The non-transitory computer-readable medium of claim 9 , wherein the object of interest includes a machine-readable representation of data, wherein the operations further comprise:

decoding, by the second model, the machine-readable representation of data from the image region.

12 . The non-transitory computer-readable medium of claim 9 , wherein the operations further comprise:

receiving first image data from a first camera;

detecting whether the object of interest is included in the first image data; and

in response to the object of interest being detected in the first image data, activating a second camera to obtain second image data, the second image data having a resolution higher than the first image data, wherein the first image frame is selected from the second image data.

13 . The non-transitory computer-readable medium of claim 12 , wherein the operations further comprise:

performing, by the second model, facial recognition using the image region of the second image frame.

14 . The non-transitory computer-readable medium of claim 12 , wherein the operations further comprise:

receiving light condition information from a light condition sensor of the wearable device, the light condition information indicating a level of ambient light; and

activating the first camera in response to the level of ambient light achieving a threshold level.

15 . A wearable device for distributed image recognition, the wearable device comprising:

at least one processor; and

a non-transitory computer-readable medium storing executable instructions that cause the at least one processor to:

determine that a condition for motion of a wearable device achieves a threshold for stability based on sensor data from a motion sensor of the wearable device;

detect whether or not an object of interest is included within a first image frame;

in response to the object of interest being detected in the first image frame and the condition achieving the threshold, transmit the first image frame to a computing device;

receive location data generated by a second model on the computing device, the location data identifying a location of the object of interest in the first image frame;

identify an image region in a second image frame using the location data of the first image frame, the second image frame being subsequent to the first image frame; and

transmit the image region to the computing device.

16 . The wearable device of claim 15 , wherein the executable instructions include instructions that cause the at least one processor to:

activate a first camera in response to a level of ambient light achieving a threshold condition;

receive first image data from the first camera;

detect whether the object of interest is included in the first image data; and

in response to the object of interest being detected in the first image data, activate a second camera to obtain second image data, the second image data having a resolution higher than the first image data, wherein the first image frame is selected from the second image data.

17 . The wearable device of claim 15 , wherein the executable instructions include instructions that cause the at least one processor to:

generate a compressed image region from the image region, wherein the compressed image region is transmitted to the computing device.

18 . The wearable device of claim 15 , wherein the executable instructions include instructions that cause the at least one processor to:

receive classification data associated with the object of interest, the classification data being determined by the second model; and

render the classification data at a location on a display of the wearable device that corresponds to a location of the object of interest.

19 . The wearable device of claim 15 , wherein the wearable device includes an extended reality device configured to be worn on a head of a user, the computing device includes a mobile user device wirelessly connected to the extended reality device.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 6, 2021
From: OLWAL, ALEX; BALKE, KEVIN; VOTINTCEV, DMITRII
To: GOOGLE LLC
Reel/Frame 057720/0039 →
Continuity (1)
Related Publication 20220165054A1 · May 26, 2022
References Cited (76)
US 8655404B1 · Singh · 2014 [cited by examiner]
US 9324322B1 · Torok et al. · 2016 [cited by applicant]
US 10515623B1 · Grizzel · 2019 [cited by applicant]
US 10896597B1 · Kocher · 2021 [cited by examiner]
US 11144749B1 · Lo · 2021 [cited by examiner]
US 11402871B1 · Berliner · 2022 [cited by examiner]
US 11647167B2 · Lee · 2023 [cited by examiner]
US 20070069976A1 · Willins · 2007 [cited by examiner]
US 20080180537A1 · Weinberg · 2008 [cited by examiner]
US 20090204409A1 · Mozer et al. · 2009 [cited by applicant]
US 20120221330A1 · Thambiratnam et al. · 2012 [cited by applicant]
US 20120275690A1 · Melvin et al. · 2012 [cited by applicant]
US 20130169536A1 · Wexler et al. · 2013 [cited by applicant]
US 20140214596A1 · Acker, Jr. · 2014 [cited by examiner]
US 20160026853A1 · Wexler et al. · 2016 [cited by applicant]
US 20160026870A1 · Wexler · 2016 [cited by examiner]
US 20160080295A1 · Davies · 2016 [cited by examiner]
US 20160080652A1 · Shirota et al. · 2016 [cited by applicant]
US 20160094814A1 · Gousev et al. · 2016 [cited by applicant]
US 20160125877A1 · Foerster et al. · 2016 [cited by applicant]
US 20160203691A1 · Arnold et al. · 2016 [cited by applicant]
US 20170213145A1 · Pathak · 2017 [cited by examiner]
US 20170249863A1 · Murgia et al. · 2017 [cited by applicant]
US 20170293480A1 · Wexler · 2017 [cited by examiner]
US 20180018451A1 · Spizhevoy · 2018 [cited by examiner]
US 20180307880A1 · Gao · 2018 [cited by examiner]
US 20190035125A1 · Bellows · 2019 [cited by examiner]
US 20190087690A1 · Srivastava · 2019 [cited by applicant]
US 20190093985A1 · Robinson · 2019 [cited by examiner]
US 20190222756A1 · Moloney et al. · 2019 [cited by applicant]
US 20190331914A1 · Lee et al. · 2019 [cited by applicant]
US 20190355379A1 · Ravindran et al. · 2019 [cited by applicant]
US 20190373998A1 · Knittel · 2019 [cited by examiner]
US 20190379822A1 · Leong · 2019 [cited by examiner]
US 20200004291A1 · Wexler · 2020 [cited by examiner]
US 20200034615A1 · Croxford et al. · 2020 [cited by applicant]
US 20200069281A1 · Chan et al. · 2020 [cited by applicant]
US 20200244885A1 · Li · 2020 [cited by examiner]
US 20200357516A1 · Kirby · 2020 [cited by examiner]
US 20200380952A1 · Zhang et al. · 2020 [cited by applicant]
US 20210081749A1 · Claire · 2021 [cited by examiner]
US 20210124045A1 · Xiao · 2021 [cited by examiner]
US 20210350566A1 · Hu · 2021 [cited by examiner]
US 20210365673A1 · Stockton · 2021 [cited by examiner]
US 20220060666A1 · Lee · 2022 [cited by examiner]
US 20220240824A1 · Maier · 2022 [cited by examiner]
US 20220254176A1 · Utsugida · 2022 [cited by examiner]
US 20220294971A1 · Dahlgren · 2022 [cited by examiner]
US 20220294992A1 · Manzari · 2022 [cited by examiner]
US 20220318541A1 · Dai · 2022 [cited by examiner]
CN 105182535A · 2015 [cited by applicant]
CN 107978316A · 2018 [cited by applicant]
CN 108962240A · 2018 [cited by applicant]
CN 109475294A · 2019 [cited by applicant]
CN 109588060A · 2019 [cited by applicant]
CN 110570864A · 2019 [cited by applicant]
CN 111048062A · 2020 [cited by applicant]
CN 111242354A · 2020 [cited by applicant]
CN 111727438A · 2020 [cited by applicant]
EP 3182349A2 · 2017 [cited by applicant]
EP 3327616A1 · 2018 [cited by applicant]
GB 2447246A · 2008 [cited by applicant]
TW I695312B · 2020 [cited by applicant]
TW M596382U · 2020 [cited by applicant]
WO 2019194863A1 · 2019 [cited by applicant]
Sawhney, et al., “Speaking and Listening on the Run: Design for Wearable Audio Computing”, Digest of Papers, Second International Symposium on Wearable Computers (Cat. No. 98EX215), 1998, 8 pages. [cited by applicant]
International Search Report and Written Opinion for PCT Application No. PCT/US2020/055374, mailed on Jun. 18, 2021, 16 pages. [cited by applicant]
International Search Report and Written Opinion for PCT Application No. PCT/US2020/055378, mailed on Jun. 17, 2021, 23 pages. [cited by applicant]
Jagatheesan, et al., “Hierarchical Automatic Speech Recognition Powered by Data Infrastructure”, The 11th Annual IEEE Consumer Communications and Networking Conference—Demos, 2014, pp. 1140-1141. [cited by applicant]
Lu, et al., “Watchar: 6-DOF Tracked Watch for AR Interaction”, 2020, 2 pages. [cited by applicant]
Wang, et al., “Hierarchical Convolutional Neural Network for Face Detection”, SpringerLink, ICIG 2015: Image and Graphics, retrieved on Jul. 16, 2020 from https://link.springer.com/chapter/10.1007/978-3-319-21963-9_34, … [cited by applicant]
Communication pursuant to Article 94(3) EPC for European Application No. 20803695.4, mailed Sep. 19, 2024, 6 pages. [cited by applicant]
First Office Action (with English translation) for Chinese Application No. 202080035869.3, mailed Aug. 21, 2024, 29 pages. [cited by applicant]
De Godoy, et al., “Paws: A Wearable Acoustic System for Pedestrian Safety”, 2018 IEEE/ACM Third International Conference on Internet-of-Things Design and Implementation, 2018, pp. 237-248. [cited by applicant]
Xia, et al., “Improving Pedestrian Safety in Cities Using Intelligent Wearable Systems”, IEEE Internet of Things Journal, vol. 6, No. 5, Oct. 2019, pp. 7497-7514. [cited by applicant]
First Office Action with English translation for Chinese Application No. 202080035680.4, mailed Mar. 14, 2025, 26 pages. [cited by applicant]