IP Library Granted Patent US 12,233,784
Granted Patent B2
US 12,233,784 · App. 18/202,120 · Granted Feb 25, 2025

Data augmentation for driver monitoring

Inventors: Sai Akhil Suggu (Cupertino, CA); Inderjot Singh Saggu (Cupertino, CA)
Assignee: PlusAI, Inc.
B60R1/20G06V10/82G06V20/597
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,233,784
App. No.
18/202,120
Filed
May 25, 2023
Granted
Feb 25, 2025
Kind
B2
Art Unit
2483
USPC
348/148
Abstract

This application is directed to augmenting training images used for generating a model for monitoring vehicle drivers. A computer system obtains a first image of a first driver in an interior of a first vehicle and separates, from the first image, a first driver image from a first background image of the interior of the first vehicle. The computer system obtains a second background image and generates a second image by overlaying the first driver image on the second background image. The second image is added to a corpus of training images to be used by a machine learning system to generate a model for monitoring vehicle drivers. In some embodiments, at least one of the first driver image and the second background image is adjusted to match lighting conditions, average intensities, and sizes of the first driver image and the second background image.

Claims (70)

1. A method for augmenting training images for model training, comprising:

at a computer system including one or more processors and memory:

obtaining a driver image and a background image selected from a corpus of background images, where the background image comprises an image of an interior of a first vehicle and the driver image comprises an image of a driver extracted from an image of an interior of a second vehicle;

generating a first image by overlaying the driver image at a first position of the background image;

generating a second image by overlaying the driver image at a second position of the background image; and

adding the first image and the second image to a corpus of training images to generate a model for autonomously monitoring vehicle drivers.

2. The method of claim 1 , further comprising:

obtaining an initial image of a first driver in an interior of a first vehicle; and

separating, from the initial image, the driver image from a driver background image of the interior of the first vehicle.

3. The method of claim 2 , separating the driver image from the driver background image, further comprising:

applying a segmentation model to generate a segmentation mask that (1) associates a plurality of first pixels of the initial image with the driver image or (2) associates a plurality of second pixels of the initial image with the driver background image.

4. The method of claim 3 , wherein the segmentation model includes a U-Net that is based on a fully convolutional network.

5. The method of claim 1 , further comprising:

training the model for autonomously monitoring vehicle drivers to determine whether a vehicle driver is looking forward at a road ahead of a vehicle.

6. The method of claim 1 , further comprising:

training the model for autonomously monitoring vehicle drivers to determine whether a vehicle driver is looking forward at a road, looking to the left, looking to the right, looking down, closing eyes, or talking.

7. The method of claim 1 , further comprising, prior to overlaying the driver image onto the background image, implementing one or more of:

adjusting one or more image properties of at least one of the driver image and the background image to match lighting conditions of the driver image and the background image; and

normalizing at least one of the driver image and the background image to match average brightness levels of the driver image and the background image.

8. The method of claim 1 , further comprising, prior to overlaying the driver image onto the background image,

scaling at least one of the driver image and the background image.

9. The method of claim 8 , wherein the driver image includes a first driver image of a first driver, the method further comprising:

obtaining a second driver image of a second driver distinct from the first driver;

generating a third image by overlaying the second driver image onto the background image; and

adding the third image to the corpus of training images to generate the model for monitoring vehicle drivers, wherein the corpus of training images includes the first image and the second image.

10. The method of claim 1 , further comprising:

collecting a plurality of background images;

mapping each of the background images onto a respective point in a multidimensional space having a distance metric d;

clustering the plurality of background images using the distance metric d to form a plurality of image clusters;

for each of the image clusters, identifying one or more background images in the image cluster that are most distant according to the distance metric d;

forming a set of candidate background images comprising the identified one or more most distant background images in each of the image clusters; and

selecting the background image from the set of candidate background images.

11. The method of claim 10 , wherein clustering the plurality of background images comprises:

selecting a positive integer number K;

selecting K cluster centers; and

for each of the plurality of background images:

determining a distance of the respective background image from each of the cluster centers; and

assigning the respective background image to a respective image cluster associated with a respective cluster center to which the respective background image has a shortest distance.

12. The method of claim 1 , further comprising:

collecting a plurality of background images;

mapping each of the background images onto a respective point in a multidimensional space having a distance metric d;

clustering the plurality of background images using the distance metric d to form a plurality of image clusters; and

for each of the plurality of background images, determining, on a clustering plot, a respective distance between the respective background image and a corresponding cluster center of an image cluster to which the respective background image belongs, wherein the background image is selected from the plurality of background images based on the respective distance between the background image and the corresponding cluster center.

13. The method of claim 1 , further comprising:

training the model for autonomously monitoring vehicle drivers to determine whether a vehicle driver is sitting on a driver seat or a passenger seat, and in accordance with a determination whether the vehicle driver is sitting on the driver eat or the passenger seat, classify the vehicle driver as a distracted driver or a distracted passenger.

14. A computer system, comprising:

one or more processors; and

memory storing one or more programs configured for execution by the one or more processors, the one or more programs comprising instructions for:

obtaining a driver image and a background image selected from a corpus of background images, where the background image comprises an image of an interior of a first vehicle and the driver image comprises an image of a driver extracted from an image of an interior of a second vehicle;

generating a first image by overlaying the driver image at a first position of the background image;

generating a second image by overlaying the driver image at a second position of the background image; and

adding the first image and the second image to a corpus of training images to generate a model for autonomously monitoring vehicle drivers.

15. The computer system of claim 14 , the one or more programs further comprising instructions for:

obtaining an initial image of a first driver in an interior of a first vehicle; and

separating, from the initial image, the driver image from a driver background image of the interior of the first vehicle.

16. The computer system of claim 15 , separating the driver image from the driver background image, further comprising:

applying a segmentation model to generate a segmentation mask that ( 1 ) associates a plurality of first pixels of the initial image with the driver image or ( 2 ) associates a plurality of second pixels of the initial image with the driver background image.

17. The computer system of claim 14 , the one or more programs further comprising instructions for:

training the model for autonomously monitoring vehicle drivers to determine whether a vehicle driver is looking forward at a road ahead of a vehicle.

18. A non-transitory computer-readable storage medium storing one or more programs configured for execution by one or more processors of a computer system, the one or more programs comprising instructions for:

obtaining a driver image and a background image selected from a corpus of background images, where the background image comprises an image of an interior of a first vehicle and the driver image comprises an image of a driver extracted from an image of an interior of a second vehicle;

generating a first image by overlaying the driver image at a first position of the background image;

generating a second image by overlaying the driver image at a second position of the background image; and

adding the first image and the second image to a corpus of training images to generate a model for autonomously monitoring vehicle drivers.

19. The non-transitory computer-readable storage medium of claim 18 , further comprising instructions for:

training the model for autonomously monitoring vehicle drivers to determine whether a vehicle driver is looking forward at a road, looking to the left, looking to the right, looking down, closing eyes, or talking.

20. The non-transitory computer-readable storage medium of claim 18 , further comprising instructions for, prior to overlaying the driver image onto the background image, implementing one or more of:

adjusting one or more image properties of at least one of the driver image and the background image to match lighting conditions of the driver image and the background image;

normalizing at least one of the driver image and the background image to match average brightness levels of the driver image and the background image; and

scaling at least one of the driver image and the background image.

Continuity (2)
Continuation 17855670 · Jun 30, 2022
Related Publication 20240001849A1 · Jan 4, 2024
References Cited (38)
US 10169680B1 · Sachdeva et al. · 2019 [cited by applicant]
US 11574462B1 · Bhatia · 2023 [cited by applicant]
US 11699282B1 · Saggu · 2023 [cited by applicant]
US 11814059B1 · Reschka et al. · 2023 [cited by applicant]
US 20160325680A1 · Curtis et al. · 2016 [cited by applicant]
US 20160379486A1 · Taylor · 2016 [cited by applicant]
US 20190108384A1 · Wang et al. · 2019 [cited by applicant]
US 20190130218A1 · Albright · 2019 [cited by examiner]
US 20200242379A1 · Mabuchi · 2020 [cited by applicant]
US 20210023992A1 · Broggi · 2021 [cited by applicant]
US 20210124916A1 · Boon · 2021 [cited by examiner]
US 20210209797A1 · Lee et al. · 2021 [cited by applicant]
US 20210229691A1 · Liu et al. · 2021 [cited by applicant]
US 20210248748A1 · Turgutlu · 2021 [cited by examiner]
US 20210309248A1 · Choe et al. · 2021 [cited by applicant]
US 20210380115A1 · Alpert et al. · 2021 [cited by applicant]
US 20210383096A1 · White et al. · 2021 [cited by applicant]
US 20210383616A1 · Rong et al. · 2021 [cited by applicant]
US 20220067408A1 · Sheu et al. · 2022 [cited by applicant]
US 20220101047A1 · Puri et al. · 2022 [cited by applicant]
US 20220137634A1 · Bozchalooi et al. · 2022 [cited by applicant]
US 20220180109A1 · Alpert · 2022 [cited by examiner]
US 20220269886A1 · Wu · 2022 [cited by examiner]
US 20220402520A1 · Hetang et al. · 2022 [cited by applicant]
US 20230132330A1 · de Oliveira Barbalho · 2023 [cited by applicant]
US 20230154127A1 · Martin-Bragado · 2023 [cited by applicant]
US 20240034372A1 · Hartmann et al. · 2024 [cited by applicant]
Bhatia, Notice of Allowance, U.S. Appl. No. 17/855,717, filed Sep. 28, 2022, 9 pgs. [cited by applicant]
Suggu, Office Action, U.S. Appl. No. 17/855,670, filed Nov. 4, 2022, 15 pgs. [cited by applicant]
Suggu, Notice of Allowance, U.S. Appl. No. 17/855,670, filed Feb. 23, 2023, 8 pgs. [cited by applicant]
Suggu, Office Action, U.S. Appl. No. 17/855,623, filed Feb. 1, 2023, 27 pgs. [cited by applicant]
Suggu, Notice of Allowance, U.S. Appl. No. 17/855,623, filed Mar. 2, 2023, 11 pgs. [cited by applicant]
Bhatia, Office Action, U.S. Appl. No. 18/083,187, filed Apr. 5, 2024, 15 pgs. [cited by applicant]
Suggu, Office Action, U.S. Appl. No. 18/202,116, filed Mar. 22, 2024, 39 pgs. [cited by applicant]
Bhatia, Final Office Action, U.S. Appl. No. 18/083,187, Aug. 29, 2024, 15 pgs. [cited by applicant]
Suggu, Notice of Allowance, U.S. Appl. No. 18/202,116, Jul 23, 2024, 10 pgs. [cited by applicant]
Bhatia, Office Action, U.S. Appl. No. 18/083,187, Dec. 11, 2024, 19 pgs. [cited by applicant]
Zhao, “Generating Image Sequences of Augmented Road Scenarious,” 2014 International Conference on Virtual Reality and Visulization, Shenyang, China, 2014. pp. 473-477 (Year: 2014), 5 pgs. [cited by applicant]