IP Library Granted Patent US 10,325,372
Granted Patent B2
US 10,325,372 · App. 15/385,249 · Granted Jun 18, 2019

Intelligent auto-cropping of images

Inventors: Amit Kumar Agrawal (Santa Clara, CA); Alexander Adrian Hugh Davidson (Seattle, WA); Prakash Ramu (San Mateo, CA)
Assignee: Amazon Technologies, Inc.
G06T7/11G06K9/4609G06K9/6202G06T5/50G06T7/194G06T7/90H04N13/106H04N13/15H04N13/207H04N13/356G06T2207/10012G06T2207/10024G06T2207/10028G06T2207/20132G06T2207/20221G06T2207/20224G06T2207/30196
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,325,372
App. No.
15/385,249
Granted
Jun 18, 2019
Kind
B2
Abstract

Techniques for providing an accurate auto-crop feature for images captured by an image capture device may be described herein. For example, one or more image masks for a color image captured by an image capture device may be received by a computer system. Metadata about the color image that identifies portions of the image as foreground and the color image itself may also be received by the computer system. Further, a representation of a user and a floor region associated with a user may be extracted from the color image using the one or more image masks and the metadata. A first area of the color image may be cropped with respect to the extracted representation of the user and the floor region associated with the user to generate a second area of the color image. In embodiments, a third area of the color image may be obscured based on the received metadata.

Claims (29)

1. A computer-implemented method, comprising:

receiving, by a computer system and from an image capture device, a first image mask that comprises a two-dimensional (2D) representation of a user in an image captured by the image capture device and first metadata that identifies a first subset of regions in the image as being in a foreground of the image;

receiving, by the computer system and from the image capture device, a second image mask that comprises a representation of a floor region associated with the user in the image captured by the image capture device and second metadata that identifies a second subset of regions in the image as the foreground of the image;

receiving, by the computer system and from the image capture device, a color image of the user;

extracting, by the computer system, the representation of the user and the floor region associated with the user, from the color image of the user, based at least in part on the first image mask, the second image mask, the first metadata, and the second metadata;

cropping, by the computer system, a first area of the color image of the user with respect to the extracted representation of the user and the floor region associated with the user based at least in part on the first image mask and the second image mask thereby generating a second area of the color image; and

obscuring, by the computer system, a third area of the cropped color image based at least in part on the first metadata and the second metadata thereby generating a revised color image of the user that comprises a combination of the extracted representation of the user and the floor region associated with the user.

2. The computer-implemented method of claim 1 , further comprising maintaining a plurality of cropped color images of the user.

3. The computer-implemented method of claim 1 , further comprising identifying a plurality of items included in the image of the user based at least in part on an item recognition algorithm and an item catalog.

4. The computer-implemented method of claim 3 , further comprising generating one or more item listing web pages for offering the identified plurality of items included in the image.

5. The computer-implemented method of claim 1 , wherein receiving the color image of the user includes receiving third metadata that identifies a third subset of regions in the image as a background of the image.

6. The computer-implemented method of claim 5 , wherein the first metadata, the second metadata, and the third metadata further identify a respective depth measurement for each pixel within the image captured by the image capture device.

7. The computer-implemented method of claim 1 , wherein the image capture device comprises a depth sensor for capturing a three-dimensional (3D) image of the user and is further configured to convert the 3D image of the user to the 2D representation of the user using the color image of the user.

8. A computer system, comprising:

a memory that stores computer-executable instructions;

a first sensor configured to capture a three-dimensional (3D) image of an object;

a second sensor configured to capture a color image of the object; and

at least one processor configured to access the memory and execute the computer-executable instructions to collectively:

obtain a first image mask that comprises a two-dimensional (2D) representation of a user in an image captured by the first sensor and first metadata that identifies a first subset of regions in the image as being in a foreground of the image based at least in part on a 3D image of the image captured by the first sensor;

obtain a second image mask that comprises a representation of a floor region associated with the user in the image captured by the first sensor and second metadata that identifies a second subset of regions in the image as the foreground of the image;

obtain the color image of the user from the second sensor;

extract the representation of the user and the floor region associated with the user, from the color image of the user, based at least in part on the first image mask, the second image mask, the first metadata, and the second metadata; and

remove a first area of the color image of the user with respect to the extracted representation of the user and the floor region associated with the user based at least in part on the first image mask and the second image mask thereby generating a second area of the color image.

9. The computer system of claim 8 , wherein the at least one processor is further configured to display a revised image of the user that comprises a combination of the extracted representation of the user and the floor region associated with the user within the second area of the color image.

10. The computer system of claim 8 , wherein the at least one processor is further configured to obscure a third area of the color image based at least in part on the first metadata and the second metadata.

11. The computer system of claim 8 , wherein obtaining the first image mask includes converting a 3D image of the user to a 2D image of the user.

12. The computer system of claim 8 , wherein the at least one processor is further configured to identify one or more objects in the image based at least in part on an item recognition algorithm.

13. The computer system of claim 12 , wherein the at least one processor is further configured to transmit instructions to the user for removing the identified one or more objects from the first subset of regions in the image.

14. The computer system of claim 12 , wherein the at least one processor is further configured to transmit instructions to the user for capturing another image of the user in response to an indication that the identified one or more objects have been removed from the first subset of regions in the image.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 20, 2016
From: AGRAWAL, AMIT KUMAR; DAVIDSON, ALEXANDER ADRIAN HUGH; RAMU, PRAKASH
To: AMAZON TECHNOLOGIES, INC.
Reel/Frame 040694/0374 →
Continuity (1)
Related Publication 20180174299A1 · Jun 21, 2018