IP Library Granted Patent US 10,410,089
Granted Patent B2
US 10,410,089 · App. 15/875,579 · Granted Sep 10, 2019

Training assistance using synthetic images

Inventors: Hiu Lok Szeto (Richmond Hill, CA); Syed Alimul Huda (Scarborough, CA)
Assignee: SEIKO EPSON CORPORATION
G06K9/6256G02B27/017G06K9/00671G06T19/006
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,410,089
App. No.
15/875,579
Granted
Sep 10, 2019
Kind
B2
Abstract

A method of detecting an object in a real scene using a computer includes specifying a 3D model corresponding to the object. The method further includes acquiring, from a capture camera, an image frame of a reference object captured from a first view angle. The method further includes generating a 2D synthetic image by rendering the 3D model in a second view angle that is different from the first view angle. The method further includes generating training data using (i) the image frame, (ii) the 2D synthetic image, (iii) the run-time camera parameter, and (iv) the capture camera parameter. The method further includes storing the generated training data in one or more memories.

Claims (60)

1. A non-transitory computer readable medium that embodies instructions that cause one or more processors to perform a method comprising:

(A) specifying a 3D model, stored in one or more memories, corresponding to an object;

(B) setting a run-time camera parameter for a run-time camera that is to be used in detecting a pose of the object in a real scene by:

receiving, through a data communication port, information representing an Augmented-Reality (AR) device having the run-time camera; and

acquiring, from one or more memories, the runtime camera parameter corresponding to the AR device in response to the information, a plurality of the run-time camera parameters and corresponding AR devices being associated with each other in one or more memories;

(C) setting a capture camera parameter for a capture camera that is to be used in capturing an image frame containing a reference object;

(D) acquiring, from the capture camera, the image frame of the reference object captured from a first view angle;

(E) generating a 2D synthetic image by rendering the 3D model in a second view angle that is different from the first view angle;

(F) generating training data using (i) the image frame, (ii) the 2D synthetic image, (iii) the run-time camera parameter, and (iv) the capture camera parameter; and

(G) storing the generated training data in one or more memories.

2. The non-transitory computer readable medium according to claim 1 , wherein the AR device is a head-mounted display device having the run-time camera.

3. The non-transitory computer readable medium according to claim 1 , wherein one camera functions as the run-time camera and the capture camera.

4. The non-transitory computer readable medium according to claim 1 , wherein the capture camera is a component of a smartphone or tablet.

5. The non-transitory computer readable medium according to claim 4 , wherein the one or more processors are a component of the smartphone or tablet.

6. The non-transitory computer readable medium according to claim 1 , wherein the method further comprises:

(H) selecting data representing a first view range that includes the first view angle, prior to step (D).

7. The non-transitory computer readable medium according to claim 1 , where the method further comprises:

(H) providing the generated training data to another processor connected to the run-time camera.

8. A non-transitory computer readable medium that embodies instructions that cause one or more processors to perform a method comprising:

(A) specifying a 3D model, stored in one or more memories, corresponding to an object;

(B) setting a run-time camera parameter for a run-time camera that is to be used in detecting a pose of the object in a real scene by acquiring, through a data communication port, the run-time camera parameter from an Augmented-Reality (AR) device having the run-time camera;

(C) setting a capture camera parameter for a capture camera that is to be used in capturing an image frame containing a reference object;

(D) acquiring, from the capture camera, the image frame of the reference object captured from a first view angle;

(E) generating a 2D synthetic image by rendering the 3D model in a second view angle that is different from the first view angle;

(F) generating training data using (i) the image frame, (ii) the 2D synthetic image, (iii) the run-time camera parameter, and (iv) the capture camera parameter; and

(G) storing the generated training data in one or more memories.

9. The non-transitory computer readable medium according to claim 8 , wherein the AR device is a head-mounted display device having the run-time camera.

10. The non-transitory computer readable medium according to claim 8 , wherein one camera functions as the run-time camera and the capture camera.

11. The non-transitory computer readable medium according to claim 8 , wherein the capture camera is a component of a smartphone or tablet.

12. The non-transitory computer readable medium according to claim 11 , wherein the one or more processors are a component of the smartphone or tablet.

13. The non-transitory computer readable medium according to claim 8 , wherein the method further comprises:

(H) selecting data representing a first view range that includes the first view angle, prior to step (D).

14. The non-transitory computer readable medium according to claim 8 , where the method further comprises:

(H) providing the generated training data to another processor connected to the run-time camera.

15. A method of detecting an object in a real scene using a computer, the method comprising:

(A) specifying a 3D model, stored in one or more memories, corresponding to the object, using a processor;

(B) setting a run-time camera parameter for a run-time camera that is to be used in detecting a pose of the object in the real scene, using the processor, by:

receiving, through a data communication port, information representing an Augmented-Reality (AR) device having the run-time camera; and

acquiring, from one or more memories, the runtime camera parameter corresponding to the AR device in response to the information, a plurality of the run-time camera parameters and corresponding AR devices being associated with each other in one or more memories;

(C) setting a capture camera parameter for a capture camera that is to be used in capturing an image frame containing a reference object, using the processor;

(D) acquiring, from the capture camera, the image frame of the reference object captured from a first view angle, using the processor;

(E) generating a 2D synthetic image by rendering the 3D model in a second view angle that is different from the first view angle, using the processor;

(F) generating training data using (i) the image frame, (ii) the 2D synthetic image, (iii) the run-time camera parameter, and (iv) the capture camera parameter, using the processor; and

(G) storing the generated training data in one or more memories, using the processor.

16. The method according to claim 15 , wherein the AR device is a head-mounted display device having the run-time camera.

17. The method according to claim 15 , wherein one camera functions as the run-time camera and the capture camera.

18. The method according to claim 15 , wherein the capture camera is a component of a smartphone or tablet.

19. The method according to claim 18 , wherein the processor is a component of the smartphone or tablet.

20. The method according to claim 15 , further comprising:

(H) selecting data representing a first view range that includes the first view angle, prior to step (D).

21. The method according to claim 15 , further comprising:

(H) providing the generated training data to another processor connected to the run-time camera.

22. A method of detecting an object in a real scene using a computer, the method comprising:

(A) specifying a 3D model, stored in one or more memories, corresponding to the object, using a processor;

(B) setting a run-time camera parameter for a run-time camera that is to be used in detecting a pose of the object in the real scene, using the processor by acquiring, through a data communication port, the run-time camera parameter from an Augmented-Reality (AR) device having the run-time camera;

(C) setting a capture camera parameter for a capture camera that is to be used in capturing an image frame containing a reference object, using the processor;

(D) acquiring, from the capture camera, the image frame of the reference object captured from a first view angle, using the processor;

(E) generating a 2D synthetic image by rendering the 3D model in a second view angle that is different from the first view angle, using the processor;

(F) generating training data using (i) the image frame, (ii) the 2D synthetic image, (iii) the run-time camera parameter, and (iv) the capture camera parameter, using the processor; and

(G) storing the generated training data in one or more memories, using the processor.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 19, 2018
From: SZETO, HIU LOK; HUDA, SYED ALIMUL
To: EPSON CANADA LIMITED
Reel/Frame 044671/0813 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 19, 2018
From: EPSON CANADA, LTD.
To: SEIKO EPSON CORPORATION
Reel/Frame 045095/0762 →
Continuity (1)
Related Publication 20190228263A1 · Jul 25, 2019