IP Library Granted Patent US 10,748,043
Granted Patent B2
US 10,748,043 · App. 15/832,656 · Granted Aug 18, 2020

Associating semantic identifiers with objects

Inventors: Vivek Pradeep (Snohomish, WA); Michelle Lynn Holtmann (Kent, WA); Steven Nabil Bathiche (Kirkland, WA)
Assignee: Microsoft Technology Licensing, LLC
G06K9/726A61B5/0205A61B5/0507A61B5/117A61B5/1113A61B5/7475G01S5/18G01S5/28G06F1/324G06F1/3206G06F1/3231G06F3/011G06F3/017G06F3/0304G06F3/0482G06F3/04842G06F3/167G06F21/32G06F21/35G06F40/211G06F40/35G06K9/00G06K9/00214G06K9/00255G06K9/00261G06K9/00288G06K9/00295G06K9/00342G06K9/00362G06K9/00711G06K9/00771G06K9/00973G06K9/6254G06K9/6255G06K9/6289G06K9/6296G06N5/025G06N5/047G06N20/00G06T7/248G06T7/292G06T7/60G06T7/70G06T7/74G07C9/28G08B13/1427G10L15/02G10L15/063G10L15/08G10L15/18G10L15/1815G10L15/1822G10L15/19G10L15/22G10L15/24G10L15/26G10L15/28G10L15/32G10L17/04G10L17/08G10L25/51H04L51/02H04L63/102H04L67/12H04L67/22H04N5/23219H04N5/332H04N7/181H04N7/188H04N21/231H04N21/42203H04N21/44218H04N21/44222H04R1/406H04R3/005H04W4/029H04W4/33A61B5/05A61B5/1118G01S11/14G01S13/38G01S13/867G01S13/888G06F3/0488G06F16/70G06F2203/0381G06F2221/2111G06K2209/09G06N3/0445G06T2207/10016G06T2207/10024G06T2207/20101G06T2207/30196G06T2207/30201G06T2207/30204G06T2207/30232G07C9/32G08B29/186G10L17/00G10L2015/0635G10L2015/088G10L2015/223G10L2015/225G10L2015/228H04N5/247Y02D10/126Y02D10/173
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,748,043
App. No.
15/832,656
Filed
Dec 5, 2017
Granted
Aug 18, 2020
Kind
B2
Examiner
LU, TOM Y
Art Unit
2667
USPC
382/103
Abstract

Computing devices and methods for associating a semantic identifier with an object are disclosed. In one example, a three-dimensional model of an environment comprising the object is generated. Image data of the environment is sent to a user computing device for display by the user computing device. User input comprising position data of the object and the semantic identifier is received. The position data is mapped to a three-dimensional location in the three-dimensional model at which the object is located. Based at least on mapping the position data to the three-dimensional location of the object, the semantic identifier is associated with the object.

Claims (33)

1. A base computing device communicatively coupled to a user computing device, the base computing device comprising:

a logic processor; and

a storage device holding instructions executable by the logic processor to:

generate a three-dimensional model of a physical environment, wherein the physical environment comprises a physical object that is not associated with a semantic identifier;

send image data of the physical environment to the user computing device for display by the user computing device, wherein the image data comprises an object image of the physical object;

receive from the user computing device user input comprising (1) position data of the physical object, wherein the position data is derived from a touch selection of the object image displayed on the user computing device, and (2) an initial semantic identifier, wherein the initial semantic identifier was not previously associated with the physical object;

map the position data to a three-dimensional location in the three-dimensional model at which the physical object is located; and

based at least on mapping the position data to the three-dimensional location of the physical object, associate the initial semantic identifier with the physical object.

2. The base computing device of claim 1 , wherein the initial semantic identifier is derived from voice input provided with the touch selection.

3. The base computing device of claim 1 , wherein receiving user input comprises receiving the image data from an image sensor of the base computing device, the image data indicating a user physically touching the physical object.

4. The base computing device of claim 1 , wherein receiving user input comprises receiving the initial semantic identifier via audio captured by a microphone of the base computing device.

5. The base computing device of claim 1 , wherein sending the image data of the physical environment to the user computing device comprises sending a video stream of the physical environment captured by an image sensor of the base computing device.

6. The base computing device of claim 1 , wherein sending the image data of the physical environment to the user computing device comprises sending rendered images of the three-dimensional model of the physical environment to the user computing device.

7. The base computing device of claim 1 , wherein sending the image data of the physical environment to the user computing device comprises sending a still image captured by an image sensor of the base computing device.

8. The base computing device of claim 1 , wherein mapping the position data to a three-dimensional location in the three-dimensional model further comprises performing object recognition to determine that the physical object is a known object.

9. At a base computing device, a method for associating an initial semantic identifier with a physical object, the method comprising:

generating a three-dimensional model of a physical environment, wherein the physical environment comprises the physical object, and wherein the physical object is not associated with a semantic identifier;

sending image data of the physical environment to a user computing device for display by the user computing device, wherein the image data comprises an object image of the physical object;

receiving from the user computing device user input comprising (1) position data of the physical object, wherein the position data is derived from a touch selection of the object image displayed on the user computing device, and (2) the initial semantic identifier, wherein the initial semantic identifier was not previously associated with the physical object;

mapping the position data to a three-dimensional location in the three-dimensional model at which the physical object is located; and

based at least on mapping the position data to the three-dimensional location of the physical object, associating the initial semantic identifier with the physical object.

10. The method of claim 9 , wherein the initial semantic identifier is derived from voice input provided with the touch selection.

11. The method of claim 9 , wherein receiving user input comprises receiving the image data from an image sensor of the base computing device, the image data indicating a user physically touching the physical object.

12. The method of claim 9 , wherein receiving user input comprises receiving the initial semantic identifier via audio captured by a microphone of the base computing device.

13. The method of claim 9 , wherein sending the image data of the physical environment to the user computing device comprises sending a video stream of the physical environment captured by an image sensor of the base computing device.

14. The method of claim 9 , wherein sending the image data of the physical environment to the user computing device comprises sending rendered images of the three-dimensional model of the physical environment to the user computing device.

15. The method of claim 9 , wherein mapping the position data to a three-dimensional location in the three-dimensional model further comprises performing object recognition to determine that the physical object is a known object.

16. At a base computing device communicatively coupled to a user computing device, a method comprising:

generating a three-dimensional model of a physical environment, wherein the physical environment comprises a physical object that is not associated with a semantic identifier;

sending image data of the physical environment to the user computing device for display by the user computing device, wherein the image data comprises an object image of the physical object;

receiving from the user computing device user input comprising (1) position data and (2) an initial semantic identifier, wherein the position data corresponds to a touch selection on the user computing device, and wherein the initial semantic identifier was not previously associated with the physical object;

mapping the position data to a three-dimensional location in the three-dimensional model at which the physical object is located; and

based at least on mapping the position data to the three-dimensional location of the physical object, associating the initial semantic identifier with the physical object.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 5, 2017
From: PRADEEP, VIVEK; HOLTMANN, MICHELLE LYNN; BATHICHE, STEVEN NABIL
To: MICROSOFT TECHNOLOGY LICENSING, LLC
Reel/Frame 044305/0235 →
Continuity (3)
Provisional Application 62459020 · Feb 14, 2017
Provisional Application 62482165 · Apr 5, 2017
Related Publication 20180232608A1 · Aug 16, 2018