IP Library Granted Patent US 11,675,996
Granted Patent B2
US 11,675,996 · App. 16/570,058 · Granted Jun 13, 2023

Artificial intelligence assisted wearable

Inventor: Brian Stephen Claire (Seattle, WA)
Assignee: Microsoft Technology Licensing, LLC
G06N3/004G06V40/172G10L15/1815G10L15/22H04N23/54H04R1/028H04R1/08G01S19/13G06F1/163G10L2015/223
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,675,996
App. No.
16/570,058
Granted
Jun 13, 2023
Kind
B2
Abstract

The description relates to artificial intelligence assisted wearables, such as backpacks. An example backpack may include sensors, such as a microphone and a camera. The backpack may receive a contextual voice command from a user. The contextual voice command may include a non-explicit reference to an object in an environment. The backpack may use the sensors to sense the environment, use an artificial intelligence engine to identify the object in the environment, and use a digital assistant to perform a contextual task in response to the contextual voice command. The contextual task may relate to the object in the environment. The backpack may output a response to the contextual voice command to the user.

Claims (60)

1. A backpack, comprising:

a strap including a camera that faces a front direction of a user when the user wears the backpack;

a microphone;

a speaker;

a network interface;

a processor; and

a storage having instructions which, when executed by the processor, cause the processor to:

receive a voice command from the user via the microphone, the voice command using a non-explicit reference to an object in an environment, wherein the non-explicit reference is a contextual signal that the voice command is a contextual voice command and that the contextual voice command references the environment;

capture an image of the environment including the object via the camera;

transmit the contextual voice command and the image to an artificial intelligence engine via the network interface to cause a contextual task to be performed, the contextual task including a computerized action relating to the object;

receive a response associated with the contextual task that was performed based at least on the contextual voice command; and

output the response to the user via the speaker.

2. The backpack of claim 1 , further comprising:

a compass,

wherein the instructions further cause the processor to sense a direction that the user is facing via the compass.

3. The backpack of claim 1 , further comprising:

a global positioning system (GPS) unit,

wherein the instructions further cause the processor to determine a location of the user via the GPS unit.

4. A system, comprising:

a wearable;

a sensor attached to the wearable, the sensor being fixed relative to a body of a user and capable of sensing an environment;

a processor; and

a storage having instructions which, when executed by the processor, cause the processor to:

receive a voice command that includes a pronoun that refers to an object in the environment, wherein the pronoun is a contextual signal that the voice command is a contextual voice command and that the contextual voice command references the environment;

detect the object in the environment using the sensor;

cause an artificial intelligence engine to perform a contextual task relating to the object in response to the contextual voice command; and

output a response associated with the contextual task to the user.

5. The system of claim 4 , wherein the wearable includes a backpack.

6. The system of claim 4 , wherein the sensor includes a camera.

7. The system of claim 6 , wherein the camera is located in a strap of the wearable and facing a front direction of the user.

8. The system of claim 4 , further comprising:

a speaker for outputting the response, wherein the response includes auditory feedback.

9. The system of claim 4 , further comprising:

a light emitting diode for outputting the response, wherein the response includes visual feedback.

10. The system of claim 4 , further comprising:

a haptic actuator for outputting the response, wherein the response includes haptic feedback.

11. The system of claim 4 , further comprising:

a network interface for connecting to a network through a companion device that is capable of connecting to the network.

12. The system of claim 4 , further comprising:

a battery for charging a companion device.

13. A method, comprising:

receiving a voice command that makes a reference to an object in an environment without explicitly identifying the object, wherein the reference is a contextual signal that the voice command is a contextual voice command and that the contextual voice command references the environment;

capturing a recording of the environment including the object;

using an artificial intelligence engine to determine an identification of the object and to interpret the contextual voice command based at least on the identification of the object; and

causing a contextual task to be performed in response to the contextual voice command, the contextual task including a computerized action relating to the object in the environment.

14. The method of claim 13 , further comprising:

using a speech recognition module to interpret the contextual voice command.

15. The method of claim 13 , wherein the recording includes one or more of:

an audio recording, an image recording, and/or a video recording.

16. The method of claim 13 , further comprising:

using an image recognition module to determine the identification of the object in the recording.

17. The method of claim 13 , further comprising:

using a text recognition module to determine the identification of the object in the recording.

18. The method of claim 13 , further comprising:

using a facial recognition module to determine the identification of the object in the recording.

19. The method of claim 13 , further comprising:

using a cognitive module to determine the contextual task to be performed in response to the contextual voice command.

20. The method of claim 13 , further comprising:

generating a response associated with the contextual task; and

transmitting the response to be output to a user.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 16, 2019
From: CLAIRE, BRIAN STEPHEN
To: MICROSOFT TECHNOLOGY LICENSING, LLC
Reel/Frame 050738/0236 →
Continuity (1)
Related Publication 20210081749A1 · Mar 18, 2021
Cited By (1)
US 12,694,875