IP Library Granted Patent US 11,869,231
Granted Patent B2
US 11,869,231 · App. 18/150,737 · Granted Jan 9, 2024

Auto-completion for gesture-input in assistant systems

Inventors: William Crosby Presant (Seattle, WA); Francislav P Penov (Kirkland, WA); Anuj Kumar (Santa Clara, CA)
Assignee: Meta Platforms Technologies, LLC
G06V10/82G06F3/011G06F3/013G06F3/017G06F3/167G06F7/14G06F9/453G06F16/176G06F16/2255G06F16/2365G06F16/243G06F16/248G06F16/24552G06F16/24575G06F16/24578G06F16/338G06F16/3323G06F16/3329G06F16/3344G06F16/904G06F16/9038G06F16/90332G06F16/90335G06F16/951G06F16/9535G06F18/2411G06F40/205G06F40/295G06F40/30G06F40/40G06N3/006G06N3/08G06N7/01G06N20/00G06Q50/01G06V10/764G06V20/10G06V40/28G10L15/02G10L15/063G10L15/07G10L15/16G10L15/183G10L15/187G10L15/1815G10L15/1822G10L15/22G10L15/26G10L17/06G10L17/22H04L5/02H04L12/2816H04L41/20H04L41/22H04L43/0882H04L43/0894H04L51/18H04L51/216H04L51/52H04L67/306H04L67/535H04L67/5651H04L67/75H04W12/08G06F2216/13G10L13/00G10L13/04G10L2015/223G10L2015/225H04L51/046H04L67/10H04L67/53
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,869,231
App. No.
18/150,737
Granted
Jan 9, 2024
Kind
B2
Abstract

A method includes detecting a user input comprising an incomplete three-dimensional (3D) gesture performed by one or more hands of a first user by a virtual-reality (VR) headset, selecting candidate 3D gestures from pre-defined 3D gestures based on a personalized gesture-recognition model, wherein each of the candidate 3D gestures is associated with a confidence score representing a likelihood the first user intended to input the respective candidate 3D gesture, and presenting one or more suggested inputs corresponding to one or more of the candidate 3D gestures at the VR headset.

Claims (40)

1. A method comprising, by a virtual-reality (VR) headset:

detecting a user input comprising an incomplete three-dimensional (3D) gesture in the air performed by one or more hands of a first user;

selecting, based on a personalized gesture-recognition model, one or more candidate 3D gestures from a plurality of pre-defined 3D gestures, wherein each of the candidate 3D gestures is associated with a confidence score representing a likelihood the first user intended to input the respective candidate 3D gesture by performing the incomplete 3D gesture in the air; and

presenting, at the VR headset, one or more suggested inputs corresponding to one or more of the candidate 3D gestures.

2. The method of claim 1 , further comprising:

calculating, by the VR headset for each of the one or more candidate 3D gestures, a similarity level of the candidate 3D gesture with respect to the incomplete 3D gesture.

3. The method of claim 2 , wherein the similarly level of each candidate 3D gesture with respect to the incomplete 3D gesture is based on a trajectory of the incomplete 3D gesture with respect to the VR headset.

4. The method of claim 2 , wherein the similarly level of each candidate 3D gesture with respect to the incomplete 3D gesture is based on an orientation of the incomplete 3D gesture with respect to the VR headset.

5. The method of claim 2 , wherein the similarly level of each candidate 3D gesture with respect to the incomplete 3D gesture is based on an object associated with the incomplete 3D gesture.

6. The method of claim 2 , wherein the similarly level of each candidate 3D gesture with respect to the incomplete 3D gesture is based on contextual information associated with the incomplete 3D gesture.

7. The method of claim 2 , wherein the similarly level of each candidate 3D gesture with respect to the incomplete 3D gesture is based on a position of the incomplete 3D gesture with respect to the VR headset.

8. The method of claim 1 , further comprising:

calculating, by the VR headset, one or more confidence scores for one or more intents corresponding to the incomplete 3D gesture; and

determining, by the VR headset, that each of the one or more confidence scores is below a threshold score.

9. The method of claim 8 , wherein the threshold score is based on a 3D wake-up gesture performed by the first user.

10. The method of claim 8 , wherein calculating the one or more confidence scores for the one or more intents corresponding to the incomplete 3D gesture is based on a velocity associated with the incomplete 3D gesture.

11. The method of claim 8 , wherein calculating the one or more confidence scores for the one or more intents corresponding to the incomplete 3D gesture is based on temporal information associated with the incomplete 3D gesture, and wherein the temporal information comprises a pause in the user input.

12. The method of claim 8 , wherein selecting the one or more candidate 3D gestures is further based on the one or more intents.

13. The method of claim 1 , further comprising:

receiving, at the VR headset, a user-selected input from the first user, wherein the user-selected input comprises one of the suggested inputs; and

executing, by the VR headset, one or more tasks based on the user-selected input.

14. The method of claim 1 , wherein each pre-defined 3D gesture comprises one or more of pointing, poking, tapping, waving, or swiping.

15. The method of claim 1 , further comprising:

receiving, at the VR headset, a first user-selected input from the first user, wherein the first user-selected input comprises one of the suggested inputs, and wherein the first user-selected input is associated with a first intent;

generating, by the VR headset based on the first user-selected input, one or more additional candidate 3D gestures, wherein each of the one or more additional candidate 3D gestures is associated with the first intent;

presenting, at the VR headset, one or more additional suggested inputs corresponding to one or more of the additional candidate 3D gestures;

receiving, at the VR headset, a second user-selected input from the first user, wherein the second user-selected input comprises one of the additional suggested inputs; and

executing, by the VR headset, one or more tasks based on the second user-selected input.

16. One or more computer-readable non-transitory storage media embodying software that is operable when executed to:

detect, by a virtual-reality (VR) headset, a user input comprising an incomplete three-dimensional (3D) gesture in the air performed by one or more hands of a first user;

select, based on a personalized gesture-recognition model by the VR headset, one or more candidate 3D gestures from a plurality of pre-defined 3D gestures, wherein each of the candidate 3D gestures is associated with a confidence score representing a likelihood the first user intended to input the respective candidate 3D gesture by performing the incomplete 3D gesture in the air; and

present, at the VR headset, one or more suggested inputs corresponding to one or more of the candidate 3D gestures.

17. The media of claim 16 , wherein the software is further operable when executed to:

calculate, by the VR headset for each of the one or more candidate 3D gestures, a similarity level of the candidate 3D gesture with respect to the incomplete 3D gesture.

18. The media of claim 17 , wherein the similarly level of each candidate 3D gesture with respect to the incomplete 3D gesture is based on a trajectory of the incomplete 3D gesture with respect to the VR headset.

19. The media of claim 17 , wherein the similarly level of each candidate 3D gesture with respect to the incomplete 3D gesture is based on an orientation of the incomplete 3D gesture with respect to the VR headset.

20. A system comprising: one or more processors; and a non-transitory memory coupled to the processors comprising instructions executable by the processors, the processors operable when executing the instructions to:

detect, by a virtual-reality (VR) headset, a user input comprising an incomplete three-dimensional (3D) gesture in the air performed by one or more hands of a first user;

select, based on a personalized gesture-recognition model by the VR headset, one or more candidate 3D gestures from a plurality of pre-defined 3D gestures, wherein each of the candidate 3D gestures is associated with a confidence score representing a likelihood the first user intended to input the respective candidate 3D gesture by performing the incomplete 3D gesture in the air; and

present, at the VR headset, one or more suggested inputs corresponding to one or more of the candidate 3D gestures.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 30, 2024
From: PRESANT, WILLIAM CROSBY; PENOV, FRANCISLAV P.; KUMAR, ANUJ
To: FACEBOOK TECHNOLOGIES, LLC
Reel/Frame 066296/0223 →
CHANGE OF NAME Recorded Jan 30, 2024
From: FACEBOOK TECHNOLOGIES, LLC
To: META PLATFORMS TECHNOLOGIES, LLC
Reel/Frame 066380/0672 →