IP Library Granted Patent US 12,243,176
Granted Patent B2
US 12,243,176 · App. 18/455,534 · Granted Mar 4, 2025

Visual editor for designing augmented-reality effects that utilize voice recognition

Inventors: Stef Marc Smet (London, GB); Hannes Luc Herman Verlinde (Ruislip, GB); Michael Slater (Nottingham, GB); Benjamin Patrick Blackburne (Baldock, GB); Ram Kumar Hariharan (Kirkland, WA); Chunjie Jia (Bothell, WA); Prakarn Nisarat (Seattle, WA)
Assignee: Meta Platforms, Inc.
G06T19/006G06F3/04815G10L15/22G10L2015/223
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,243,176
App. No.
18/455,534
Granted
Mar 4, 2025
Kind
B2
Abstract

In one embodiment, a computer-implemented method includes receiving, through a user interface (UI) of an artificial-reality (AR) design tool, a selection of a configurable interface element to place the AR design tool and the UI into a configure phase to configure an AR effect. The computer-implemented method further includes receiving, through the UI of the AR design tool after the AR design tool and the UI are placed into the configure phase in response to selecting the configurable interface element, instructions to add a voice-command module to the AR effect. The computer-implemented method further includes configuring, while the AR design tool and the UI are placed into the configure phase, one or more parameters of the voice-command module. The computer-implemented method further includes generating the AR effect utilizing a particular voice command at runtime based on configured one or more parameters of the voice-command module.

Claims (71)

1. A computer-implemented method comprising:

receiving, through a user interface (UI) of an artificial-reality (AR) design tool, a selection of a configurable interface element to place the AR design tool and the UI into a configure phase to configure an AR effect;

receiving, through the UI of the AR design tool after the AR design tool and the UI are placed into the configure phase in response to selecting the configurable interface element, instructions to add a voice-command module to the AR effect;

configuring, while the AR design tool and the UI are placed into the configure phase, one or more parameters of the voice-command module to generate the AR effect based on one or more voice commands; and

generating the AR effect utilizing a particular voice command at runtime based on configured one or more parameters of the voice-command module.

2. The computer-implemented method of claim 1 , wherein the one or more parameters of the voice-command module comprise:

an intent type;

at least one slot; and

one or more entities associated with the at least one slot.

3. The computer-implemented method of claim 1 , wherein generating the AR effect comprises:

determining that the particular voice command at runtime corresponds to the configured one or more parameters of the voice-command module;

selecting a particular parameter of the configured one or more parameters of the voice-command module as a runtime value; and

accessing an AR asset corresponding to selected runtime value.

4. The computer-implemented method of claim 3 , further comprising:

tracking a location of an object in a scene; and

presenting, for display, the AR asset corresponding to the selected runtime value in accordance with a tracked location of the object in the scene.

5. The computer-implemented method of claim 1 , further comprising:

tracking a location of an object in a scene,

wherein the AR effect is generated further based on a tracked location of the object in the scene.

6. The computer-implemented method of claim 1 , wherein configuring the one or more parameters of the voice-command module comprises:

receiving, through the UI, one or more slots for the voice-command module;

receiving, through the UI, one or more entities for the one or more slots; and

configuring the one or more slots to be associated with the one or more entities.

7. The computer-implemented method of claim 1 , wherein the particular voice command comprises one or more of an utterance or one or more words.

8. One or more computer-readable non-transitory non-volatile storage media embodying software that is operable when executed to:

receive, through a user interface (UI) of an artificial-reality (AR) design tool, a selection of a configurable interface element to place the AR design tool and the UI into a configure phase to configure an AR effect;

receive, through the UI of the AR design tool after the AR design tool and the UI are placed into the configure phase in response to selecting the configurable interface element, instructions to add a voice-command module to the AR effect;

configure, while the AR design tool and the UI are placed into the configure phase, one or more parameters of the voice-command module to generate the AR effect based on one or more voice commands; and

generate the AR effect utilizing a particular voice command at runtime based on configured one or more parameters of the voice-command module.

9. The computer-readable non-transitory non-volatile storage media of claim 8 , wherein the one or more parameters of the voice-command module comprise:

an intent type;

at least one slot; and

one or more entities associated with the at least one slot.

10. The computer-readable non-transitory non-volatile storage media of claim 8 , wherein generating the AR effect comprises:

determining that the particular voice command at runtime corresponds to the configured one or more parameters of the voice-command module;

selecting a particular parameter of the configured one or more parameters of the voice-command module as a runtime value; and

accessing an AR asset corresponding to selected runtime value.

11. The computer-readable non-transitory non-volatile storage media of claim 10 , wherein the software is further operable when executed to:

track a location of an object in a scene; and

present, for display, the AR asset corresponding to the selected runtime value in accordance with a tracked location of the object in the scene.

12. The computer-readable non-transitory non-volatile storage media of claim 8 , wherein the software is further operable when executed to:

track a location of an object in a scene,

wherein the AR effect is generated further based on a tracked location of the object in the scene.

13. The computer-readable non-transitory non-volatile storage media of claim 8 , wherein configuring the one or more parameters of the voice-command module comprises:

receiving, through the UI, one or more slots for the voice-command module;

receiving, through the UI, one or more entities for the one or more slots; and

configuring the one or more slots to be associated with the one or more entities.

14. The computer-readable non-transitory non-volatile storage media of claim 8 , wherein the particular voice command comprises one or more of an utterance or one or more words.

15. A system comprising: one or more processors; and a memory coupled to the processors comprising instructions executable by the processors, the processors being operable when executing the instructions to:

receive, through a user interface (UI) of an artificial-reality (AR) design tool, a selection of a configurable interface element to place the AR design tool and the UI into a configure phase to configure an AR effect;

receive, through the UI of the AR design tool after the AR design tool and the UI are placed into the configure phase in response to selecting the configurable interface element, instructions to add a voice-command module to the AR effect;

configure, while the AR design tool and the UI are placed into the configure phase, one or more parameters of the voice-command module to generate the AR effect based on one or more voice commands; and

generate the AR effect utilizing a particular voice command at runtime based on configured one or more parameters of the voice-command module.

16. The system of claim 15 , wherein the one or more parameters of the voice-command module comprise:

an intent type;

at least one slot; and

one or more entities associated with the at least one slot.

17. The system of claim 15 , wherein generating the AR effect comprises:

determining that the particular voice command at runtime corresponds to the configured one or more parameters of the voice-command module;

selecting a particular parameter of the configured one or more parameters of the voice-command module as a runtime value; and

accessing an AR asset corresponding to selected runtime value.

18. The system of claim 17 , wherein the processors are further operable when executing the instructions to:

track a location of an object in a scene; and

present, for display, the AR asset corresponding to the selected runtime value in accordance with a tracked location of the object in the scene.

19. The system of claim 15 , wherein the processors are further operable when executing the instructions to:

track a location of an object in a scene,

wherein the AR effect is generated further based on a tracked location of the object in the scene.

20. The system of claim 15 , wherein configuring the one or more parameters of the voice-command module comprises:

receiving, through the UI, one or more slots for the voice-command module;

receiving, through the UI, one or more entities for the one or more slots; and

configuring the one or more slots to be associated with the one or more entities.

Assignments (2)
CHANGE OF NAME Recorded Feb 1, 2024
From: FACEBOOK, INC.
To: META PLATFORMS, INC.
Reel/Frame 066436/0288 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 31, 2024
From: SMET, STEF MARC; VERLINDE, HANNES LUC HERMAN; SLATER, MICHAEL; BLACKBURNE, BENJAMIN PATRICK; HARIHARAN, RAM KUMAR; JIA, CHUNJIE; NISARAT, PRAKARN
To: FACEBOOK, INC.
Reel/Frame 066308/0076 →
Continuity (2)
Continuation 17138096 · Dec 30, 2020
Related Publication 20230401800A1 · Dec 14, 2023
References Cited (19)
US 9836691B1 · Narayanaswami et al. · 2017 [cited by applicant]
US 10175980B2 · Temam et al. · 2019 [cited by applicant]
US 10534607B2 · Temam et al. · 2020 [cited by applicant]
US 11081104B1 · Su et al. · 2021 [cited by applicant]
US 20080120111A1 · Doyle · 2008 [cited by examiner]
US 20150073798A1 · Karov et al. · 2015 [cited by applicant]
US 20180293221A1 · Finkelstein et al. · 2018 [cited by applicant]
US 20190004791A1 · Brebner · 2019 [cited by examiner]
US 20190034172A1 · Kostello · 2019 [cited by applicant]
US 20190103101A1 · Danila et al. · 2019 [cited by applicant]
US 20190228269A1 · Brent · 2019 [cited by examiner]
US 20190303110A1 · Brude · 2019 [cited by applicant]
US 20200005128A1 · Temam et al. · 2020 [cited by applicant]
US 20200118564A1 · Moniz et al. · 2020 [cited by applicant]
US 20210117623A1 · Aly et al. · 2021 [cited by applicant]
US 20210383800A1 · Pair · 2021 [cited by applicant]
Dettmers T., “Deep Learning in a Nutshell: Core Concepts,” Artificial Intelligence, Nov. 3, 2015, 10 pages. [cited by applicant]
International Search Report and Written Opinion for International Application No. PCT/US2021/065179 mailed Apr. 5, 2022, 11 pages. [cited by applicant]
Jiao Y., et al., High-Performance Machine Learning, IEEE, Feb. 18, 2020, pp. 136-138. [cited by applicant]