IP Library Granted Patent US 12,394,146
Granted Patent B1
US 12,394,146 · App. 17/947,503 · Granted Aug 19, 2025

Methods and systems for composing and executing a scene

Inventors: Mark E. Drummond (Palo Alto, CA); Daniel L. Kovacs (Santa Clara, CA); Shaun D. Budhram (Los Gatos, CA); Edward Ahn (San Francisco, CA); Behrooz Mahasseni (San Jose, CA); Aashi Manglik (Sunnyvale, CA); Payal Jotwani (Santa Clara, CA); Mu Qiao (Campbell, CA); Bo Morgan (Emerald Hills, CA); Noah Gamboa (San Franciso, CA); Michael J. Gutensohn (San Francisco, CA); Dan Feng (Santa Clara, CA); Siva Chandra Mouli Sivapurapu (Santa Clara, CA)
Assignee: Apple Inc.
G06T17/00G06F3/013G06V10/987
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,394,146
App. No.
17/947,503
Granted
Aug 19, 2025
Kind
B1
Abstract

In one implementation, a method of composing a scene content is performed at a device including a display, one or more processors, and non-transitory memory. The method includes generating a definition of a scene based on textual or speech input and a model of a physical environment, wherein the definition includes a constraint that defines a spatial relationship between a virtual asset and an anchor asset that corresponds to one or more physical objects in the physical environment. The method includes generating, based on the definition of the scene and the model of the physical environment, a first instance of the scene that satisfies the constraint with the virtual asset in the spatial relationship with a first one of the one or more physical objects in the physical environment. The method includes presenting, on the display, the first instance of the scene.

Claims (44)

1. A method comprising:

at a device including a display, one or more processors, and non-transitory memory:

generating a definition of a scene based on textual or speech input and a model of an environment, wherein the definition of the scene includes a constraint that defines a spatial relationship between a virtual asset and an anchor asset that corresponds to one or more objects in the environment;

generating, based on application of the definition of the scene to a first physical environment, a first instance of the scene that satisfies the constraint with the virtual asset in the spatial relationship with a first one of the one or more objects in the first physical environment;

presenting, on the display, the first instance of the scene;

generating, based on application of the definition of the scene to a second physical environment, a second instance of the scene that satisfies the constraint with the virtual asset in the spatial relationship with a second one of the one or more objects in the second physical environment; and

presenting, on the display, the second instance of the scene.

2. The method of claim 1 , further comprising:

receiving, from a user, user feedback regarding the first instance of the scene; and

modifying the definition of the scene based on the user feedback.

3. The method of claim 2 , wherein the user feedback includes a positive indication or a negative indication.

4. The method of claim 2 , wherein the user feedback includes a numerical ranking.

5. The method of claim 2 , wherein the user feedback includes a preference between the first instance of the scene and a second instance of the scene.

6. The method of claim 2 , wherein the user feedback includes gaze information of the user.

7. The method of claim 2 , wherein modifying the definition of the scene includes adding or modifying a definition of a property of the virtual asset or the anchor asset.

8. The method of claim 7 , wherein adding or modifying the definition of the property is based on a user acceptance of a proposed addition or modification.

9. The method of claim 2 , wherein modifying the definition of the scene includes modifying a reward function.

10. The method of claim 2 , wherein modifying the definition of the scene includes modifying a neural network.

11. The method of claim 1 , wherein the second instance of the scene is displayed simultaneously with the first instance of the scene.

12. A device comprising:

a display;

non-transitory memory; and

one or more processors to:

obtain a definition of a scene, wherein the definition of the scene includes a constraint that defines a spatial relationship between a virtual asset and an anchor asset that corresponds to one or more objects in an environment;

generate, based on application of the definition of the scene to a first environment, a first instance of the scene that satisfies the constraint with the virtual asset in the spatial relationship with a first one of the one or more objects in the first environment;

generate, based on application of the definition of the scene to a second environment, a second instance of the scene that satisfies the constraint with the virtual asset in the spatial relationship with a second one of the one or more objects in the second environment; and

simultaneously present, on the display, the first instance of the scene and the second instance of the scene.

13. The device of claim 12 , wherein the one or more processors are further to:

receive, from a user, user feedback regarding the first instance of the scene and the second instance of the scene; and

modify the definition of the scene based on the user feedback.

14. The device of claim 13 , wherein the user feedback includes a preference between the first instance of the scene and the second instance of the scene.

15. The device of claim 13 , wherein the one or more processors are to modify the definition of the scene by adding or modifying a definition of a property of the virtual asset or the anchor asset.

16. The device of claim 12 , wherein the first environment is a first physical environment and the second environment is a first virtual environment.

17. A non-transitory memory having instructions encoded thereon which, when executed by one or more processors of a device, cause the device to:

obtain a definition of a scene, wherein the definition of the scene includes a definition of a property of an environment and a definition of a spatial relationship between a virtual asset and an anchor asset corresponding to one or more objects in the environment;

obtain a first virtual environment with the property;

present, on a display and based on application of the definition of the scene to the first virtual environment, a first instance of the scene including the virtual asset in the spatial relationship with a first object in the first virtual environment corresponding to the anchor asset;

obtain a second virtual environment with the property; and

present, on the display and based on application of the definition of the scene to the second virtual environment, a second instance of the scene including the virtual asset in the spatial relationship with a second object in the second virtual environment corresponding to the anchor asset.

18. The non-transitory memory of claim 17 , wherein the first virtual environment is based on a physical environment remote from the device.

19. The non-transitory memory of claim 17 , wherein the instructions, when executed, further cause the device to:

receive, from a user, user feedback regarding the first instance of the scene; and

modify the definition of the scene based on the user feedback.

20. The non-transitory memory of claim 19 , wherein the user feedback includes a positive indication or a negative indication.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 26, 2023
From: DRUMMOND, MARK E.; KOVACS, DANEIL L.; BUDHRAM, SHAUN D.; AHN, EDWARD; MAHASSENI, BEHROOZ; MANGLIK, AASHI; JOTWANI, PAYAL; QIAO, MU; MORGAN, BO; GAMBOA, NOAH; GUTENSOHN, MICHAEL J.; FENG, DAN; SIVAPURAPU, SIVA CHANDRA MOULI
To: APPLE INC.
Reel/Frame 065357/0708 →
Continuity (1)
Provisional Application 63246631 · Sep 21, 2021
References Cited (15)
US 9727996B2 · Pandey et al. · 2017 [cited by applicant]
US 9898873B2 · Yu · 2018 [cited by applicant]
US 9916002B2 · Petrovskaya et al. · 2018 [cited by applicant]
US 10203762B2 · Bradski et al. · 2019 [cited by applicant]
US 10459518B2 · Lutnick et al. · 2019 [cited by applicant]
US 10634913B2 · Nair et al. · 2020 [cited by applicant]
US 10719993B1 · Ha · 2020 [cited by applicant]
US 11989404B1 · McGinnis · 2024 [cited by examiner]
US 20140282220A1 · Wantland et al. · 2014 [cited by applicant]
US 20150185825A1 · Mullins · 2015 [cited by applicant]
US 20160253844A1 · Petrovskaya · 2016 [cited by examiner]
US 20210044636A1 · Miller · 2021 [cited by applicant]
US 20210110610A1 · Xu · 2021 [cited by examiner]
US 20240320489A1 · Vandikas · 2024 [cited by examiner]
EP 2546806B1 · 2019 [cited by applicant]