IP Library Granted Patent US 11,816,820
Granted Patent B2
US 11,816,820 · App. 17/711,808 · Granted Nov 14, 2023

Gaze direction-based adaptive pre-filtering of video data

Inventors: Can Jin (San Jose, CA); Nicolas Peirre Marie Frederic Bonnier (Campbell, CA); Hao Pan (Sunnyvale, CA)
Assignee: Apple Inc.
G06T5/20G02B27/017G06F3/013G06T19/006
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,816,820
App. No.
17/711,808
Granted
Nov 14, 2023
Kind
B2
Abstract

A multi-layer low-pass filter is used to filter a first frame of video data representing at least a portion of an environment of an individual. A first layer of the filter has a first filtering resolution setting for a first subset of the first frame, while a second layer of the filter has a second filtering resolution setting for a second subset. The first subset includes a data element positioned along a direction of a gaze of the individual, and the second subset of the frame surrounds the first subset. A result of the filtering is compressed and transmitted via a network to a video processing engine configured to generate a modified visual representation of the environment.

Claims (35)

1. A method, comprising:

obtaining, at a first device using a first set of sensors, an environment data set comprising a first video frame corresponding to a scene visible to an individual;

obtaining, at the first device using a second set of sensors, a behavior data set comprising representations of one or more behaviors of the individual;

filtering at least a portion of the environment data set using a multi-layer filter at the first device, wherein the filtering comprises (a) applying a first layer of the multi-layer filter to a first subset of the first video frame and (b) applying a second layer of the multi-layer filter to a second subset of the first video frame, wherein a filtering resolution of the first layer differs from a filtering resolution of the second layer; and

obtaining, at the first device, content to be displayed to the individual, wherein at least a portion of the content is generated at a second device based at least in part on (a) a result of the filtering and (b) the behavior data set.

2. The method as recited in claim 1 , wherein the one or more behaviors includes a hand gesture of the individual.

3. The method as recited in claim 1 , wherein the one or more behaviors includes a face gesture of the individual.

4. The method as recited in claim 1 , wherein the one or more behaviors includes a head movement of the individual.

5. The method as recited in claim 1 , wherein the behavior data set includes an indication of an expression of the individual.

6. The method as recited in claim 1 , wherein the first device comprises one or more of: (a) a wearable device or (b) a head-mounted display.

7. The method as recited in claim 1 , wherein the second device comprises a processing engine of one or more of: (a) a mixed reality application, (b) a virtual reality application or (c) an augmented reality application.

8. A system, comprising:

one or more processors; and

one or more memories;

wherein the one or more memories store program instructions that when executed on or across the one or more processors perform a method comprising:

obtaining, at a first device using a first set of sensors, an environment data set comprising a first video frame corresponding to a scene visible to an individual;

obtaining, at the first device using a second set of sensors, a behavior data set comprising representations of one or more behaviors of the individual;

filtering at least a portion of the environment data set using a multi-layer filter at the first device, wherein the filtering comprises (a) applying a first layer of the multi-layer filter to a first subset of the first video frame and (b) applying a second layer of the multi-layer filter to a second subset of the first video frame, wherein a filtering resolution of the first layer differs from a filtering resolution of the second layer; and

obtaining, at the first device, content to be displayed to the individual, wherein at least a portion of the content is generated at a second device based at least in part on (a) a result of the filtering and (b) the behavior data set.

9. The system as recited in claim 8 , wherein the one or more behaviors includes a hand gesture of the individual.

10. The system as recited in claim 8 , wherein the one or more behaviors includes a face gesture of the individual.

11. The system as recited in claim 8 , wherein the one or more behaviors includes a head movement of the individual.

12. The system as recited in claim 8 , wherein the behavior data set includes an indication of an expression of the individual.

13. The system as recited in claim 8 , wherein the first device comprises one or more of: (a) a wearable device or (b) a head-mounted display.

14. The system as recited in claim 8 , wherein the second device comprises a processing engine of one or more of: (a) a mixed reality application, (b) a virtual reality application or (c) an augmented reality application.

15. One or more non-transitory computer-accessible storage media storing program instructions that when executed on or across one or more processors cause the one or more processors to perform a method comprising:

obtaining, at a first device using a first set of sensors, an environment data set comprising a first video frame corresponding to a scene visible to an individual;

obtaining, at the first device using a second set of sensors, a behavior data set comprising representations of one or more behaviors of the individual;

filtering at least a portion of the environment data set using a multi-layer filter at the first device, wherein the filtering comprises (a) applying a first layer of the multi-layer filter to a first subset of the first video frame and (b) applying a second layer of the multi-layer filter to a second subset of the first video frame, wherein a filtering resolution of the first layer differs from a filtering resolution of the second layer; and

obtaining, at the first device, content to be displayed to the individual, wherein at least a portion of the content is generated at a second device based at least in part on (a) a result of the filtering and (b) the behavior data set.

16. The one or more non-transitory computer-accessible storage media as recited in claim 15 , wherein the one or more behaviors includes a hand gesture of the individual.

17. The one or more non-transitory computer-accessible storage media as recited in claim 15 , wherein the one or more behaviors includes a face gesture of the individual.

18. The one or more non-transitory computer-accessible storage media as recited in claim 15 , wherein the one or more behaviors includes a head movement of the individual.

19. The one or more non-transitory computer-accessible storage media as recited in claim 15 , wherein the behavior data set includes an indication of an expression of the individual.

20. The one or more non-transitory computer-accessible storage media as recited in claim 15 , wherein the second device comprises a processing engine of one or more of: (a) a mixed reality application, (b) a virtual reality application or (c) an augmented reality application.

Continuity (4)
Continuation 17112708 · Dec 4, 2020
Continuation 16040496 · Jul 19, 2018
Provisional Application 62535734 · Jul 21, 2017
Related Publication 20220222790A1 · Jul 14, 2022
Cited By (1)
US 12,579,613