IP Library Granted Patent US 11,823,472
Granted Patent B2
US 11,823,472 · App. 18/149,591 · Granted Nov 21, 2023

Arrangement for producing head related transfer function filters

Inventors: Tomi Huttunen (Kuopio, FI); Antti Vanne (Los Gatos, CA)
Assignee: Apple Inc.
G06V20/64G06F18/2135G06T3/0068G06T7/60G06T17/00G06T19/20G06V10/7715G06V40/10H04S7/303G06T2207/10012G06T2207/10016G06T2207/10028G06T2219/2004G06T2219/2016G06T2219/2021H04S1/005H04S2420/01
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,823,472
App. No.
18/149,591
Granted
Nov 21, 2023
Kind
B2
Abstract

When three-dimensional audio is produced by using headphones, particular HRTF-filters are used to modify sound for the left and right channels of the headphone. As the morphology of every ear is different, it is beneficial to have HRTF-filters particularly designed for the user of headphones. Such filters may be produced by deriving ear geometry from a plurality of images taken with an ordinary camera, detecting necessary features from images and fitting said features to a model that has been produced from accurately scanned ears comprising representative values for different sizes and shapes. Taken images are sent to a server ( 52 ) that performs the necessary computations and submits the data further or produces the requested filter.

Claims (35)

1. A method performed by a programmed processor of a mobile device, the method comprising:

capturing, using a camera of the mobile device, at least one image of a pinna;

determining at least one geometrical value of the pinna as output from a statistical model using input based on the at least one image, the statistical model is trained on a dataset including a plurality of pinna geometries produced from scanned ears of different sizes and shapes, wherein the statistical model determines the at least one geometrical value by reducing a set of data points representative of the pinna;

determining, based on the at least one geometrical value, a head-related transfer function (HRTF) filter; and

producing, using the HRTF filter, three-dimensional sound on the mobile device or transmitting the HRTF filter to an electronic device for production of the three-dimensional sound at the electronic device.

2. The method of claim 1 , wherein the mobile device is configured to instruct a user to move the mobile device around a head of the user to capture the at least one image.

3. The method of claim 1 , wherein the at least one image comprises a plurality of images captured using the camera, wherein the mobile device is configured to instruct a user to move the mobile device such that the plurality of images are captured at a plurality of different angles to cover the pinna.

4. The method of claim 1 , wherein the at least one geometrical value corresponds to at least one point of significance in the pinna.

5. The method of claim 4 , wherein the statistical model is trained on a dataset that have the at least one point of significance.

6. The method of claim 1 , wherein determining the at least one geometrical value comprises solving a minimization problem for basis functions of the statistical model.

7. The method of claim 1 , wherein the statistical model includes a pre-computed principle component analysis basis.

8. A non-transitory machine-readable medium storing instructions that, when executed by one or more processors of a mobile device, cause the mobile device to:

capture, using a camera of the mobile device, at least one image of a pinna;

determine at least one geometrical value of the pinna as output from a statistical model using input based on the at least one image, the statistical model is trained on a dataset including a plurality of pinna geometries produced from scanned ears of different sizes and shapes, wherein the statistical model determines the at least one geometrical value by reducing a set of data points representative of the pinna;

determine, based on the at least one geometrical value, a head-related transfer function (HRTF) filter; and

produce, using the HRTF filter, three-dimensional sound on the mobile device or transmit the HRTF filter to an electronic device for production of the three-dimensional sound at the electronic device.

9. The non-transitory machine-readable medium of claim 8 , wherein the mobile device is configured to instruct a user to move the mobile device around a head of the user to capture the at least one image.

10. The non-transitory machine-readable medium of claim 8 , wherein the at least one image comprises a plurality of images captured using the camera, wherein the mobile device is configured to instruct a user to move the mobile device such that the plurality of images are captured at a plurality of different angles to cover the pinna.

11. The non-transitory machine-readable medium of claim 8 , wherein the at least one geometrical value corresponds to at least one point of significance in the pinna.

12. The non-transitory machine-readable medium of claim 11 , wherein the statistical model is trained on a dataset that have the at least one point of significance.

13. The non-transitory machine-readable medium of claim 8 , wherein the instructions to determine the at least one geometrical value comprises instructions to solve a minimization problem for basis functions of the statistical model.

14. The non-transitory machine-readable medium of claim 8 , wherein the statistical model includes a pre-computed principle component analysis basis.

15. A mobile device comprising:

a camera;

a processor; and

memory having stored therein instructions which when executed by the processor causes the mobile device to:

capture, using the camera, at least one image of a pinna;

determine at least one geometrical value of the pinna as output from a statistical model using input based on the at least one image, the statistical model is trained on a dataset including a plurality of pinna geometries produced from scanned ears of different sizes and shapes, wherein the statistical model determines the at least one geometrical value by reducing a set of data points representative of the pinna;

determine, based on the at least one geometrical value, a head-related transfer function (HRTF) filter; and

produce, using the HRTF filter, three-dimensional sound on the mobile device or transmit the HRTF filter to an electronic device for production of the three-dimensional sound at the electronic device.

16. The mobile device of claim 15 , wherein the mobile device is configured to instruct a user to move the mobile device around a head of the user to capture the at least one image.

17. The mobile device of claim 15 , wherein the at least one image comprises a plurality of images captured using the camera, wherein the mobile device is configured to instruct a user to move the mobile device such that the plurality of images are captured at a plurality of different angles to cover the pinna.

18. The mobile device of claim 15 , wherein the at least one geometrical value corresponds to at least one point of significance in the pinna.

19. The mobile device of claim 18 , wherein the statistical model is trained on a dataset that have the at least one point of significance.

20. The mobile device of claim 15 , wherein determining the at least one geometrical value comprises solving a minimization problem for basis functions of the statistical model.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 20, 2023
From: OWNSURROUND OY
To: APPLE INC.
Reel/Frame 062441/0402 →
Priority Claims (1)
FI 20165211 · Mar 15, 2016 · national
Continuity (3)
Continuation 17073212 · Oct 16, 2020
Continuation 16084707
Related Publication 20230222819A1 · Jul 13, 2023