IP Library › Granted Patent US 12,328,569
Granted Patent B2
US 12,328,569 · App. 18/151,430 · Granted Jun 10, 2025

Transforming computer game audio using impulse response of a virtual 3D space generated by NeRF input to a convolutional reverberation engine

Inventor: Michael Taylor (San Mateo, CA)
Assignee: Sony Interactive Entertainment Inc.
H04S7/305G06T15/205
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,328,569
App. No.
18/151,430
Granted
Jun 10, 2025
Kind
B2
Abstract

A 3D neural radiance field (NeRF) is used to generate an impulse response (IR) characterization, which can then be input to a convolutional reverberation engine to create an audio experience that reflects the in-game world on a 2.0 stereo speaker system. The NeRF recreates a background geometry and the impulse response of a virtual 3D space generated using NeRF is input to the convolutional reverberation engine to transform game sounds/music to appear as though they are occurring inside the 3D space of the game. The same may be done for the player's real-world space in which the virtual IR and real IR are blended together in real-time, and the real-life player tracked as he moves around the room to create audio as it would sound were the player moving through the virtual space while adjusting for the acoustics of the real space.

Claims (45)

1. A device comprising:

at least one computer storage that is not a transitory signal and that comprises instructions executable by at least one processor to:

generate a three dimensional (3D) neural radiance field (NeRF) from at least one image of a virtual space in a computer simulation;

use at least part of the NeRF to generate an impulse response (IR) characterizing sound transmission in the virtual space;

process audio from the computer simulation using the IR; and

play the audio after processing using the IR on at least one speaker.

2. The device of claim 1 , wherein the instructions are executable to process the audio from the computer simulation at least in part using at least one convolutional reverberation engine decoding the IR characterizing sound transmission in the virtual space.

3. The device of claim 1 , comprising the at least one processor.

4. The device of claim 1 , wherein the instructions are executable to:

for at least one object in the virtual space, correlate at least one surface characteristic to at least one acoustic reflection property; and

use the at least one acoustic reflection property to generate the IR.

5. The device of claim 4 , wherein the surface characteristic comprises at least one texture.

6. The device of claim 1 , wherein the at least one speaker comprises a speaker in a stereo speaker system.

7. The device of claim 1 , wherein the instructions are executable to:

access at least one image of a physical space in which the speaker is disposed;

based at least in part on the image of the physical space, generate a physical space IR; and

use the physical space IR to process sound from the computer simulation such that audio from the computer simulation is played as it would sound were a player moving through the virtual space.

8. The device of claim 7 , wherein the instructions are executable to:

based at least in part on the image of the physical space, generate a physical space IR at least in part using a NeRF representing the physical space.

9. An apparatus comprising:

at least one processor programmed with instructions to:

generate a neural radiance field (NeRF) representation of a virtual space from a computer simulation;

using at least one virtual sound source and at least one virtual microphone in the virtual space, generate at least one impulse response (IR) representation of the virtual space;

process audio from the computer simulation at least in part using the IR representation of the virtual space; and

play the audio from the computer simulation on at least one real world (RW) speaker in a RW space.

10. The apparatus of claim 9 , wherein the instructions are executable to process audio from the computer simulation at least in part using at least one convolutional reverberation engine programmed with the IR representation of the virtual space.

11. The apparatus of claim 9 , wherein the instructions are executable to:

for at least one object in the virtual space, correlate at least one surface characteristic of the object to at least one acoustic reflection property; and

use the at least one acoustic reflection property to generate the IR representation of the virtual space.

12. The apparatus of claim 11 , wherein the surface characteristic comprises at least one texture.

13. The apparatus of claim 9 , wherein the at least one RW speaker comprises a speaker in a stereo speaker system.

14. The apparatus of claim 9 , wherein the instructions are executable to:

access at least one image of a physical space in which the RW speaker is disposed;

based at least in part on the image of the physical space, generate a physical space IR; and

use the physical space IR to process sound from the computer simulation such that audio from the computer simulation is played as it would sound were a player moving through the virtual space.

15. The apparatus of claim 14 , wherein the instructions are executable to:

based at least in part on the image of the physical space, generate a physical space IR at least in part using a NeRF representing the physical space.

16. A method comprising:

generating a neural radiance field (NeRF) representation of a virtual space;

based at least in part on the NeRF representation, generating information representing acoustic transmission in the virtual space; and

playing, on at least one speaker, audio processed using the information representing acoustic transmission in the virtual space.

17. The method of claim 16 , wherein the information representing acoustic transmission in the virtual space comprises at least one impulse response (IR) element.

18. The method of claim 17 , comprising processing the audio using at least one convolutional reverberation engine programmed with the IR.

19. The method of claim 16 , wherein the virtual space and the audio are both sourced from a computer simulation.

20. The method of claim 16 , comprising processing the audio using at least one impulse response (IR) representation of a physical space in which the speaker is located.

Assignments (3)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 6, 2025
From: TAYLOR, MICHAEL
To: SONY INTERACTIVE ENTERTAINMENT INC.
Reel/Frame 071031/0121 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 25, 2025
From: TAYLOR, MICHAEL
To: SONY INTERACTIVE ENTERTAINMENT INC.
Reel/Frame 070621/0980 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 10, 2023
From: TAYLOR, MICHAEL
To: SONY INTERACTIVE ENTERTAINMENT INC.
Reel/Frame 062326/0097 →
Continuity (1)
Related Publication 20240236608A1 · Jul 11, 2024
References Cited (5)
US 20220101623A1 · Walsh · 2022 [cited by examiner]
US 20220139036A1 · Bertel et al. · 2022 [cited by applicant]
US 20220303715A1 · Watanabe et al. · 2022 [cited by applicant]
US 20230386202A1 · Suresh · 2023 [cited by examiner]
“International Search Report and Written Opinion”, dated Mar. 21, 2024, from the counterpart PCT application PCT/US23/82498. [cited by applicant]
Cited By (1)
US 12,444,140