IP Library Granted Patent US 12,330,058
Granted Patent B2
US 12,330,058 · App. 17/899,766 · Granted Jun 17, 2025

Extraction and classification of audio events in gaming systems

Inventors: Nathan Souviraa-Labastie (Mons-en-Baroeul, FR); Damien Kevin Granger (Lille, FR); Raphaël Greff (Lille, FR)
Assignee: STEELSERIES ApS
A63F13/424A63F13/213A63F13/215A63F13/54A63F13/79A63F13/533A63F13/537A63F13/87A63F2300/308A63F2300/5546A63F2300/572A63F2300/6072A63F2300/6081G06N3/02G06N3/08G10L15/16G10L15/26H04R3/12H04R5/033H04R5/04H04S7/302H04S7/308H04S2400/01H04S2420/01
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,330,058
App. No.
17/899,766
Granted
Jun 17, 2025
Kind
B2
Abstract

A system that incorporates the subject disclosure may include, for example, receiving an input audio stream from a gaming system, the input audio stream including gaming audio of a video game played by a game player, the input audio stream including a plurality of classes of sounds, providing the input audio stream to a neural network, extracting, by the neural network, sounds of a selected class of sounds of the plurality of classes of sounds, and providing a plurality of output audio streams including providing a first audio stream including the sounds of the selected class of sounds of the input audio stream and a second audio stream including remaining sounds of the input audio stream. Additional embodiments are disclosed.

Claims (61)

1. A device, comprising:

a processing system including a processor; and

a memory that stores executable instructions that, when executed by the processing system, facilitate performance of operations, the operations comprising:

receiving an input audio stream from a gaming system, the input audio stream including gaming audio of a video game played by a game player, the input audio stream including a plurality of classes of sounds;

providing the input audio stream to a neural network;

extracting, by the neural network, sounds of a selected class of sounds of the plurality of classes of sounds;

providing a plurality of output audio streams including providing a first audio stream including the sounds of the selected class of sounds of the input audio stream and a second audio stream including remaining sounds of the input audio stream;

receiving control information from the game player;

modifying one or more of the plurality of output audio streams, forming modified output audio, wherein the modifying includes modifying a gain of sounds of a selected class of sounds according to the control information from the game player to form the modified output audio; and

providing the modified output audio to the game player.

2. The device of claim 1 , wherein the receiving an input audio stream from a gaming system comprises receiving sounds of the gaming audio having an apparent location, and wherein the providing a first audio stream comprises providing the sounds of the selected class of sounds with a same apparent location as the sounds of the gaming audio.

3. The device of claim 2 , wherein the receiving an input audio stream from a gaming system comprises receiving a multichannel audio signal from the gaming system.

4. The device of claim 1 , wherein the receiving an input audio stream from a gaming system comprises receiving surround sound audio from the gaming system and wherein the providing a first audio stream comprises providing the sounds of the selected class of sounds in surround sound audio.

5. The device of claim 2 , wherein the operations further comprise:

receiving footstep sounds having an apparent direction of origin in relation to the game player;

extracting the footstep sounds from the input audio stream; and

providing the first audio stream including the footstep sounds to the game player with same apparent direction of origin in relation to the game player.

6. The device of claim 1 , wherein the modified output audio includes footstep sounds.

7. The device of claim 6 , wherein the modified output audio includes gunshot sounds.

8. The device of claim 6 , wherein the modifying one or more of the plurality of output audio streams comprises:

equalizing sounds of a selected class of sounds according to the control information from the game player to form the modified output audio.

9. The device of claim 6 , wherein the modifying one or more of the plurality of output audio streams comprises:

modifying sounds of a selected class of sounds according to the control information from the game player, forming modified sounds of interest;

modifying sounds of the remaining sounds of the input audio stream according to the control information from the game player, forming modified remaining sounds;

combining the modified sounds of interest and the modified remaining sounds to form the modified output audio, forming combined output audio; and

providing the combined output audio to the game player.

10. The device of claim 9 , wherein the operations further comprise:

providing one of the combined output audio and the input audio stream to the game player according to the control information from the game player.

11. A non-transitory machine-readable medium, comprising executable instructions that, when executed by a processing system including a processor, facilitate performance of operations, the operations comprising:

receiving an input audio stream from a gaming system, the input audio stream including gaming audio of a video game played by a game player, the input audio stream including a plurality of classes of sounds;

extracting, from the input audio stream, sounds of a selected class of sounds of the plurality of classes of sounds;

receiving audio processing control information from the game player;

providing a plurality of output audio streams to the game player according to the audio processing control information, including selectively providing one of a first audio stream including the sounds of the selected class of sounds of the input audio stream and a second audio stream including remaining sounds of the input audio stream;

receiving gain control and equalization control information from the game player;

modifying a gain or a frequency spectrum of the sounds of the selected class of sounds according to the gain control and equalization control information from the game player, forming modified output sounds; and

providing the modified output sounds to the game player.

12. The non-transitory machine-readable medium of claim 11 , wherein the modified output sounds include at least one of footsteps or gunshot sounds.

13. The non-transitory machine-readable medium of claim 11 , wherein the operations further comprise:

selecting the selected class of sounds in response to the audio processing control information from the game player.

14. The non-transitory machine-readable medium of claim 11 , wherein the operations further comprise:

receiving a multichannel audio stream as the input audio stream from a gaming system, the multichannel audio stream including location information of sounds of the gaming audio of the video game; and

providing the plurality of output audio streams to the game player with same location information of the sounds of the gaming audio of the video game.

15. The non-transitory machine-readable medium of claim 11 , wherein the operations further comprise:

selecting one of footstep sounds and gunshot sounds of the gaming audio of the video game as the selected class of sounds, wherein the selecting is responsive to the audio processing control information from the game player.

16. The non-transitory machine-readable medium of claim 11 , wherein the operations further comprise:

providing the input audio stream to a neural network; and

receiving, from the neural network, the sounds of a selected class of sounds of the plurality of classes of sounds.

17. A method, comprising:

receiving, by a processing system including a processor, an input audio stream of a gaming system, the input audio stream including gaming audio of a video game played by a game player, the input audio stream including gaming sounds, the gaming sounds organizable as a plurality of classes of sounds;

providing, by the processing system, the input audio stream to a neural network, the neural network trained to recognize gaming sounds of at least one class of the plurality of classes of sounds;

receiving, by the processing system, from the neural network, extracted gaming sounds of the at least one class of the plurality of classes of sounds;

providing, by the processing system, to the game player, a first output audio stream including the extracted gaming sounds and a second output audio stream including remaining sounds of the gaming audio of the video game;

receiving, by the processing system, audio processing control information from the game player; and

modifying, by the processing system, the extracted gaming sounds according to the audio processing control information, wherein the modifying the extracted gaming sounds comprises:

enhancing, by the processing system, the extracted gaming sounds of the gaming audio of the video game; reducing, by the processing system, the remaining sounds of the gaming audio of the video game; or a combination thereof.

18. The method of claim 17 , wherein the extracted gaming sounds include footstep sounds.

19. The method of claim 18 , wherein the modifying the extracted gaming sounds comprises:

modifying, by the processing system, a gain of the extracted gaming sounds;

equalizing, by the processing system, the extracted gaming sounds; or

compressing, by the processing system, the extracted gaming sounds, or a combination of these.

20. The method of claim 18 , wherein the extracted gaming sounds include footstep sounds.

Assignments (2)
NUNC PRO TUNC ASSIGNMENT Recorded Jan 23, 2025
From: STEELSERIES APS
To: GN STORE NORD A/S
Reel/Frame 069988/0950 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 7, 2022
From: GRANGER, DAMIEN KEVIN; SOUVIRAA-LABASTIE, NATHAN; GREFF, RAPHAËL
To: STEELSERIES APS
Reel/Frame 061010/0874 →
Continuity (2)
Provisional Application 63240004 · Sep 2, 2021
Related Publication 20230064627A1 · Mar 2, 2023
References Cited (97)
US 7479063B2 · Pryzby et al. · 2009 [cited by applicant]
US 8811629B1 · Kulavik et al. · 2014 [cited by applicant]
US 9566505B2 · Perry · 2017 [cited by applicant]
US 9675871B1 · Jetter et al. · 2017 [cited by applicant]
US 9814936B1 · Bucolo · 2017 [cited by applicant]
US 9884258B2 · Huang et al. · 2018 [cited by applicant]
US 10237615B1 · Gudmundsson et al. · 2019 [cited by applicant]
US 10713543B1 · Skuin et al. · 2020 [cited by applicant]
US 11007445B2 · Schwarz et al. · 2021 [cited by applicant]
US 11011015B2 · Achmueller et al. · 2021 [cited by applicant]
US 11045727B2 · Mahlmeister et al. · 2021 [cited by applicant]
US 11071914B2 · Mahlmeister et al. · 2021 [cited by applicant]
US 11185786B2 · Mahlmeister et al. · 2021 [cited by applicant]
US 11260298B2 · Woidan et al. · 2022 [cited by applicant]
US 11311806B2 · Mahlmeister et al. · 2022 [cited by applicant]
US 11375256B1 · Dorner · 2022 [cited by applicant]
US 11484789B2 · Mahlmeister et al. · 2022 [cited by applicant]
US 11590420B2 · Mahlmeister et al. · 2023 [cited by applicant]
US 11766612B2 · Mahlmeister et al. · 2023 [cited by applicant]
US 20030044002A1 · Yeager et al. · 2003 [cited by applicant]
US 20030100359A1 · Loose et al. · 2003 [cited by applicant]
US 20050043090A1 · Pryzby et al. · 2005 [cited by applicant]
US 20050277469A1 · Pryzby et al. · 2005 [cited by applicant]
US 20080139301A1 · Holthe · 2008 [cited by applicant]
US 20080268961A1 · Brook et al. · 2008 [cited by applicant]
US 20080320545A1 · Schwartz · 2008 [cited by applicant]
US 20100040240A1 · Bonanno et al. · 2010 [cited by applicant]
US 20100100820A1 · Bryant et al. · 2010 [cited by applicant]
US 20100120533A1 · Bracken et al. · 2010 [cited by applicant]
US 20100261523A1 · Bonney et al. · 2010 [cited by applicant]
US 20110281645A1 · Wolfson et al. · 2011 [cited by applicant]
US 20110312424A1 · Burckart et al. · 2011 [cited by applicant]
US 20120014553A1 · Bonanno · 2012 [cited by applicant]
US 20120134651A1 · Cottrell · 2012 [cited by applicant]
US 20130331188A1 · Bonney et al. · 2013 [cited by applicant]
US 20140073429A1 · Meneses et al. · 2014 [cited by applicant]
US 20140153727A1 · Walsh et al. · 2014 [cited by applicant]
US 20140155171A1 · Laakkonen et al. · 2014 [cited by applicant]
US 20140179439A1 · Miura et al. · 2014 [cited by applicant]
US 20150005073A1 · Cudak et al. · 2015 [cited by applicant]
US 20150099586A1 · Huang et al. · 2015 [cited by applicant]
US 20150106711A1 · Virolainen · 2015 [cited by applicant]
US 20150141139A1 · Trombetta et al. · 2015 [cited by applicant]
US 20150217196A1 · McCarthy et al. · 2015 [cited by applicant]
US 20150222239A1 · Zhang et al. · 2015 [cited by applicant]
US 20150302684A1 · Loose et al. · 2015 [cited by applicant]
US 20160303474A1 · Kuruba Buchannagari et al. · 2016 [cited by applicant]
US 20170046906A1 · Hilbert et al. · 2017 [cited by applicant]
US 20170106283A1 · Malyuk et al. · 2017 [cited by applicant]
US 20180032845A1 · Polak et al. · 2018 [cited by applicant]
US 20180193747A1 · Bracken et al. · 2018 [cited by applicant]
US 20180310115A1 · Romigh · 2018 [cited by applicant]
US 20180353855A1 · Niemeyer et al. · 2018 [cited by applicant]
US 20180367484A1 · Rodriguez et al. · 2018 [cited by applicant]
US 20190122404A1 · Freeman et al. · 2019 [cited by applicant]
US 20190303796A1 · Balasubramanian et al. · 2019 [cited by applicant]
US 20190308096A1 · Kuruba Buchannagari et al. · 2019 [cited by applicant]
US 20200043282A1 · Keilwert et al. · 2020 [cited by applicant]
US 20200061477A1 · Mahlmeister et al. · 2020 [cited by applicant]
US 20200097312A1 · Kotteri et al. · 2020 [cited by applicant]
US 20200147487A1 · Mahlmeister et al. · 2020 [cited by applicant]
US 20200147489A1 · Mahlmeister et al. · 2020 [cited by applicant]
US 20200147500A1 · Mahlmeister et al. · 2020 [cited by applicant]
US 20200147501A1 · Mahlmeister et al. · 2020 [cited by applicant]
US 20200186897A1 · Dareddy · 2020 [cited by examiner]
US 20200188784A1 · Woidan et al. · 2020 [cited by applicant]
US 20200228911A1 · Baszucki · 2020 [cited by examiner]
US 20200242887A1 · Achmueller et al. · 2020 [cited by applicant]
US 20200353361A1 · Wang · 2020 [cited by applicant]
US 20210394061A1 · Mahlmeister et al. · 2021 [cited by applicant]
US 20220203232A1 · Wiggeshoff · 2022 [cited by applicant]
US 20220366170A1 · Wang et al. · 2022 [cited by applicant]
US 20220410004A1 · van Welzen · 2022 [cited by examiner]
US 20230060590A1 · Mahlmeister et al. · 2023 [cited by applicant]
US 20230060642A1 · Mahlmeister et al. · 2023 [cited by applicant]
US 20230063610A1 · Mahlmeister et al. · 2023 [cited by applicant]
US 20230066209A1 · Greff et al. · 2023 [cited by applicant]
US 20230067090A1 · Greff et al. · 2023 [cited by applicant]
US 20230068527A1 · Mahlmeister et al. · 2023 [cited by applicant]
US 20230330527A1 · Mahlmeister et al. · 2023 [cited by applicant]
US 20230364508A1 · Mahlmeister et al. · 2023 [cited by applicant]
US 20230405458A1 · Mahlmeister et al. · 2023 [cited by applicant]
US 20240066400A1 · Jeffrey et al. · 2024 [cited by applicant]
US 20240091638A1 · Greff et al. · 2024 [cited by applicant]
“PS5 3D Audio—Customize These Settings to Get the Perfect Sound for Your PlayStation 5 Gaming Setup”, https://www.youtube.com/watch?v=SjCv3-WNHGw, Jan. 11, 2021, 1 page. [cited by applicant]
Hennequin, Romain et al., “Spleeter: a fast and efficient music source separation tool with pre-trained models”, Journal of Open Source Software, 1 (5(50), 2154, https://doi.org/10.21105/joss.02154, Jun. 24, 2020, 4 pag… [cited by applicant]
Iwaya, “Individualization of head-related transfer functions with tournament-style listening test: Listening with other's ears”, Acoust. Sci. & Tech., vol. 27, No. 6, Nov. 2006, pp. 340-343., 5 pages. [cited by applicant]
Lundqvist, Jonas, “In a real game situation, is HRTF-enhanced game audio preferable over regular stereo game audio?”, Luleå University of Technology | Bachelor thesis | Audio Technology |Department of Music and media |D… [cited by applicant]
Mills, Matt, “What is HRTF: Listening to Positional Audio in Games”, https://itigic.com/hrtf-listening-to-positional-audio-in-games/, 7 pages. [cited by applicant]
Moore-Coyler, Roland, “PS5 and 3D audio: Everything you need to know”, https://www.tomsguide.com/features/ps5-and-3d-audio-everything-you-need-to-know, Oct. 6, 2020, 17 pages. [cited by applicant]
Poirier-Quinot, et al., “Impact of HRTF individualization on player performance in a VR shooter game I”, Conference on Spatial Reproduction, Aug. 2018, 7 pages. [cited by applicant]
Seeber, et al., “Subjective Selection of Non-Individual Head-Related Transfer Functions”, Proceedings of the 2003 International Conference on Auditory Display, Jul. 2003, 4 pages. [cited by applicant]
Shukla, et al., “User selection of optimal HRTF sets via holistic comparative evaluation”, Conference on Audio for Virtual and Augmented Reality, Aug. 2018., 10 pages. [cited by applicant]
“Teamspeak: 3D Sound—what is it?”, technical-tips, Nov. 11, 2020, technical-tips.com/blog/software/teamspeak-3d-sound-1201., Nov. 11, 2020, 2 pgs. [cited by applicant]
Rothbucher , et al., “Integrating a H RTF-based Sound Synthesis System into Mumble”, MMSP, pp. 24-28., Oct. 1, 2010, 5 pgs. [cited by applicant]
Int'l Search Report & Written Opinion for PCT/US2022/042292 mailed Jan. 25, 2023, 9 pages. [cited by applicant]
International Preliminary Report on Patentability for PCT/2022/042292 mailed Mar. 14, 2024. [cited by applicant]