IP Library › Granted Patent US 12,549,918
Granted Patent B2
US 12,549,918 · App. 18/517,862 · Granted Feb 10, 2026

Augmented reality virtual reality ray tracing sensory enhancement system, apparatus and method

Inventors: Joydeep Ray (Folsom, CA); Travis T. Schluessler (Hillsboro, OR); Prasoonkumar Surti (Folsom, CA); John H. Feit (Folsom, CA); Nikos Kaburlasos (Lincoln, CA); Jacek Kwiatkowski (Santa Clara, CA); Abhishek R. Appu (El Dorado Hills, CA); James M. Holland (Folsom, CA); Jeffery S. Boles (Folsom, CA); Jonathan Kennedy (Bristol, GB); Louis Feng (San Jose, CA); Atsuo Kuwahara (Hillsboro, OR); Barnan Das (San Jose, CA); Narayan Biswal (Folsom, CA); Stanley J. Baran (Elk Grove, CA); Gokcen Cilingir (Sunnyvale, CA); Nilesh V. Shah (Folsom, CA); Archie Sharma (Folsom, CA); Mayuresh M. Varerkar (Folsom, CA)
Assignee: Intel Corporation
H04S7/303G06F3/016G06T1/20G06T15/06G09B21/003G09B21/006G09B21/008H04R1/406H04R3/005H04S2400/11H04S2420/01
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,549,918
App. No.
18/517,862
Granted
Feb 10, 2026
Kind
B2
Abstract

Systems, apparatuses and methods may provide away to render augmented reality (AR) and/or virtual reality (VR) sensory enhancements using ray tracing. More particularly, systems, apparatuses and methods may provide a way to normalize environment information captured by multiple capture devices, and calculate, for an observer, the sound sources or sensed events vector paths. The systems, apparatuses and methods may detect and/or manage one or more capture devices and assign one or more the capture devices based on one or more conditions to provide observer an immersive VR/AR experience.

Claims (52)

1 . A system comprising:

a power source to supply power to the system;

a memory comprising environment information, the memory coupled to a processor; and

a graphics pipeline apparatus to:

normalize environment information to a position of an observer within a device coordinate space to generate normalized environment information relative to the observer, wherein the environment information is captured by one or more capture devices;

calculate, for the observer, ray tracing vector paths of sound sources or sensed events, wherein the ray tracing vector paths include vector path attributes that include one or more of positional information and directional information relative to the observer;

detect, for the capture devices, one or more of a capture device mode of operation, a capture device capability, a change to the capture device mode, and a change to the capture device capability based on one or more operating parameters wherein the one or more operating parameters include a priority ranking of one or more of multiple n-dimensional environments, the capture devices, the sound sources, and the sensed events; and

play back the normalized environment information to the observer as augmented reality sensory enhancements (AR) or virtual reality (VR) sensory enhancements based on the ray tracing vector paths based on the operating parameters.

2 . The system of claim 1 , wherein a number of ray bundles include a number of surround sound bands for surround sound type acoustic effects.

3 . The system of claim 2 , the graphics pipeline apparatus further to:

adjust a number of frequency bands within the ray bundles based on acoustic properties of one or more of the sound sources.

4 . The system of claim 2 , wherein the environment information including one more of sounds from one or more sound sources or sensed events from one or more event sources, wherein, during the playing back of the environment information, the graphics pipeline apparatus is to attenuate at least one of the sounds to produce a surround sound experience for the observer based on superposition of the ray bundles, wherein normalizing the environment information includes determining, by at least one of the capture devices, the position of the observer.

5 . The system of claim 4 , the graphics pipeline apparatus further to:

superimpose the ray bundles to visual information presented during playback, and wherein attenuating at least one of the sounds is based on the superposition of the ray bundles, produces a ray trace Doppler effect for the observer, by applying a relative motion filter to the normalized environment information.

6 . The system of claim 5 , the graphics pipeline apparatus further to:

adjust fidelity of the sound, by adjusting one or more of a number of ray bundles or a number of frequency bands based on the operating parameters.

7 . The system of claim 5 , wherein user preferences include render adjustments, wherein the graphics pipeline apparatus is trained during a training session with the observer, and

the system further comprising feedback devices, wherein the feedback devices include one or more speakers to output the sounds, or actuators to relay tactile information to a tactile surface based on one or more of the render adjustments or the sounds, wherein the tactile surface is two-dimensional or three-dimensional.

8 . The system of claim 4 , further comprising a capture device manager to:

assign one or more of the capture devices to capture at least a first microphone of the microphones to a first sound of the sounds or a first sensor of the sensors to a first sensed event of the sensed events based on one or more of user preferences, application parameters, attributes of the first sound or the first sensed event, or location or direction of travel the sound source of the first sound; and

assign at least a second microphone to the first sound or a second sensor to the first sensed event based on one or more of the user preferences, the application parameters, the attributes of the first sound or the first sensed event, a change in the location or the direction of travel of the sound source of the first sound, or the changes to the capture device capabilities.

9 . The system of claim 8 , the graphics pipeline apparatus further to:

attenuate output of one or more of the sounds based on one or more of the assigning of the one or more capture devices or the operating parameters, wherein the capture device manager adjusts the capture device mode of one or more of the capture devices based on the change to the capture device capabilities or the capture device mode, and wherein one of more of the graphics pipeline apparatus or the device capture manager are trained based on the operating parameters, the capture device mode, the capture device capabilities, the change to the capture device mode, or the change to the capture device capabilities.

10 . The system of claim 1 , the graphics pipeline apparatus further to:

apply two or more of an absorption filter, an attenuation filter, and a reflective filter to the normalized environment information to form filtered normalized environment information associated with acoustic properties of objects and surfaces, based on attributes of the ray tracing vector paths of the sound sources or sensed events.

11 . An apparatus comprising:

a memory comprising environment information; and

a graphics pipeline apparatus to:

normalize environment information to a position of an observer within a device coordinate space to generate normalized environment information relative to the observer, wherein the environment information is captured by one or more capture devices;

calculate, for the observer, ray tracing vector paths of sound sources or sensed events, wherein the ray tracing vector paths include vector path attributes that include one or more of positional information and directional information relative to the observer;

detect, for the capture devices, one or more of a capture device mode of operation, a capture device capability, a change to the capture device mode, and a change to the capture device capability based on one or more operating parameters, wherein the one or more operating parameters include a priority ranking of one or more of multiple n-dimensional environments, the capture devices, the sound sources, and the sensed events; and

play back the normalized environment information to the observer as augmented reality sensory enhancements (AR) or virtual reality (VR) sensory enhancements based on the ray tracing vector paths based on the operating parameters.

12 . The apparatus of claim 11 , wherein a number of ray bundles include a number of surround sound bands for surround sound type acoustic effects.

13 . The apparatus of claim 12 , the graphics pipeline apparatus further to:

adjust a number of frequency bands within the ray bundles based on acoustic properties of one or more of the sound sources.

14 . The apparatus of claim 12 , wherein the environment information including one more of sounds from one or more sound sources or sensed events from one or more event sources, wherein, during the playing back of the environment information, the graphics pipeline apparatus is to attenuate at least one of the sounds to produce a surround sound experience for the observer based on superposition of the ray bundles, wherein normalizing the environment information includes determining, by at least one of the capture devices, the position of the observer.

15 . The apparatus of claim 14 , the graphics pipeline apparatus further to:

superimpose the ray bundles to visual information presented during playback, and wherein attenuating at least one of the sounds is based on the superposition of the ray bundles, produces a ray trace Doppler effect for the observer, by applying a relative motion filter to the normalized environment information.

16 . The apparatus of claim 15 , the graphics pipeline apparatus further to:

adjust fidelity of the sound, by adjusting one or more of a number of ray bundles or a number of frequency bands based on the operating parameters.

17 . The apparatus of claim 15 , wherein user preferences include render adjustments, wherein the graphics pipeline apparatus is trained during a training session with the observer, and

the apparatus further comprising feedback devices, wherein the feedback devices include one or more speakers to output the sounds, or actuators to relay tactile information to a tactile surface based on one or more of the render adjustments or the sounds, wherein the tactile surface is two-dimensional or three-dimensional.

18 . The apparatus of claim 14 , further comprising a capture device manager to:

assign one or more of the capture devices to capture at least a first microphone of the microphones to a first sound of the sounds or a first sensor of the sensors to a first sensed event of the sensed events based on one or more of user preferences, application parameters, attributes of the first sound or the first sensed event, or location or direction of travel the sound source of the first sound; and

assign at least a second microphone to the first sound or a second sensor to the first sensed event based on one or more of the user preferences, the application parameters, the attributes of the first sound or the first sensed event, a change in the location or the direction of travel of the sound source of the first sound, or the changes to the capture device capabilities,

wherein the graphics pipeline apparatus is further to: attenuate output of one or more of the sounds based on one or more of the assigning of the one or more capture devices or the operating parameters, wherein the capture device manager adjusts the capture device mode of one or more of the capture devices based on the change to the capture device capabilities or the capture device mode, and wherein one of more of the graphics pipeline apparatus or the device capture manager are trained based on the operating parameters, the capture device mode, the capture device capabilities, the change to the capture device mode, or the change to the capture device capabilities.

19 . At least one non-transitory computer readable storage medium comprising a set of instructions, which when executed, cause a computing device to:

normalize environment information to a position of an observer within a device coordinate space to generate normalized environment information relative to the observer, wherein the environment information is captured by one or more capture devices;

calculate, for the observer, ray tracing vector paths of sound sources or sensed events, wherein the ray tracing vector paths include vector path attributes that include one or more of positional information and directional information relative to the observer;

detect, for the capture devices, one or more of a capture device mode of operation, a capture device capability, a change to the capture device mode, and a change to the capture device capability based on one or more operating parameters, wherein the one or more operating parameters include a priority ranking of one or more of multiple n-dimensional environments, the capture devices, the sound sources, and the sensed events; and

play back the normalized environment information to the observer as augmented reality sensory enhancements (AR) or virtual reality (VR) sensory enhancements based on the ray tracing vector paths based on the operating parameters.

20 . The least one non-transitory computer readable storage medium of claim 19 , wherein a number of ray bundles include a number of surround sound bands for surround sound type acoustic effects.

Continuity (5)
Continuation 17816960 · Aug 2, 2022
Continuation 17135850 · Dec 28, 2020
Continuation 16269778 · Feb 7, 2019
Continuation 15494742 · Apr 24, 2017
Related Publication 20240163631A1 · May 16, 2024
References Cited (110)
US 5450057A · Watanabe · 1995 [cited by applicant]
US 7146296B1 · Carlbom et al. · 2006 [cited by applicant]
US 8502864B1 · Watkins · 2013 [cited by applicant]
US 8933931B2 · Balan et al. · 2015 [cited by applicant]
US 9129443B2 · Gruen et al. · 2015 [cited by applicant]
US 9165399B2 · Uralsky et al. · 2015 [cited by applicant]
US 9177413B2 · Tatarinov et al. · 2015 [cited by applicant]
US 9241146B2 · Neill · 2016 [cited by applicant]
US 9262797B2 · Minkin et al. · 2016 [cited by applicant]
US 9342857B2 · Kubisch et al. · 2016 [cited by applicant]
US 9355483B2 · Lum et al. · 2016 [cited by applicant]
US 9437040B2 · Lum et al. · 2016 [cited by applicant]
US 10251011B2 · Joydeep et al. · 2019 [cited by applicant]
US 10880666B2 · Joydeep et al. · 2020 [cited by applicant]
US 11438722B2 · Ray et al. · 2022 [cited by applicant]
US 20010022830A1 · Sommer et al. · 2001 [cited by applicant]
US 20030202667A1 · Sekine · 2003 [cited by applicant]
US 20060001532A1 · Nagata · 2006 [cited by applicant]
US 20070008312A1 · Zhou et al. · 2007 [cited by applicant]
US 20090129603A1 · Cho · 2009 [cited by applicant]
US 20090161527A1 · Yu · 2009 [cited by applicant]
US 20110249029A1 · Baumgart · 2011 [cited by applicant]
US 20120327115A1 · Chhetri et al. · 2012 [cited by applicant]
US 20130113803A1 · Bakedash et al. · 2013 [cited by applicant]
US 20130169626A1 · Balan et al. · 2013 [cited by applicant]
US 20140063016A1 · Howson et al. · 2014 [cited by applicant]
US 20140118351A1 · Uralsky et al. · 2014 [cited by applicant]
US 20140125650A1 · Neill · 2014 [cited by applicant]
US 20140168035A1 · Luebke et al. · 2014 [cited by applicant]
US 20140168242A1 · Kubisch et al. · 2014 [cited by applicant]
US 20140168783A1 · Luebke et al. · 2014 [cited by applicant]
US 20140218390A1 · Rouet et al. · 2014 [cited by applicant]
US 20140253555A1 · Lum et al. · 2014 [cited by applicant]
US 20140267238A1 · Lum et al. · 2014 [cited by applicant]
US 20140267315A1 · Minkin et al. · 2014 [cited by applicant]
US 20140292771A1 · Kubisch et al. · 2014 [cited by applicant]
US 20140300636A1 · Miyazaya et al. · 2014 [cited by applicant]
US 20140347359A1 · Gruen et al. · 2014 [cited by applicant]
US 20140354675A1 · Lottes · 2014 [cited by applicant]
US 20150002508A1 · Tatarinov et al. · 2015 [cited by applicant]
US 20150009306A1 · Moore · 2015 [cited by applicant]
US 20150022537A1 · Lum et al. · 2015 [cited by applicant]
US 20150049104A1 · Lum et al. · 2015 [cited by applicant]
US 20150130915A1 · More et al. · 2015 [cited by applicant]
US 20150138065A1 · Alfierri · 2015 [cited by applicant]
US 20150138228A1 · Lum et al. · 2015 [cited by applicant]
US 20150170408A1 · He et al. · 2015 [cited by applicant]
US 20150170409A1 · He et al. · 2015 [cited by applicant]
US 20150187129A1 · Sloan · 2015 [cited by applicant]
US 20150194128A1 · Hicok · 2015 [cited by applicant]
US 20150264299A1 · Leech et al. · 2015 [cited by applicant]
US 20150289065A1 · Jensen et al. · 2015 [cited by applicant]
US 20150305701A1 · Wendler et al. · 2015 [cited by applicant]
US 20150317827A1 · Crassin et al. · 2015 [cited by applicant]
US 20150326966A1 · Mehra et al. · 2015 [cited by applicant]
US 20150332505A1 · Wang et al. · 2015 [cited by applicant]
US 20150378019A1 · Schissler et al. · 2015 [cited by applicant]
US 20160034248A1 · Schissler et al. · 2016 [cited by applicant]
US 20160048999A1 · Patney et al. · 2016 [cited by applicant]
US 20160049000A1 · Patney et al. · 2016 [cited by applicant]
US 20160071242A1 · Uralsky et al. · 2016 [cited by applicant]
US 20160071246A1 · Uralsky et al. · 2016 [cited by applicant]
US 20160093107A1 · Yamamoto et al. · 2016 [cited by applicant]
US 20160133055A1 · Fateh · 2016 [cited by applicant]
US 20160139666A1 · Rubin et al. · 2016 [cited by applicant]
US 20160140764A1 · Bickerstaff et al. · 2016 [cited by applicant]
US 20160142830A1 · Hu · 2016 [cited by applicant]
US 20160189427A1 · Wu et al. · 2016 [cited by applicant]
US 20160249989A1 · Devam et al. · 2016 [cited by applicant]
US 20170045941A1 · Tokubo et al. · 2017 [cited by applicant]
US 20170236332A1 · Kipman et al. · 2017 [cited by applicant]
US 20170311080A1 · Kolb et al. · 2017 [cited by applicant]
US 20170340394A1 · Gemmel et al. · 2017 [cited by applicant]
US 20180084359A1 · Lyren et al. · 2018 [cited by applicant]
CN 102254338A · 2011 [cited by applicant]
CN 103412045A · 2013 [cited by applicant]
Office Action issued for Patent Application No. CN 201810371353.0, mailed Feb. 2, 2024, 22 pages. [cited by applicant]
Nicholas Wilt, “The CUDA Handbook: A Comprehensive Guide to GPU Programming”, 522 pages, Jun. 2013, Addison-Wesley, USA. [cited by applicant]
Shane Cook, “CUDA Programming: A Developer's Guide to Parallel Computing with GPUs”, 591 pages, 2013, Elsevier, USA. [cited by applicant]
Mariella Moon, “Scientists are making VR displays that match your eyesight”, retrieved from engadget.com/2017/02/14/personalized-adjustable-vr-display, Feb. 14, 2017, 1 page. [cited by applicant]
Partial European Search Report for European Patent Application No. 18168352.5, mailed Oct. 31, 2018, 21 pages. [cited by applicant]
Damon Shing-Min Liu et al., “Visibility Preprocessing Suitable for Virtual Reality Sound Propagation with a Moving Receiver and Multiple Sources”, 2016 IEEE International Conference on Multimedia & Expo Workshops (ICMEW… [cited by applicant]
Okada et al., “A ray tracing Simulation of sound diffraction based on analytic secondary source model”, 19th European Signal Processing Conference, Aug. 29, 2011, p. 1653-1657, Barcelona Spain. [cited by applicant]
Extended European Search Report for European Patent Application No. 18168352.5, mailed Jan. 31, 2019, 17 pages. [cited by applicant]
Office Action for U.S. Appl. No. 15/494,742, mailed May 7, 2018, 22 pages. [cited by applicant]
Notice of Allowance for U.S. Appl. No. 15/494,742, mailed Nov. 7, 2018, 7 pages. [cited by applicant]
Office Action for U.S. Appl. No. 16/269,778, mailed Mar. 27, 2020, 16 pages. [cited by applicant]
“Visibility preprocessing suitable for virtual reality sound propagation with a moving receiver and multiple sources” Damon Shing-Min Liu et al., 2016 IEEE International Conference. [cited by applicant]
A Ray Tracing Simulation of Sound Diffraction Based on the Analytic Secondary Source Model, Masashi Okada et al., IEEE Transactions on Audio, Speech, and Language Processing Year: 2012, vol. 20, Issue: 9. [cited by applicant]
Notice of Allowance for U.S. Appl. No. 16/269,778, mailed Aug. 26, 2020, 7 pages. [cited by applicant]
Office Action for U.S. Appl. No. 17/135,850, mailed Jan. 5, 2022, 40 pages. [cited by applicant]
Notice of Allowance for U.S. Appl. No. 17/135,850, mailed May 2, 2022, 7 pages. [cited by applicant]
Brief Communication, EP App. No. 18168352.5, Feb. 16, 2024, 16 pages. [cited by applicant]
Brief Communication, EP App. No. 22209903.8, Jun. 17, 2024, 33 pages. [cited by applicant]
Decision of Rejection, CN App. No. 201810371353, Aug. 22, 2024, 16 pages (12 pages of English Translation and 4 pages of Original Document). [cited by applicant]
Decision to Refuse, EP App. No. 18168352.5, Mar. 14, 2024, 12 pages. [cited by applicant]
Decision to Refuse, EP App. No. 22209903.8, Jun. 21, 2024, 19 pages. [cited by applicant]
European Search Report, EP App. No. 22209903.8, Mar. 1, 2023, 3 pages. [cited by applicant]
Funkhouser, Thomas, “Survey of Methods for Modeling Sound Propagation in Interactive Virtual Environment Systems”, Department of Computer Science of Princeton University, Jan. 2003, 54 pages. [cited by applicant]
Mariette et al., “SoundDelta—Large Scale, Multi-User Audio Augmented Reality”, Proceedings of the EAA Symposium on Auralization, Jun. 15-17, 2009, pp. 1-6. [cited by applicant]
Non-Final Office Action, U.S. Appl. No. 17/816,960, Aug. 14, 2023, 8 pages. [cited by applicant]
Notice of Allowance, CN App. No. 201810371353, Feb. 18, 2025, 7 pages (3 pages of English Translation and 4 pages of Original Document). [cited by applicant]
Notice of Allowance, U.S. Appl. No. 17/816,960, Sep. 25, 2023, 7 pages. [cited by applicant]
Office Action, EP App. No. 18168352.5, Jun. 23, 2021, 16 pages. [cited by applicant]
Office Action, EP App. No. 22209903.8, Aug. 16, 2023, 9 pages. [cited by applicant]
Office Action, EP App. No. 22209903.8, Mar. 28, 2023, 8 pages. [cited by applicant]
Second Office Action, CN App. No. 201810371353, Jun. 19, 2024, 24 pages (14 pages of English Translation and 10 pages of Original Document). [cited by applicant]
Summons to Attend Oral Proceedings, EP App. No. 18168352.5, Sep. 18, 2023, 16 pages. [cited by applicant]
Summons to Attend Oral Proceedings, EP App. No. 22209903.8, Mar. 14, 2024, 10 pages. [cited by applicant]
Taylor et al., “Guided Multiview Ray Tracing for Fast Auralization”, IEEE Transactions on Visualization and Computer Graphics, vol. 18, No. 11, Nov. 2012, pp. 1797-1810. [cited by applicant]