IP Library Granted Patent US 12,571,898
Granted Patent B2
US 12,571,898 · App. 18/055,772 · Granted Mar 10, 2026

Systems and methods for using ultrawideband audio sensing systems

Inventors: Ziqi Wang (San Diego, CA); Mani Srivastava (Sherman Oaks, CA); Akash Deep Singh (Austin, TX); Luis Garcia (Midvale, UT); Zhe Chen (Singapore, SG); Jun Luo (Singapore, SG)
Assignees: The Regents of the University of California; Nanyang Technological University
G01S13/0209G01S7/2813G01S7/2923G01S7/414G01S13/888
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,571,898
App. No.
18/055,772
Granted
Mar 10, 2026
Kind
B2
Abstract

Systems and methods for simultaneously recovering and separate sounds from multiple sources using Impulse Radio Ultra-Wideband (IR-UWB) signals are described. In one embodiment, a device can be configured for generating an audio signal based on audio source ranging using ultrawideband signals. In an embodiment the device includes, a transmitter circuitry, a receiver circuitry, memory and a processor. The processor configured to generate a radio signal. The radio signal including an ultra-wideband Gaussian pulse modulated on a radio-frequency carrier. The processor further configured to transmit the radio signal using the transmitter circuitry, receive one or more backscattered signals at the receiver circuitry, demodulate the one or more backscattered signals to generate one or more baseband signals, and generate a set of data frames based on the one or more baseband signals.

Claims (51)

1 . A device for generating an audio signal based on audio source ranging using ultrawideband signals, the device comprising:

a transmitter circuitry;

a receiver circuitry;

memory; and

a processor, the processor configured to:

generate a radio signal comprising an ultra-wideband Gaussian pulse modulated on a radio-frequency carrier and transmit the radio frequency signal using the transmitter circuitry;

receive one or more backscattered signals at the receiver circuitry;

demodulate the one or more backscattered signals to generate one or more baseband signals, wherein generating one or more baseband signals comprises generating an in-phase component and a quadrature component, and wherein the set of frames is generated based on a highest spectrum concentration component, the component selected from a list of the in-phase component and the quadrature component;

generate a set of data frames based on the one or more baseband signals, wherein a data frame of the set of data frames is a time series generated based on backscattered signal data collected during a time interval, wherein the time interval is defined as a period from a transmission of the radio signal to a receipt of a last backscattered signal;

determine one or more distance bins associated with the one or more backscattered signals based on the set of data frames;

generate a slice based on a selected distance bin, the selected distance bin selected from among the determined one or more distance bins, wherein the slice includes a selected portion of the data from the set of frames, the portion of the data selected corresponds to the selected distance bin, the data defining a slice waveform; and

generate a sound waveform based on the generated slice.

2 . The device of claim 1 , wherein generating one or more baseband signals comprises generating an in-phase component and a quadrature component.

3 . The device of claim 1 , wherein the generated sound waveform has an amplitude proportional to an amplitude of the slice waveform associated with the generated slice.

4 . The device of claim 1 , wherein determining one or more distances associated with the one or more backscattered signals based on the set of frames comprises determining a time of flight for each of the one or more backscattered signals.

5 . The device of claim 1 , wherein the radio signal is an ultrawideband signal.

6 . The device of claim 1 , wherein the radio signal is an impulse radio ultrawideband signal.

7 . The device of claim 1 , wherein the transmitter circuitry is configured to transmit signals using an impulse radio ultra-wideband.

8 . The device of claim 1 , wherein the transmitter circuitry is configured to transmit a discontinuous wireless pulse signal.

9 . The device of claim 1 , wherein the transmitter circuitry is configured to use a sub 10 GHz range.

10 . The device of claim 1 , wherein the set of frames is a two-dimensional data structure.

11 . The device of claim 1 , wherein the slice is a one-dimensional data structure.

12 . The device of claim 1 , wherein generating a sound waveform based on the slice comprises generating a waveform having an amplitude proportional to the amplitude of a waveform described by data included in the slice.

13 . The device of claim 1 , wherein generating a sound waveform based on the slice comprises generating a waveform having an amplitude proportional to an amplitude of an in-phase part of the slice waveform.

14 . The device of claim 1 , wherein generating a sound waveform based on the slice comprises generating a waveform having an amplitude proportional to an amplitude of a quadrature part of the slice waveform.

15 . The device of claim 1 , wherein generating a sound waveform based on the slice comprises generating a waveform having an amplitude proportional to the amplitude of a waveform generated based on filtering the slice.

16 . The device of claim 1 , wherein the processor is further configured to perform a phase noise correction on the set of data frames, wherein performing the phase noise correction comprises:

generating a reference slice based on the set of data frames, the reference slice corresponding to a first distance bin, wherein the first distance bin is a closest distance bin to the transmitter circuitry;

determining a standard reference phase based on the reference slice; and

for each data frame of the set of data frames:

determining a phase difference between a data frame first distance bin and the determined standard reference phase; and

offsetting a phase error for all samples in the data frame based on the determined phase difference.

17 . The device of claim 16 , wherein calculating a standard reference phase based on the reference slice comprises calculating a mean phase of the reference slice.

18 . The device of claim 16 , wherein offsetting the phase error for all samples in the data frame comprises multiplying all samples in the data frame by e jΔφ , where e is Euler's number, j is a constant, and jΔφ is the determined phase difference.

19 . The device of claim 1 , wherein the processor is further configured to perform a static clutter suppression on the set of data frames, wherein performing the static clutter suppression comprises:

applying a Butterworth finite impulse response filter (FIR) on each distance bin, each distance bin corresponding to a distance bin portion of data from the set of data frames.

20 . The device of claim 19 , wherein the FIR is applied with a stopping frequency at 20 Hz.

21 . The device of claim 19 , wherein the FIR is applied with a passing frequency at 70 Hz.

22 . The device of claim 19 , wherein the FIR is applied using a stop-band attenuation set at −80 dB.

23 . The device of claim 19 , wherein performing static clutter suppression further comprises applying a high-pass filter.

24 . The device of claim 1 , wherein the processor is further configured to perform a vibrating activity localization on the set of data frames, wherein performing the vibrating activity localization comprises:

performing a discrete Fourier transform over each distance bin to generate a spectrum for each distance bin, each distance bin corresponding to a distance bin portion of data from the set of data frames;

generating a Herfindahl-Hirschman index for each distance bin based on the generated spectrums; and

selecting those distance bins which correspond to Herfindahl-Hirschman indices exceeding a threshold value.

25 . The device of claim 1 , wherein the processor is further configured to perform a denoising on the generated slice, wherein the denoising comprises:

applying spectral subtraction on the generated slice; and

applying a normalization on an output of the spectral subtraction.

26 . The device of claim 25 , wherein the applied spectral subtraction is an algorithm selected from the list of linear spectral subtraction, non-linear spectral subtraction, and multi-band spectral subtraction.

27 . The device of claim 1 , wherein the processor is further configured to perform audio recovery on the generated slice, wherein the audio recovery comprises:

outputting an audio file based on the generated slice.

28 . The device of claim 1 , wherein each frame of the set of frames is associated with the transmission of an RF probe pulse.

Assignments (3)
CONFIRMATORY LICENSE Recorded Feb 5, 2025
From: UNIVERSITY OF CALIFORNIA LOS ANGELES
To: NATIONAL SCIENCE FOUNDATION
Reel/Frame 070115/0791 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jun 30, 2023
From: CHEN, ZHE; LUO, JUN
To: NANYANG TECHNOLOGICAL UNIVERSITY
Reel/Frame 064129/0213 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 23, 2023
From: WANG, ZIQI; SRIVASTAVA, MANI; SINGH, AKASH DEEP; GARCIA, LUIS
To: THE REGENTS OF THE UNIVERSITY OF CALIFORNIA
Reel/Frame 063082/0389 →
Continuity (2)
Provisional Application 63264079 · Nov 15, 2021
Related Publication 20230288549A1 · Sep 14, 2023
References Cited (120)
US 7391819B1 · von der Embse · 2008 [cited by examiner]
US 12272375B2 · Ozturk · 2025 [cited by examiner]
US 20220292888A1 · Reiter · 2022 [cited by examiner]
US 20230003835A1 · Rong · 2023 [cited by examiner]
“Dcase: Sound event detection in domestic environments”, Retrieved from https://dcase.community/challenge2019/task-sound-event-detection-in-domestic-environments, retrieved on Feb. 20, 2025, 12 pgs. [cited by applicant]
“WARP v3 User Guide: RF Interfaces”, Retrieved from https://warpproject.org/trac/wiki/HardwareUsersGuides/WARPv3/RF, retrieved on Feb. 18, 2025, 3 pgs. [cited by applicant]
Abdel-Hamid et al., “Convolutional Neural Networks for Speech Recognition”, IEEE/ACM Transactions on Audio, Speech, and Language Processing, vol. 22, No. 10, Oct. 2014, pp. 1533-1545, doi: 10.1109/TASLP.2014.2339736. [cited by applicant]
Abowd et al., “Towards a Better Understanding of Context and Context-Awareness”, HUC '99: Proceedings of the 1st international symposium on Handheld and Ubiquitous Computing, Jan. 1999, pp. 304-307, doi: 10.1007/3-540-4… [cited by applicant]
Ahmad et al., “CarMap: Fast 3D Feature Map Updates for Automobiles”, Proceedings of the 17th USENIX Symposium on Networked Systems Design and Implementation (NSDI '20), Feb. 25-27, 2020, Santa Clara, CA, pp. 1063-1081. [cited by applicant]
Alanwar et al., “D-SLATS: Distributed Simultaneous Localization and Time Synchronization”, Mobihoc '17: Proceedings of the 18th ACM International Symposium on Mobile Ad Hoc Networking and Computing, Jul. 10-14, 2017, 10… [cited by applicant]
Amodei et al., “Deep Speech 2 : End-to-End Speech Recognition in English and Mandarin”, Proceedings of the 33rd International Conference on Machine Learning, vol. 48, 2016, pp. 173-182. [cited by applicant]
Andersen et al., “A 118-mW Pulse-Based Radar SoC in 55-nm CMOS for Non-Contact Human Vital Signs Detection”, IEEE Journal of Solid-State Circuits, vol. 52, No. 12, Dec. 2017, pp. 3421-3433, doi: 10.1109/JSSC.2017.276405… [cited by applicant]
Benedetto et al., “UWB Communication Systems: A Comprehensive Overview”, EURASIP Book Series on Signal Processing and Communications, vol. 5, 2006, 506 pgs. (Presented in 2 Parts). [cited by applicant]
Bengio et al., “Representation Learning: A Review and New Perspectives”, IEEE Transactions on Pattern Analysis and Machine Intelligence, vol. 35, No. 8, Aug. 2013, pp. 1798-1828, doi: 10.1109/TPAMI.2013.50. [cited by applicant]
Berouti et al., “Enhancement of Speech Corrupted by Acoustic Noise”, ICASSP'79, IEEE International Conference on Acoustics, Speech, and Signal Processing, vol. 4, Apr. 2-4, 1979, pp. 208-211, doi: 10.1109/ICASSP.1979.11… [cited by applicant]
Boll, Steven F. “A Spectral Subtraction Algorithm for Suppression of Acoustic Noise in Speech”, ICASSP '79, IEEE International Conference on Acoustics, Speech, and Signal Processing, Apr. 2-4, 1979, pp. 200-203, doi: 10… [cited by applicant]
Cai et al., “AcuTe: Acoustic Thermometer Empowered by a Single Smartphone”, SenSys '20: Proceedings of the 18th Conference on Embedded Networked Sensor Systems, Nov. 16-19, 2020, pp. 28-41, doi: 10.1145/3384419.3430714. [cited by applicant]
Chen et al., “Bathroom Activity Monitoring Based on Sound”, International Conference on Pervasive Computing, May 8-13, 2005, pp. 47-61, doi: 10.1007/11428572_4. [cited by applicant]
Chen et al., “ResNet and Model Fusion for Automatic Spoofing Detection”, Interspeech, Aug. 20-24, 2017, pp. 102-106, doi: 10.21437/Interspeech.2017-1085. [cited by applicant]
Chen et al., “Robust Deep Feature for Spoofing Detection—The SJTU System for ASV spoof 2015 Challenge”, Sixteenth Annual Conference of the International Speech Communication Association, Sep. 6-10, 2015, pp. 2097-2101. [cited by applicant]
Cheng et al., “Wide & Deep Learning for Recommender Systems”, Proceedings of the 1st Workshop on Deep Learning for Recommender Systems, Sep. 15, 2016, pp. 7-10, doi: 10.1145/2988450.2988454. [cited by applicant]
Davis et al., “The Visual Microphone: Passive Recovery of Sound from Video”, Journal of ACM Transactions on Graphics, vol. 33, No. 4, Article 79, Jul. 2014, 10 pgs. [cited by applicant]
Dekkers et al., “The Sins Database for Detection of Daily Activities in a Home Environment Using an Acoustic Sensor Network”, Proceedings of the Detection and Classification of Acoustic Scenes and Events, Nov. 16, 2017,… [cited by applicant]
Dhekne et al., “LiquID: A Wireless Liquid IDentifier”, MobiSys '18: Proceedings of the 16th Annual International Conference on Mobile Systems, Applications, and Services, Jun. 10-15, 2018, pp. 442-454, doi: 10.1145/3210… [cited by applicant]
Ding et al., “RF-Net: A Unified Meta-Learning Framework for RF-enabled One-Shot Human Activity Recognition”, SenSys '20: Proceedings of the 18th Conference on Embedded Networked Sensor Systems, Nov. 16-19, 2020, pp. 517… [cited by applicant]
Donahue et al., “Exploring Speech Enhancement with Generative Adversarial Networks for Robust Speech Recognition”, Proceedings of IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), Apr. 1… [cited by applicant]
Dotlic et al., “Angle of Arrival Estimation Using Decawave DW1000 Integrated Circuits”, Proceedings of 14th Workshop on Positioning, Navigation and Communications (WPNC), Oct. 25-26, 2017, 6 pgs., doi: 10.1109/WPNC.2017… [cited by applicant]
Esteva et al., “Dermatologist-level classification of skin cancer with deep neural networks”, Nature, vol. 542, No. 7639, Feb. 2, 2017, pp. 115-118, doi: 10.1038/nature21056. [cited by applicant]
Evans et al., “Spoofing and countermeasures for automatic speaker verification”, Interspeech, Aug. 25-29, 2013, pp. 925-929. [cited by applicant]
Germain et al., “Speech Denoising with Deep Feature Losses”, arXiv:1806.10522v2 [eess.AS], Sep. 14, 2018, pp. 1-6. [cited by applicant]
Goodfellow et al., “Deep Learning”, The MIT Press, Cambridge, Massachusetts, 2016, 66 pgs. [cited by applicant]
Griffin et al., “Localizing multiple audio sources in a wireless acoustic sensor network”, Journal of Signal Processing, vol. 107, Feb. 2015, pp. 54-67, doi: 10.1016/j.sigpro.2014.08.013. [cited by applicant]
He et al., “Deep Residual Learning for Image Recognition”, Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, 2016, pp. 770-778, doi: 10.1109/CVPR.2016.90. [cited by applicant]
He et al., “Delving Deep into Rectifiers: Surpassing Human-Level Performance on ImageNet Classification”, Proceedings of the IEEE International Conference on Computer Vision, Dec. 7-13, 2015, pp. 1026-1034. [cited by applicant]
Hollosi et al., “Voice Activity Detection Driven Acoustic Event Classification for Monitoring in Smart Homes”, 3rd International Symposium on Applied Sciences in Biomedical and Communication Technologies (ISABEL 2010), … [cited by applicant]
Huang et al., “WaveNet Vocoder and its Applications in Voice Conversion”, Conference on Computational Linguistics and Speech Processing ROCLING, 2018, pp. 96-110. [cited by applicant]
Ioffe et al., “Batch Normalization: Accelerating Deep Network Training by Reducing Internal Covariate Shift”, arXiv:1502.03167v3 [cs.LG], Mar. 2, 2015, pp. 1-11. [cited by applicant]
Jiménez et al., “Comparing Decawave and Bespoon UWB location systems: indoor/outdoor performance analysis”, International Conference on Indoor Positioning and Indoor Navigation (IPIN), Oct. 4-7, 2016, 8 pgs., doi: 10.11… [cited by applicant]
Juvela et al., “Speech Waveform Synthesis from MFCC Sequences with Generative Adversarial Networks”, IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), Apr. 15-20, 2018, pp. 5679-5683, do… [cited by applicant]
Kamath et al., “A Multi-band Spectral Subtraction Method for Enhancing Speech Corrupted by Colored Noise”, IEEE International Conference on Acoustics Speech and Signal Processing, vol. 4, May 13, 2002, pg. IV-4164, doi:… [cited by applicant]
Kavalerov et al., “Universal Sound Separation”, IEEE Workshop on Applications of Signal Processing to Audio and Acoustics (WASPAA), Oct. 20-23, 2019, pp. 175-179, doi: 10.1109/WASPAA.2019.8937253. [cited by applicant]
Kervel, Fredrik “X4driver example for raspbian”, Retrieved from https://github.com/novelda/Legacy-SW/tree/master/Examples/X4Driver_RaspberryPi, retrieved on Feb. 18, 2025, 3 pgs. [cited by applicant]
Kervel, Fredrik “XeThru X4 Phase Noise Correction”, Retrieved from https://github.com/novelda/Legacy-Documentation/blob/master/Application-Notes/XTAN-14_XeThru_X4_Phase_Noise_Correction_rev_a.pdf, Date Published—Sep. 27… [cited by applicant]
Kervel, Fredrik “XeThru X4 Radar User Guide”, Retrieved from https://github.com/novelda/Legacy-Documentation/blob/master/Application-Notes/XTAN-13_XeThruX4RadarUserGuide_rev_a.pdf, Date Published—Sep. 25, 2018, 13 pgs. [cited by applicant]
Kingma et al., “Adam: A Method for Stochastic Optimization”, arXiv:1412.6980v1 [cs.LG], Dec. 22, 2014, 9 pgs. [cited by applicant]
Kinnunen et al., “ASVspoof 2017: Automatic Speaker Verification Spoofing and Countermeasures Challenge Evaluation Plan”, Retrieved from https://www.researchgate.net/publication/364279170_ASVspoof_2017_Automatic_Speaker_… [cited by applicant]
Kinnunen et al., “t-DCF: a Detection Cost Function for the Tandem Assessment of Spoofing Countermeasures and Automatic Speaker Verification”, arXiv:1804.09618v1 [eess.AS], Apr. 25, 2018, 8 pgs. [cited by applicant]
Kshetrimayum, Rakhesh S. “An introduction to UWB communication systems”, IEEE Potentials, vol. 28, No. 2, Mar.-Apr. 2009, pp. 9-13. [cited by applicant]
Kwong et al., “Hard Drive of Hearing: Disks that Eavesdrop with a Synthesized Microphone”, IEEE Symposium on Security and Privacy (SP), May 19-23, 2019, pp. 905-919, doi: 10.1109/SP.2019.00008. [cited by applicant]
Lavrentyeva, “Audio replay attack detection with deep learning frameworks”, Interspeech, Aug. 20-24, 2017, pp. 82-86, doi: 10.21437/Interspeech.2017-360. [cited by applicant]
Lewis et al., “Effects of Noise on Speech Recognition and Listening Effort in Children With Normal Hearing and Children With Mild Bilateral or Unilateral Hearing Loss”, Journal of Speech, Language, and Hearing Research,… [cited by applicant]
Li et al., “An Overview of Noise-Robust Automatic Speech Recognition”, IEEE/ACM Transactions on Audio, Speech, and Language Processing, vol. 22, No. 4, Apr. 2014, pp. 745-777. [cited by applicant]
Liang et al., “Ultra-Wideband Impulse Radar Through-Wall Detection of Vital Signs”, Scientific Reports, vol. 8, No. 13367, Sep. 6, 2018, pp. 1-21, doi: 10.1038/s41598-018-31669-y. [cited by applicant]
Liu et al., “Design of Low-Power, 1GS/s Throughput FFT Processor for MIMO-OFDM UWB Communication System”, IEEE International Symposium on Circuits and Systems (ISCAS), May 27-30, 2007, pp. 2594-2597, doi: 10.1109/ISCAS.… [cited by applicant]
Lorenzo-Trueba et al., “The Voice Conversion Challenge 2018: Promoting Development of Parallel and Nonparallel Methods”, arXiv:1804.04262v1 [eess.AS], Apr. 12, 2018, 10 pgs. [cited by applicant]
McFee et al., “librosa: Audio and Music Signal Analysis in Python”, Proceedings of the 14th python in science conference, vol. 8, 2015, pp. 18-25. [cited by applicant]
Mesaros et al., “TUT Database for Acoustic Scene Classification and Sound Event Detection”, 24th European Signal Processing Conference (EUSIPCO), Aug. 29, 2016-Sep. 2, 2016, pp. 1128-1132, doi: 10.1109/EUSIPCO.2016.7760… [cited by applicant]
Miller, Dan “Voice Biometrics Census: Steady Growth of Global Enrollments”, Retrieved from https://www.nuance.com/content/dam/nuance/en_au/collateral/enterprise/report/ar-opus-vb-census-en-US.pdf?srsltid=AfmBOooQvPKR_ol… [cited by applicant]
Nachmani et al., “Voice Separation with an Unknown Number of Multiple Speakers”, arXiv:2003.01531v4 [eess.AS], Sep. 1, 2020, 12 pgs. [cited by applicant]
Nandakumar et al., “Contactless Sleep Apnea Detection on Smartphones”, MobiSys '15: Proceedings of the 13th Annual International Conference on Mobile Systems, Applications, and Services, May 18, 2015, pp. 45-57, doi: 10… [cited by applicant]
Nguyen et al., “Energy intelligent buildings based on user activity: A survey”, Energy and Buildings, vol. 56, Jan. 2013, pp. 244-257, doi: 10.1016/j.enbuild.2012.09.005. [cited by applicant]
Nishikawa et al., “Blind Source Separation of Acoustic Signals Based on Multistage ICA Combining Frequency-Domain ICA and Time-Domain ICA”, IEICE Transactions on Fundamentals of Electronics, Communications and Computer … [cited by applicant]
Oord et al., “Wavenet: A Generative Model for Raw Audio”, arXiv:1609.03499v2 [cs.SD], Sep. 19, 2016, pp. 1-15. [cited by applicant]
Paszke et al., “Automatic differentiation in PyTorch”, 31st Conference on Neural Information Processing Systems (NIPS), 2017, pp. 1-4. [cited by applicant]
Pavlidi et al., “Real-Time Multiple Sound Source Localization and Counting Using a Circular Microphone Array”, IEEE Transactions on Audio, Speech, and Language Processing, vol. 21, No. 10, Oct. 2013, pp. 2193-2206, doi:… [cited by applicant]
Qiu et al., “AVR: Augmented Vehicular Reality”, MobiSys '18: Proceedings of the 16th Annual International Conference on Mobile Systems, Applications, and Services, Jun. 10, 2018, pp. 81-95, doi: 10.1145/3210240.3210319. [cited by applicant]
Rajagopal et al., “Demo Abstract: Welcome to My World: Demystifying Multi-User AR with the Cloud”, 17th ACM/IEEE International Conference on Information Processing in Sensor Networks (IPSN), Apr. 11-13, 2018, pp. 146-14… [cited by applicant]
Rajpurkar et al., “Cardiologist-Level Arrhythmia Detection with Convolutional Neural Networks”, arXiv: 1707.01836v1 [cs.CV], Jul. 6, 2017, 9 pgs. [cited by applicant]
Reiff, Christian G. “Acoustic Source Localization and Cueing from an Aerostat during the NATO SET-093 Field Experiment”, Proceedings of SPIE, vol. 7333, Unattended Ground, Sea, and Air Sensor Technologies and Applicatio… [cited by applicant]
Rethage et al., “A Wavenet for Speech Denoising”, IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), Apr. 15-20, 2018, pp. 5069-5073, doi: 10.1109/ICASSP.2018.8462417. [cited by applicant]
Reynolds et al., “Robust Text-Independent Speaker Identification Using Gaussian Mixture Speaker Models”, IEEE Transactions on Speech and Audio Processing, vol. 3, No. 1, Jan. 1995, pp. 72-83, doi: 10.1109/89.365379. [cited by applicant]
Reynolds et al., “Speaker Verification Using Adapted Gaussian Mixture Models”, Digital Signal Processing, vol. 10, No. 1-3, Jan. 2000, pp. 19-41, doi: 10.1006/dspr.1999.0361. [cited by applicant]
Robjohns, Hugh “A brief history of microphones”, Microphone Data Book, 2001, pp. 1-7. [cited by applicant]
Ruiz et al., “Comparing Ubisense, BeSpoon, and DecaWave UWB Location Systems: Indoor Performance Analysis”, IEEE Transactions on Instrumentation and Measurement, vol. 66, No. 8, Aug. 2017, pp. 1-12, doi: 10.1109/TIM.201… [cited by applicant]
Sabath et al., “Definition and Classification of Ultra-Wideband Signals and Devices”, URSI Radio Science Bulletin, vol. 2005, No. 313, Jun. 2005, pp. 12-26, doi: 10.23919/URSIRSB.2005.7909522. [cited by applicant]
Saha et al., “Deep Convolutional Bidirectional LSTM for Complex Activity Recognition with Missing Data”, Human Activity Recognition Challenge, Smart Innovation, Systems and Technologies, vol. 199, Nov. 21, 2020, pp. 39-… [cited by applicant]
Salous et al., “Millimeter-Wave Propagation Characterization and modeling toward fifth-generation systems”, IEEE Antennas and Propagation Magazine, vol. 58, No. 6, Dec. 1, 2016, pp. 115-127, doi: 10.1109/MAP.2016.260981… [cited by applicant]
Sawada et al., “Blind Extraction of Dominant Target Sources Using ICA and Time-Frequency Masking”, IEEE Transactions on Audio, Speech, and Language Processing, vol. 14, No. 6, Nov. 2006, pp. 2165-2173, doi: 10.1109/TASL… [cited by applicant]
Shojaei et al., “Effect of signal to noise ratio on the speech perception ability of older adults”, Medical Journal of the Islamic Republic of Iran, vol. 30, Mar. 9, 2016, pp. 1-7. [cited by applicant]
Singh et al., “RadHAR: Human Activity Recognition from Point Clouds Generated through a Millimeter-wave Radar”, mmNets '19: Proceedings of the 3rd ACM Workshop on Millimeter-wave Networks and Sensing Systems, Oct. 7, 20… [cited by applicant]
Srivastava et al., “Dropout: A Simple Way to Prevent Neural Networks from Overfitting”, Journal of Machine Learning Research, vol. 15, No. 1, Jun. 1, 2014, pp. 1929-1958. [cited by applicant]
Stefanakis et al., “Perpendicular Cross-Spectra Fusion for Sound Source Localization with a Planar Microphone Array”, IEEE/ACM Transactions on Audio, Speech, and Language Processing, vol. 25, No. 9, Sep. 2017, pp. 1821-… [cited by applicant]
Szegedy et al., “Going Deeper with Convolutions”, IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Jun. 7-12, 2015, pp. 1-9, doi: 10.1109/CVPR.2015.7298594. [cited by applicant]
Tamai et al., “Three Ring Microphone Array for 3D Sound Localization and Separation for Mobile Robot Audition”, IEEE/RSJ International Conference on Intelligent Robots and Systems, Aug. 2-6, 2005, pp. 4172-4177, doi: 10… [cited by applicant]
Thomas et al., “Revisiting Trilateration for Robot Localization”, IEEE Transactions on Robotics, vol. 21, No. 1, Feb. 7, 2005, pp. 93-101, doi: 10.1109/TRO.2004.833793. [cited by applicant]
Toda et al., “The Voice Conversion Challenge 2016”, Interspeech, Sep. 8-12, 2016, pp. 1632-1636, doi: 10.21437/Interspeech.2016-1066. [cited by applicant]
Todisco et al., “ASV spoof 2019: Future Horizons in Spoofed and Fake Audio Detection”, Interspeech, Sep. 15-19, 2019, pp. 1008-1012, doi: 10.21437/Interspeech.2019-2249. [cited by applicant]
Todisco et al., “Constant Q cepstral coefficients: A spoofing countermeasure for automatic speaker verification”, Computer Speech & Language, vol. 45, Sep. 2017, pp. 516-535, doi: 10.1016/j.csl.2017.01.001. [cited by applicant]
Turpault et al., “Sound Event Detection in Domestic Environments with Weakly Labeled Data and Soundscape Synthesis”, Workshop on Detection and Classification of Acoustic Scenes and Events, Oct. 25-26, 2019, 5 pgs. [cited by applicant]
Tzinis et al., “Improving Universal Sound Separation Using Sound Classification”, ICASSP 2020—2020 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), May 4-8, 2020, pp. 96-100, doi: 10.11… [cited by applicant]
Vacher et al., “Sound Detection and Classification for Medical Telesurvey”, 2nd Conference on Biomedical Engineering, Feb. 2004, 5 pgs. [cited by applicant]
Valin et al., “Robust Sound Source Localization Using a Microphone Array on a Mobile Robot”, IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS 2003) (Cat. No.03CH37453), Oct. 27-31, 2003, pp. 122… [cited by applicant]
Venkatesh et al., “Implementation and Analysis of Respiration-rate Estimation Using Impulse-based UWB”, IEEE Military Communications Conference, Oct. 17-20, 2005, pp. 3314-3320, doi: 10.1109/MILCOM.2005.1606167. [cited by applicant]
Villalba et al., “Spoofing Detection with DNN and One-class SVM for the ASVspoof 2015 Challenge”, Sixteenth Annual Conference of the International Speech Communication Association, Interspeech, pp. 2067-2071, doi: 10.21… [cited by applicant]
Viswanathan et al., “Blind Navigation Proposal using Sonar”, IEEE International Conference on Computer Graphics, Vision and Information Security (CGVIS), Nov. 2-3, 2015, pp. 151-156, doi: 10.1109/CGVIS.2015.7449912. [cited by applicant]
Wang, Qiaosong “Towards Real-time 3D Reconstruction using Consumer UAVs”, 28th Workshop on Information Technologies and Systems, Dec. 16-18, 2018, arXiv:1902.09733, 7 pgs. [cited by applicant]
Wang et al., “Relative phase information for detecting human speech and spoofed speech”, 16th Annual Conference of the International Speech Communication Association, Interspeech, Sep. 6-10, 2015, pp. 2092-2096. [cited by applicant]
Wang et al., “Syncope Detection in Toilet Environments Using Wi-Fi Channel State Information”, UbiComp '18: Proceedings of the 2018 ACM International Joint Conference and 2018 International Symposium on Pervasive and Ub… [cited by applicant]
Wang et al., “Target Classification and Localization in Habitat Monitoring”, IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP '03), Apr. 6-10, 2003, pp. 844-847, doi: 10.1109/ICASSP.2003… [cited by applicant]
Watanabe et al., “ESPnet: End-to-End Speech Processing Toolkit”, Interspeech, Sep. 2-6, 2018, pp. 2207-2211, doi: 10.21437/Interspeech.2018-1456. [cited by applicant]
Wei et al., “Acoustic Eavesdropping through Wireless Vibrometry”, MobiCom '15: Proceedings of the 21st Annual International Conference on Mobile Computing and Networking, Sep. 7, 2015, pp. 130-141, doi: 10.1145/2789168.… [cited by applicant]
Wisdom et al., “What's All the Fuss About Free Universal Sound Separation Data?”, arXiv:2011.00803v1, Nov. 2, 2020, 5 pgs. [cited by applicant]
Wu et al., “ASVspoof 2015: the First Automatic Speaker Verification Spoofing and Countermeasures Challenge”, Sixteenth Annual Conference of the International Speech Communication Association, Interspeech, Sep. 2015, 5 p… [cited by applicant]
Wu et al., “Merlin: An Open Source Neural Network Speech Synthesis System”, 9th ISCA Speech Synthesis Workshop, Sep. 13-15, 2016, pp. 202-207, doi: 10.21437/SSW.2016-33. [cited by applicant]
Xing et al., “DeepCEP: Deep Complex Event Processing Using Distributed Multimodal Information”, IEEE International Conference on Smart Computing (Smartcomp), Jun. 12-15, 2019, pp. 87-92, doi: 10.1109/SMARTCOMP.2019.0003… [cited by applicant]
Xu et al., “WaveEar: Exploring a mmWave-based Noise-resistant Speech Sensing for Voice-User Interface”, MobiSys '19: Proceedings of the 17th Annual International Conference on Mobile Systems, Applications, and Services,… [cited by applicant]
Yamagishi et al., “ASVspoof 2019: Automatic Speaker Verification Spoofing and Countermeasures Challenge Evaluation Plan”, Retrieved from https://www.asvspoof.org/asvspoof2019/asvspoof2019_evaluation_plan.pdf, Date Publi… [cited by applicant]
Yang et al., “Making Sense of Mechanical Vibration Period with Sub-millisecond Accuracy Using Backscatter Signals”, MobiCom '16: Proceedings of the 22nd Annual International Conference on Mobile Computing and Networking… [cited by applicant]
Ye et al., “Urban sound event classification based on local and global features aggregation”, Applied Acoustics, vol. 117, Feb. 2017, 11 pgs., doi: 10.1016/j.apacoust.2016.08.002. [cited by applicant]
Zavarehei, Esfandiar “Berouti Spectral Subtraction”, Retrieved from https://in.mathworks.com/matlabcentral/fileexchange/7653-berouti-spectral-subtraction, retrieved on Feb. 18, 2025, 3 pgs. [cited by applicant]
Zavarehei, Esfandiar “Boll Spectral Subtraction”, Retrieved from https://in.mathworks.com/matlabcentral/fileexchange/7675-boll-spectral-subtraction, retrieved on Feb. 18, 2025, 3 pgs. [cited by applicant]
Zavarehei, Esfandiar “Multi-band Spectral Subtraction”, Retrieved from https://in.mathworks.com/matlabcentral/fileexchange/7674-multi-band-spectral-subtraction, retrieved on Feb. 18, 2025, 3 pgs. [cited by applicant]
Zhang et al., “Accurate UWB Indoor Localization System Utilizing Time Difference of Arrival Approach”, IEEE Radio and Wireless Symposium, Oct. 17-19, 2006, pp. 515-518, doi: 10.1109/RWS.2006.1615207. [cited by applicant]
Zhang et al., “BreathTrack: Tracking Indoor Human Breath Status via Commodity WiFi”, IEEE Internet of Things Journal, vol. 6, No. 2, Apr. 2019, pp. 3899-3911, doi: 10.1109/JIOT.2019.2893330. [cited by applicant]
Zhang et al., “Character-level Convolutional Networks for Text Classification”, arXiv:1509.01626v3 [cs.LG], Apr. 4, 2016, Advances in Neural Information Processing Systems, 2015, pp. 649-657. [cited by applicant]
Zhang et al., “Vibrosight: Long-Range Vibrometry for Smart Environment Sensing”, UIST '18: Proceedings of the 31st Annual ACM Symposium on User Interface Software and Technology, Oct. 14-17, 2018, pp. 225-236, doi: 10.1… [cited by applicant]
Zheng et al., “V2iFi: in-Vehicle Vital Sign Monitoring via Compact RF Sensing”, Proceedings of the ACM on Interactive, Mobile, Wearable and Ubiquitous Technologies, vol. 4, No. 2, Jun. 15, 2020, 27 pgs., doi: 10.1145/33… [cited by applicant]
“X4C007—Datasheet, Ultra Wideband (UWB) Impuse Radar Sensor”, Novelda, Retrieved from: https://fcc.report/FCC-ID/2AD9Q-X4C007/4756670.pdf, 2020, 13 pgs. [cited by applicant]
Bennett, Rachel, “Introducing the Amazon Alexa Premium Far-Field Voice Development Kit”, Amazon.com, Retrieved from: https://developer.amazon.com/en-US/blogs/alexa/post/80facfd2-1176-4c4f-94ac-4c5c781011ca/amazon-alexa-… [cited by applicant]
Klemm et al., “Clinical trials of a UWB imaging radar for breast cancer”, Proceedings of the Fourth European Conference on Antennas and Propagation, Apr. 12-16, 2010, 4 pgs, doi: 10.14361/transcript.9783839415696.fm. [cited by applicant]