IP Library › Granted Patent US 12,681,917
Granted Patent B2
US 12,681,917 · App. 19/020,796 · Granted Jul 14, 2026

System and method for correction of a query using a replacement phrase

Inventors: Pranav Singh (Sunnvale, CA); Olivia Bettaglio (Santa Clara, CA)
Assignee: SOUNDHOUND AI IP, LLC
G06F16/2365G06F16/24522G06N7/00
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,681,917
App. No.
19/020,796
Filed
Jan 14, 2025
Granted
Jul 14, 2026
Kind
B2
Art Unit
2159
USPC
707/690
Abstract

Systems and methods are provided for natural language processing using neural network models and natural language virtual assistants. The system and method include receiving a natural language phrase including a word sequence, computing corresponding error probabilities that the words are errors, and for a word with a corresponding error probability above a threshold, then computing a replacement phrase with a low error probability to provide a response from the virtual assistant depending on the replacement phrase.

Claims (32)

1 . A virtual assistant comprising:

one or more processors;

a memory for storing software code, which, when executed by the one or more processors implements a trained neural network model, wherein training the model uses queries to detect an error and learn to correct the error, for natural language processing to identify errors for a query presented in a natural language form, wherein the virtual assistant calculates a vector distance between a first sentiment vector and a second sentiment vector and computes a transcription error probability and a natural language understanding misinterpretation probability using the trained neural network model, wherein the transcription error probability and the natural language understanding misinterpretation probability represent an error profile for the query.

2 . The virtual assistant of claim 1 , wherein probabilities are related to the vector distance and inversely related to an edit distance, and the transcription error probability exceeds the natural language understanding misinterpretation probability for large vector distances.

3 . A computer-implemented method comprising:

receiving a natural language query and transcribing it into a first word sequence;

using a statistical model, wherein training the model includes using natural language expressions identified as errors, on words within the first word sequence to compute corresponding error probabilities that the words are errors;

deriving a second word sequence having a replacement phrase for a word with a corresponding error probability above a threshold, the replacement phrase having a lower error probability;

wherein the replacement phrase is derived from a phonetic closeness score between a candidate replacement phrase and a hypothesized erroneous phrase.

4 . The computer-implemented method of claim 3 , further comprising the step of transmitting a virtual assistant query response depending on the replacement phrase.

5 . The computer-implemented method of claim 3 , further comprising receiving acoustic model scores for words within the first word sequence, wherein the corresponding error probability is inversely related to the acoustic model score.

6 . The computer-implemented method of claim 3 , further comprising computing a sentiment vector from the replacement phrase, wherein the virtual assistant query response depends on the sentiment vector.

7 . The computer-implemented method of claim 3 , further comprising:

computing a sentiment vector from the replacement phrase;

computing the distance between the computed sentiment vector and a sentiment vector from a previous natural language query; and

determining the virtual assistant query response based on the distance being below a threshold.

8 . The computer-implemented method of claim 3 , further comprising computing an error score by aggregating a plurality of error indicators.

9 . The computer-implemented method of claim 8 , further comprising normalizing weighting of error indicators using the error score.

10 . A computer-implemented method comprising:

transcribing a natural language query and into a first word sequence;

using a statistical model on words within the first word sequence to compute corresponding error probabilities that the words are errors, wherein training the model includes using natural language expressions identified as errors;

deriving a second word sequence having a replacement phrase for one or more words in the natural language query where the one or more words have a corresponding error probability above a threshold; and

computing a sentiment vector from the replacement phrase, wherein the virtual assistant query response depends on the sentiment vector.

11 . The computer-implemented method of claim 10 , wherein the replacement phrase is derived from a phonetic closeness score between a candidate replacement phrase and a hypothesized erroneous phrase.

12 . The computer-implemented method of claim 10 , further comprising the step of transmitting a virtual assistant query response depending on the replacement phrase.

13 . The computer-implemented method of claim 10 , further comprising receiving acoustic model scores for words within the first word sequence, wherein the corresponding error probability is inversely related to the acoustic model score.

14 . The computer-implemented method of claim 10 , further comprising:

computing a sentiment vector from the replacement phrase;

computing the distance between the computed sentiment vector and a sentiment vector from a previous natural language query; and

determining the virtual assistant query response based on the distance being below a threshold.

15 . The computer-implemented method of claim 10 , further comprising computing an error score by aggregating a plurality of error indicators.

16 . The computer-implemented method of claim 15 , further comprising normalizing weighting of error indicators using the error score.

Assignments (3)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 23, 2025
From: BETTAGLIO, OLIVIA; SINGH, PRANAV
To: SOUNDHOUND, INC.
Reel/Frame 070920/0961 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 23, 2025
From: SOUNDHOUND, INC.
To: SOUNDHOUND AI IP HOLDING, LLC
Reel/Frame 071009/0933 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 23, 2025
From: SOUNDHOUND AI IP HOLDING, LLC
To: SOUNDHOUND AI IP, LLC
Reel/Frame 071010/0032 →
Continuity (3)
Continuation 17581846 · Jan 21, 2022
Continuation 16561020 · Sep 5, 2019
Related Publication 20250156398A1 · May 15, 2025
References Cited (60)
US 5712957A · Waibel et al. · 1998 [cited by applicant]
US 7702512B2 · Gopinath et al. · 2010 [cited by applicant]
US 8660849B2 · Gruber et al. · 2014 [cited by applicant]
US 9953637B1 · Fabbrizio et al. · 2018 [cited by applicant]
US 10902490B2 · He · 2021 [cited by examiner]
US 20030216912A1 · Chino · 2003 [cited by applicant]
US 20040024601A1 · Gopinath et al. · 2004 [cited by applicant]
US 20040225650A1 · Cooper et al. · 2004 [cited by applicant]
US 20050159950A1 · Roth et al. · 2005 [cited by applicant]
US 20060206337A1 · Paek et al. · 2006 [cited by applicant]
US 20070073540A1 · Hirakawa et al. · 2007 [cited by applicant]
US 20080052073A1 · Goto et al. · 2008 [cited by applicant]
US 20090125299A1 · Wang · 2009 [cited by applicant]
US 20090228273A1 · Wang et al. · 2009 [cited by applicant]
US 20090326938A1 · Marila et al. · 2009 [cited by applicant]
US 20100125458A1 · Franco et al. · 2010 [cited by applicant]
US 20100271598A1 · Murayama et al. · 2010 [cited by applicant]
US 20110295897A1 · Gao et al. · 2011 [cited by applicant]
US 20120016678A1 · Gruber et al. · 2012 [cited by applicant]
US 20130179166A1 · Fujibayashi · 2013 [cited by applicant]
US 20130283168A1 · Brown et al. · 2013 [cited by applicant]
US 20140277735A1 · Breazeal · 2014 [cited by applicant]
US 20140310005A1 · Brown et al. · 2014 [cited by applicant]
US 20150039309A1 · Braho et al. · 2015 [cited by applicant]
US 20160063998A1 · Krishnamoorthy et al. · 2016 [cited by applicant]
US 20160179801A1 · Venkataraman et al. · 2016 [cited by applicant]
US 20160253989A1 · Kuo et al. · 2016 [cited by applicant]
US 20160260436A1 · Lemay et al. · 2016 [cited by applicant]
US 20160267128A1 · Dumoulin et al. · 2016 [cited by applicant]
US 20170229120A1 · Engelhardt · 2017 [cited by applicant]
US 20180315415A1 · Mosley et al. · 2018 [cited by applicant]
US 20180342233A1 · Li et al. · 2018 [cited by applicant]
US 20190035385A1 · Lawson et al. · 2019 [cited by applicant]
US 20190035386A1 · Leeb et al. · 2019 [cited by applicant]
US 20190050771A1 · Meharwade · 2019 [cited by examiner]
US 20200160866A1 · Szymanski · 2020 [cited by examiner]
JP 2001228894 · 2001 [cited by applicant]
JP 2002182680 · 2002 [cited by applicant]
JP 200524829 · 2005 [cited by applicant]
JP 2010044239 · 2010 [cited by applicant]
JP 2011002656 · 2011 [cited by applicant]
WO 2011028842 · 2011 [cited by applicant]
WO 2018083777 · 2018 [cited by applicant]
WO 2018160505 · 2018 [cited by applicant]
WO 2018217194 · 2018 [cited by applicant]
Ofek, Sentiment Analysis in Transcribed Utterance, pp. 27-38 (Year: 2015). [cited by examiner]
Larson, Outlier Detection for Improved Data Quality and Diversity in Dialog Systems June pp. 517-527 (Year: 2019). [cited by examiner]
Kumar, Sentiment analysis on speaker specific speech data, pp. 1-5 (Year: 2017). [cited by examiner]
Erica Sadun, et al., Talking to Siri: Mastering the Language of Apple's Intelligent Assistant, Third Edition, Mar. 2014, pp. 23-27. [cited by applicant]
Griol, David et al.; A framework for improving error detection and correction in spoken dialog systems; Group of Applied Artificial Intelligence {GIAA), Computer Science Department, Carlos III University of Madrid, Avda… [cited by applicant]
Laurent Prevot, A SIP of CoFee : A Sample of Interesting Productions of Conversational Feedback, Proceedings of the SIGDIAL 2015 Conference, pp. 149-153, Prague, Czech Republic, Sep. 2-4, 2015. [cited by applicant]
Levow, Gina-Anne; Characterizing and Recognizing Spoken Corrections in Human-Computer Dialogue; MIT AI Laboratory Room 769, 545 Technology Sq. Cambridge, MA 02139. [cited by applicant]
Matthias Scheutz, Robust Natural Language Dialogues for Instruction Tasks, Proceedings of SPIE, 2010. [cited by applicant]
Omar Zia Khan, Making Personal Digital Assistants Aware of What They Do Not Know, Interspeech, ISCA, Jul. 22, 2016. [cited by applicant]
Raveesh Meena, Data-driven Methods for Spoken Dialogue Systems, Doctoral Thesis, KTH Royal Institute of Technology, School of Computer Science and Communication, Department of Speech, Music and Hearing, 100 44 Stockholm… [cited by applicant]
Office Action dated Jan. 31, 2023 in U.S. Appl. No. 17/581,846. [cited by applicant]
Response to Office Action dated May 31, 2023 in U.S. Appl. No. 17/581,846. [cited by applicant]
Final Office Action dated Aug. 25, 2023 in U.S. Appl. No. 17/581,846. [cited by applicant]
Response to Final Office Action dated Jul. 11, 2024 in U.S. Appl. No. 17/581,846. [cited by applicant]
Notice of Allowance dated Sep. 11, 2024 in U.S. Appl. No. 17/581,846. [cited by applicant]