IP Library › Granted Patent US 12,299,402
Granted Patent B2
US 12,299,402 · App. 18/659,606 · Granted May 13, 2025

Techniques for out-of-domain (OOD) detection

Inventors: Thanh Long Duong (Melbourne, AU); Mark Edward Johnson (Sydney, AU); Vishal Vishnoi (Redwood City, CA); Crystal C. Pan (Palo Alto, CA); Vladislav Blinov (Melbourne, AU); Cong Duy Vu Hoang (Melbourne, AU); Elias Luqman Jalaluddin (Seattle, WA); Duy Vu (Melbourne, AU); Balakota Srinivas Vinnakota (Sunnyvale, CA)
Assignee: Oracle International Corporation
G06F40/30G06F40/289G06N20/00H04L51/02G06F40/205
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,299,402
App. No.
18/659,606
Granted
May 13, 2025
Kind
B2
Abstract

The present disclosure relates to techniques for identifying out-of-domain utterances. One particular technique includes receiving an utterance and a target domain of a chatbot, generating a sentence embedding for the utterance, obtaining an embedding representation for each cluster of in-domain utterances associated with the target domain, predicting, using a metric learning model, a first probability that the utterance belongs to the target domain based on a similarity or difference between the sentence embedding and each embedding representation for each cluster, predicting, using an outlier detection model, a second probability that the utterance belongs to the target domain based on a determined distance or density deviation between the sentence embedding and embedding representations for neighboring clusters, evaluating the first probability and the second probability to determine a final probability, and classifying the utterance as in-domain or out-of-domain for the chatbot based on the final probability.

Claims (37)

1. A computer-implemented method comprising:

generating an embedding for an input utterance;

predicting, based on the embedding and a distance or density algorithm, a first probability as to whether the input utterance belongs to a target domain of a chatbot;

predicting, based on the embedding and a deep learning network, a second probability as to whether the input utterance belongs to the target domain of the chatbot;

determining, based on the first probability and the second probability, a final probability as to whether the input utterance belongs to the target domain of the chatbot; and

classifying, based on the first probability and the second probability, the input utterance is an in-domain utterance or out-of-domain utterance for the chatbot.

2. The method of claim 1 , wherein the first probability is predicted based on a distance or density deviation between the embedding and embedding representations for clusters of in-domain utterances associated with the target domain.

3. The method of claim 2 , wherein each embedding representation of the embedding representations for the clusters is associated with a respective cluster of the clusters.

4. The method of claim 2 , wherein the embedding representations for the clusters are obtained by generating a sentence embedding for each in-domain utterance of the in-domain utterances and inputting the sentence embedding for each in-domain utterance into an unsupervised clustering model configured to identify clusters.

5. The method of claim 1 , wherein the second probability is predicted based on a similarity or difference between the embedding and embedding representations for clusters of in-domain utterances associated with the target domain.

6. The method of claim 5 , wherein each embedding representation of the embedding representations for the clusters is associated with a respective cluster of the clusters.

7. The method of claim 1 , wherein the deep learning network comprises at least one of a stacked highway network and a combination of a linear model and a deep neural network.

8. A computer-program product tangibly embodied in a non-transitory machine-readable storage medium, including instructions configured to cause one or more data processors to perform actions including:

generating an embedding for an input utterance;

predicting, based on the embedding and a distance or density algorithm, a first probability as to whether the input utterance belongs to a target domain of a chatbot;

predicting, based on the embedding and a deep learning network, a second probability as to whether the input utterance belongs to the target domain of the chatbot;

determining, based on the first probability and the second probability, a final probability as to whether the input utterance belongs to the target domain of the chatbot; and

classifying, based on the first probability and the second probability, the input utterance is an in-domain utterance or out-of-domain utterance for the chatbot.

9. The computer-program product of claim 8 , wherein the first probability is predicted based on a distance or density deviation between the embedding and embedding representations for clusters of in-domain utterances associated with the target domain.

10. The computer-program product of claim 9 , wherein each embedding representation of the embedding representations for the clusters is associated with a respective cluster of the clusters.

11. The computer-program product of claim 9 , wherein the embedding representations for the clusters are obtained by generating a sentence embedding for each in-domain utterance of the in-domain utterances and inputting the sentence embedding for each in-domain utterance into an unsupervised clustering model configured to identify clusters.

12. The computer-program product of claim 8 , wherein the second probability is predicted based on a similarity or difference between the embedding and embedding representations for clusters of in-domain utterances associated with the target domain.

13. The computer-program product of claim 12 , wherein each embedding representation of the embedding representations for the clusters is associated with a respective cluster of the clusters.

14. The computer-program product of claim 8 , wherein the deep learning network comprises at least one of a stacked highway network and a combination of a linear model and a deep neural network.

15. A system comprising:

one or more data processors; and

a non-transitory computer readable storage medium containing instructions which, when executed on the one or more data processors, cause the one or more data processors to perform actions including:

generating an embedding for an input utterance;

predicting, based on the embedding and a distance or density algorithm, a first probability as to whether the input utterance belongs to a target domain of a chatbot;

predicting, based on the embedding and a deep learning network, a second probability as to whether the input utterance belongs to the target domain of the chatbot;

determining, based on the first probability and the second probability, a final probability as to whether the input utterance belongs to the target domain of the chatbot; and

classifying, based on the first probability and the second probability, the input utterance is an in-domain utterance or out-of-domain utterance for the chatbot.

16. The system of claim 15 , wherein the first probability is predicted based on a distance or density deviation between the embedding and embedding representations for clusters of in-domain utterances associated with the target domain.

17. The system of claim 16 , wherein each embedding representation of the embedding representations for the clusters is associated with a respective cluster of the clusters.

18. The system of claim 16 , wherein the embedding representations for the clusters are obtained by generating a sentence embedding for each in-domain utterance of the in-domain utterances and inputting the sentence embedding for each in-domain utterance into an unsupervised clustering model configured to identify clusters.

19. The system of claim 15 , wherein the second probability is predicted based on a similarity or difference between the embedding and embedding representations for clusters of in-domain utterances associated with the target domain.

20. The system of claim 19 , wherein each embedding representation of the embedding representations for the clusters is associated with a respective cluster of the clusters.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 10, 2024
From: DUONG, THANH LONG; JOHNSON, MARK EDWARD; VISHNOI, VISHAL; PAN, CRYSTAL C; BLINOV, VLADISLAV; HOANG, CONG DUY VU; JALALUDDIN, ELIAS LUQMAN; VU, DUY; VINNAKOTA, BALAKOTA SRINIVAS
To: ORACLE INTERNATIONAL CORPORATION
Reel/Frame 067370/0791 →
Continuity (4)
Continuation 18364298 · Aug 2, 2023
Continuation 17217909 · Mar 30, 2021
Provisional Application 63002139 · Mar 30, 2020
Related Publication 20240289555A1 · Aug 29, 2024
References Cited (34)
US 5912989A · Watanabe · 1999 [cited by applicant]
US 6061652A · Tsuboka et al. · 2000 [cited by applicant]
US 11003959B1 · Levner et al. · 2021 [cited by applicant]
US 11430467B1 · Vasudevan et al. · 2022 [cited by applicant]
US 11763092B2 · Duong et al. · 2023 [cited by applicant]
US 20020016798A1 · Sakai et al. · 2002 [cited by applicant]
US 20140244254A1 · Ju et al. · 2014 [cited by applicant]
US 20160103833A1 · Sanders et al. · 2016 [cited by applicant]
US 20170351965A1 · Kurniadi et al. · 2017 [cited by applicant]
US 20190050395A1 · Min · 2019 [cited by applicant]
US 20190147853A1 · Gunasekara et al. · 2019 [cited by applicant]
US 20190266287A1 · Chen et al. · 2019 [cited by applicant]
US 20200005118A1 · Chen et al. · 2020 [cited by applicant]
US 20200104650A1 · Huang · 2020 [cited by applicant]
US 20200179808A1 · Lee et al. · 2020 [cited by applicant]
US 20200285702A1 · Padhi · 2020 [cited by examiner]
US 20200320053A1 · He et al. · 2020 [cited by applicant]
US 20220199070A1 · Hsu · 2022 [cited by examiner]
US 20220253747A1 · Ren · 2022 [cited by examiner]
US 20220405682A1 · Yoon et al. · 2022 [cited by applicant]
US 20230040084A1 · Cherukara et al. · 2023 [cited by applicant]
JP 2005164836A · 2010 [cited by applicant]
JP 2010191046A · 2010 [cited by applicant]
JP 2017534941A · 2017 [cited by applicant]
JP 2018147189A · 2018 [cited by applicant]
JP 2019070957A · 2019 [cited by applicant]
Office Action received in Japanese Patent Application No. 2022-559631; 3 pages. [cited by applicant]
U.S. Appl. No. 17/217,909, Non-Final Office Action Action mailed on Apr. 18, 2023, 15 pages. [cited by applicant]
U.S. Appl. No. 17/217,909, Notice of Allowance Action mailed on Jun. 8, 2023, 27 pages. [cited by applicant]
U.S. Appl. No. 18/364,298, Notice of Allowance Action mailed on Apr. 10, 2024, 15 pages. [cited by applicant]
Kim , et al., Joint Learning of Domain Classification and Out-of-Domain Detection with Dynamic Class Weighting for Satisficing False Acceptance Rates, Available online at: https://arxiv.org/pdf/1807.00072.pdf, Jun. 29, … [cited by applicant]
Lane , et al., Out-of-Domain Utterance Detection Using Classification Confidences of Multiple Topics, Institute of Electrical and Electronics Engineers, Transactions on Audio, Speech, and Language Processing, vol. 15, N… [cited by applicant]
International Application No. PCT/US2021/024917, International Preliminary Report on Patentability mailed on Oct. 13, 2022, 9 pages. [cited by applicant]
International Application No. PCT/US2021/024917, International Search Report and Written Opinion mailed on Jul. 12, 2021, 13 pages. [cited by applicant]