IP Library › Granted Patent US 12,443,859
Granted Patent B2
US 12,443,859 · App. 17/807,653 · Granted Oct 14, 2025

Dialogue model training method and device therefor

Inventors: Seok Jun Seo (Seoul, KR); Seung Ju Han (Seoul, KR); Beom Su Kim (Seoul, KR); Bu Ru Chang (Seoul, KR); Enkhbayar Erdenee (Seoul, KR)
Assignee: Hyperconnect LLC
G06N5/022G06F16/3329G06F40/35G06N3/096G06N3/0455
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,443,859
App. No.
17/807,653
Granted
Oct 14, 2025
Kind
B2
Abstract

Disclosed is a method of training a dialogue model in an electronic device, the method including selecting a first context from a first dialogue data set including at least one pair of a context and a response corresponding to the context, generating a first response corresponding to the first context through a first dialogue model, generating an augmented dialogue dataset by incorporating a pair of the first context and the first response corresponding to the first context into the first dialogue data set, and training a second dialogue model based on the augmented dialogue dataset.

Claims (51)

1. A method of training a dialogue model in an electronic device, the method comprising:

selecting a first context from a first dialogue data set including at least one pair of a context and a response corresponding to the context;

generating a first response corresponding to the first context through a first dialogue model, wherein the first dialogue model is a generative-based dialogue model that generates a response to a given context;

generating an augmented dialogue dataset by incorporating a pair of the first context and the first response corresponding to the first context into the first dialogue data set;

generating an augmented response set including a response of the first dialogue data set and the first response; and

training a second dialogue model based on the augmented dialogue dataset, wherein the second dialogue model is a retrieval-based dialogue model that searches for a response to the given context, wherein the training comprises:

acquiring a response set including a first response subset corresponding to a second context included in the augmented dialogue dataset and a second response subset selected arbitrarily,

calculating a first score for a response included in the response set with respect to the second context based on the first dialogue model,

calculating a second score for a response included in the response set with respect to the second context based on the second dialogue model, and

training the second dialogue model based on the first score and the second score.

2. The method of claim 1 , wherein the training of the second dialogue model based on the first score and the second score comprises:

calculating a loss based on the first score and the second score; and

training the second dialogue model such that the loss is minimized.

3. A method of training a dialogue model in an electronic device, the method comprising:

acquiring, from a first dialogue data set, a response set including a first response subset corresponding to a first context and a second response subset selected arbitrarily;

calculating a first score for a response included in the response set with respect to the first context based on a first dialogue model, wherein the first dialogue model is a generative-based dialogue model that generates a response to a given context;

calculating a second score for a response included in the response set with respect to the first context based on a second dialogue model, wherein the second dialogue model is a retrieval-based dialogue model that searches for a response to the given context; and

training the second dialogue model based on the first score and the second score, wherein the training of the second dialogue model includes

calculating a loss based on the first score and the second score; and

training the second dialogue model such that the loss is minimized.

4. The method of claim 3 , further comprising:

selecting a second context from a second dialogue data set including at least one pair of a context and a response corresponding to the context;

generating a response corresponding to the second context through the first dialogue model; and

generating the first dialogue data set by incorporating a pair of the second context and a response corresponding to the second context into the second dialogue data set.

5. The method of claim 3 , wherein the calculating of the second score comprises:

encoding the first context and a response included in the response set based on fixed-length embedding; and

calculating a relevance score with respect to the first context for each response included in the response set based on an embedding value corresponding to the first context and an embedding value corresponding to each response included in the response set.

6. The method of claim 3 , wherein the first score is calculated using a normalized log-likelihood based on a length of each response included in the response set.

7. The method of claim 3 , wherein the first score is calculated based on a mutual information score with respect to the first context for each response included in the response set.

8. The method of claim 3 , wherein the loss comprises a cross-entropy loss for a score corresponding to the first response subset and a knowledge distillation loss for a score corresponding to a response included in the response set.

9. The method of claim 8 , wherein the training of the second dialogue model such that the loss is minimized comprises training by maximizing a score corresponding to the first response subset so that the cross-entropy loss is minimized.

10. The method of claim 8 , wherein the training of the second dialogue model such that the loss is minimized comprises training by matching the first score and the second score so that the knowledge distillation loss is minimized.

11. A non-transitory computer-readable recording medium comprising a computer program to execute the method of claim 3 .

12. An electronic device for training a dialogue model, the electronic device comprising:

a storage device; and

a controller,

wherein the controller is configured to

acquire a response set including a first response subset corresponding to a first context and a second response subset selected arbitrarily, from a first dialogue data set through the storage device;

calculate a first score for a response included in the response set with respect to the first context based on a first dialogue model, wherein the first dialogue model is a generative-based dialogue model that generates a response to a given context;

calculate a second score for a response included in the response set with respect to the first context based on a second dialogue model, wherein the second dialogue model is a retrieval-based dialogue model that searches for a response to the given context; and

train the second dialogue model based on the first score and the second score, wherein, to train the second dialogue model, the controller is configured to

calculate a loss based on the first score and the second score, and

train the second dialogue model such that the loss is minimized.

13. The electronic device of claim 12 , wherein the controller is configured to:

select a second context from a second dialogue data set including at least one pair of a context and a response corresponding to the context;

generate a response corresponding to the second context through the first dialogue model; and

generate the first dialogue data set by incorporating a pair of the second context and a response corresponding to the second context into the second dialogue data set.

14. The electronic device of claim 13 , wherein the controller is configured to:

generate an augmented response set including a response of the second dialogue data set and a response corresponding to the second context; and

store the augmented response set through the storage device.

15. The electronic device of claim 12 , wherein the loss comprises a cross-entropy loss for a score corresponding to the first response subset and a knowledge distillation loss for a score corresponding to a response included in the response set.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 12, 2022
From: SEO, SEOK JUN; HAN, SEUNG JU; KIM, BEOM SU; CHANG, BU RU; ERDENEE, ENKHBAYAR
To: HYPERCONNECT INC.
Reel/Frame 060799/0559 →
Priority Claims (2)
KR 10-2021-0112541 · Aug 25, 2021 · national
KR 10-2021-0161615 · Nov 22, 2021 · national
Continuity (1)
Related Publication 20230080930A1 · Mar 16, 2023
References Cited (136)
US 6571234B1 · Knight et al. · 2003 [cited by applicant]
US 6731307B1 · Strubbe et al. · 2004 [cited by applicant]
US 6735615B1 · Iwayama et al. · 2004 [cited by applicant]
US 6804647B1 · Heck et al. · 2004 [cited by applicant]
US 6804675B1 · Knight et al. · 2004 [cited by applicant]
US 7277855B1 · Acker et al. · 2007 [cited by applicant]
US 7685237B1 · Weaver et al. · 2010 [cited by applicant]
US 10176819B2 · Sun et al. · 2019 [cited by applicant]
US 10930263B1 · Mahyar · 2021 [cited by applicant]
US 11615777B2 · Ahn et al. · 2023 [cited by applicant]
US 11645547B2 · Tian et al. · 2023 [cited by applicant]
US 20020120450A1 · Junqua et al. · 2002 [cited by applicant]
US 20040111271A1 · Tischer · 2004 [cited by applicant]
US 20050144247A1 · Christensen et al. · 2005 [cited by applicant]
US 20060149558A1 · Kahn et al. · 2006 [cited by applicant]
US 20060210034A1 · Beadle et al. · 2006 [cited by applicant]
US 20060235932A1 · Celi, Jr. et al. · 2006 [cited by applicant]
US 20070005754A1 · Horvitz et al. · 2007 [cited by applicant]
US 20070071206A1 · Gainsboro et al. · 2007 [cited by applicant]
US 20080082333A1 · Nurminen et al. · 2008 [cited by applicant]
US 20080147385A1 · Nurminen et al. · 2008 [cited by applicant]
US 20080183473A1 · Nagano et al. · 2008 [cited by applicant]
US 20080207242A1 · Ekberg · 2008 [cited by applicant]
US 20080235024A1 · Goldberg et al. · 2008 [cited by applicant]
US 20090037179A1 · Liu et al. · 2009 [cited by applicant]
US 20090171657A1 · Tian et al. · 2009 [cited by applicant]
US 20090177473A1 · Aaron et al. · 2009 [cited by applicant]
US 20090204510A1 · Hwang · 2009 [cited by applicant]
US 20100161327A1 · Chandra et al. · 2010 [cited by applicant]
US 20120189272A1 · Kunigita et al. · 2012 [cited by applicant]
US 20120226500A1 · Balasubramanian et al. · 2012 [cited by applicant]
US 20130332167A1 · Kilgore · 2013 [cited by applicant]
US 20140195227A1 · Rudzicz et al. · 2014 [cited by applicant]
US 20140303958A1 · Lee et al. · 2014 [cited by applicant]
US 20150379654A1 · Deshmukh et al. · 2015 [cited by applicant]
US 20160005403A1 · Agiomyrgiannakis et al. · 2016 [cited by applicant]
US 20160036962A1 · Rand · 2016 [cited by applicant]
US 20160098992A1 · Renard · 2016 [cited by examiner]
US 20160104474A1 · Bunn et al. · 2016 [cited by applicant]
US 20160203827A1 · Leff et al. · 2016 [cited by applicant]
US 20160379643A1 · Ito et al. · 2016 [cited by applicant]
US 20170171509A1 · Huang et al. · 2017 [cited by applicant]
US 20170171599A1 · Peng · 2017 [cited by applicant]
US 20170249953A1 · Yassa et al. · 2017 [cited by applicant]
US 20170301340A1 · Yassa et al. · 2017 [cited by applicant]
US 20180048865A1 · Taylor et al. · 2018 [cited by applicant]
US 20180063556A1 · Kalmanson et al. · 2018 [cited by applicant]
US 20180090126A1 · Peterson et al. · 2018 [cited by applicant]
US 20180130471A1 · Trufinescu et al. · 2018 [cited by applicant]
US 20180204576A1 · Dhoot et al. · 2018 [cited by applicant]
US 20180316964A1 · Dillon et al. · 2018 [cited by applicant]
US 20190013017A1 · Kang · 2019 [cited by examiner]
US 20190044985A1 · Jo et al. · 2019 [cited by applicant]
US 20190079941A1 · Sarkar et al. · 2019 [cited by applicant]
US 20190221225A1 · Bricklin et al. · 2019 [cited by applicant]
US 20190251952A1 · Arik et al. · 2019 [cited by applicant]
US 20190334842A1 · Sato · 2019 [cited by applicant]
US 20190354594A1 · Foster et al. · 2019 [cited by applicant]
US 20200013422A1 · Matkin · 2020 [cited by applicant]
US 20200082807A1 · Kim et al. · 2020 [cited by applicant]
US 20200193085A1 · Tanaka · 2020 [cited by examiner]
US 20200197810A1 · Kung et al. · 2020 [cited by applicant]
US 20200265829A1 · Liu et al. · 2020 [cited by applicant]
US 20200395008A1 · Cohen et al. · 2020 [cited by applicant]
US 20210043187A1 · Ahn et al. · 2021 [cited by applicant]
US 20210182489A1 · Barkan · 2021 [cited by examiner]
US 20220199068A1 · Ahn et al. · 2022 [cited by applicant]
US 20230077528A1 · Erdenee et al. · 2023 [cited by applicant]
US 20230154453A1 · Erdenee et al. · 2023 [cited by applicant]
US 20230215418A1 · Ahn et al. · 2023 [cited by applicant]
US 20230229864A1 · Kim et al. · 2023 [cited by applicant]
US 20230229964A1 · Kim et al. · 2023 [cited by applicant]
CN 110633357A · 2019 [cited by examiner]
CN 111708871A · 2020 [cited by examiner]
CN 112541060A · 2021 [cited by applicant]
CN 117411109A · 2024 [cited by examiner]
JP 2003202885A · 2003 [cited by applicant]
JP 2019179257A · 2019 [cited by applicant]
JP 2020160319A · 2020 [cited by applicant]
KR 20000036463A · 2000 [cited by applicant]
KR 20010091677A · 2001 [cited by applicant]
KR 20090028151A · 2009 [cited by applicant]
KR 101632435B1 · 2016 [cited by applicant]
KR 20170107683A · 2017 [cited by applicant]
KR 20180059322A · 2018 [cited by applicant]
KR 20190008137A · 2019 [cited by applicant]
KR 20190085882A · 2019 [cited by applicant]
KR 102170563B1 · 2020 [cited by applicant]
KR 102173553B1 · 2020 [cited by applicant]
WO 2018074516A1 · 2018 [cited by applicant]
WO 2019139430A1 · 2019 [cited by applicant]
WO 2019222591A1 · 2019 [cited by applicant]
Gou, Jianping, et al. “Knowledge distillation: A survey.” International Journal of Computer Vision 129.6 (2021): 1789-1819 (Year: 2021). [cited by examiner]
Zhang, Rongsheng, et al. “Dialogue distillation: Open-domain dialogue augmentation using unpaired data.” arXiv preprint arXiv:2009.09427 (2020). (Year: 2020). [cited by examiner]
Fakoor, Rasool, et al. “Fast, accurate, and simple models for tabular data via augmented distillation.” Advances in Neural Information Processing Systems 33 (2020): 8671-8681. (Year: 2020). [cited by examiner]
Kim, Yoon, and Alexander M. Rush. “Sequence-level knowledge distillation.” arXiv preprint arXiv:1606.07947 (2016). (Year: 2016). [cited by examiner]
Adiwardana, Daniel, et al. “Towards a human-like open-domain chatbot.” arXiv preprint arXiv:2001.09977 (2020). (Year: 2020). [cited by examiner]
European Patent Application 22207004.7 dated Mar. 9, 2023, 9 pgs. [cited by applicant]
Extended European Search Report for Application No. 20189677.6, Dated Sep. 28, 2020, 9 pgs. [cited by applicant]
Extended European Search Report for Application No. 22189981.8, mailed Jan. 17, 2023, 9 pgs. [cited by applicant]
Japanese Office Action for Application No. 2020-134046, Dated Sep. 10, 2021, 8 pgs. [cited by applicant]
Korean Office Action for Application No. 10-2019-0097398, Dated Aug. 18, 2021, 15 pgs. [cited by applicant]
Korean Office Action for Application No. 10-2019-0097398, Dated Jun. 25, 2020, 11 pgs. [cited by applicant]
Office Action for Japanese Patent Application No. 2021-083959 dated Sep. 28, 2022, 2 pgs. [cited by applicant]
Adiwardana et al., “Towards a Human-like Open-Domain Chatbot”, arXiv:2001.09977v3 [cs.CL], Feb. 27, 2020, 38 pgs. [cited by applicant]
Brown et al., “Language Models are Few-Shot Learners”, arXiv:2005.14165v4 [cs.CL], Jul. 22, 2020, 75 pgs. [cited by applicant]
Cai et al., “Retrieval-guided Dialogue Response Generation via a Matching-to-Generation Framework”, Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint … [cited by applicant]
Cai et al., “Skeleton-to-Response: Dialogue Generation Guided by Retrieval Memory”, arXiv:1809.05296v5 [cs.CL], Feb. 28, 2020, 8 pgs. [cited by applicant]
Choi et al., “Attentron: Few-Shot Text-to-Speech Utilizing Attention-Based Variable-Length Embedding”, ArXiv abs/2005.08484, Aug. 12, 2020 (Version 2), 5 pgs. [cited by applicant]
Choi et al., “Attentron: Few-Shot Text-to-Speech Utilizing Attention-Based Variable-Length Embedding”, ArXiv abs/2005.08484, May 18, 2020 (Version 1), 5 pgs. [cited by applicant]
Cooper et al., “Zero-Shot Multi-Speaker Text-to-Speech with State-of-the-Art Neural Speaker Embeddings”, arXiv:1910.10838v2, Feb. 4, 2020, 5 pgs. [cited by applicant]
Fan et al., “Augmenting Transformers with KNN-Based Composite Memory for Dialog”, Transactions of the Association for Computational Linguistics, vol. 9, Mar. 1, 2021, pp. 82-99, https://doi.org/10.1162/tacl_a_00356. [cited by applicant]
Fu et al., “Stylistic Retrieval-based Dialogue System with Unparallel Training Data”, arXiv:2109.05477, Sep. 12, 2021, 9 pgs. [cited by applicant]
Gupta et al., “Controlling Dialogue Generation with Semantic Exemplars”, Proceedings of the 2021 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Ju… [cited by applicant]
Guu et al., “REALM: Retrieval-Augmented Language Model Pre-Training”, arXiv:2002.08909v1 [cs.CL], Feb. 10, 2020, 12 pgs. [cited by applicant]
Han et al., “Meet Your Favorite Character: Open-domain Chatbot Mimicking Fictional Characters with only a Few Utterances”, arXiv:2204.10825, Apr. 22, 2022, 19 pgs. [cited by applicant]
Holtzman et al., “The Curious Case of Neural Text Degeneration”, arXiv:1904.09751v2 [cs.CL], Feb. 14, 2020, 16 pgs. [cited by applicant]
Hsu et al., “Hierarchical Generative Modeling for Controllable Speech Synthesis”, arXiv preprint arXiv:1810.07217v2, Dec. 27, 2018, 27 pgs. [cited by applicant]
Humeau et al., “Poly-Encoders: Architectures and Pre-Training Strategies for Fast and Accurate Multi-Sentence Scoring”, International Conference on Learning Representations, Apr. 30, 2020, 14 pgs. [cited by applicant]
Kim et al., “Distilling the Knowledge of Large-scale Generative Models into Retrieval Models for Efficient Open-domain Conversation”, Findings of the Association for Computational Linguistics, EMNLP 2021, Nov. 7-11, 202… [cited by applicant]
Kim et al., “Sequence-Level Knowledge Distillation”, Proceedings of the 2016 Conference on Empirical Methods in Natural Language Processing, Austin, TX, Nov. 1-5, 2016, pp. 1317-1327. [cited by applicant]
Lee et al., “Robust and Fine-Grained Prosody Control of End-to-End Speech Synthesis”, arXiv:1811.02122v2, Feb. 18, 2019, 5 pgs. [cited by applicant]
Lewis et al., “Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks”, arXiv:2005.11401v4 [cs.CL], Apr. 12, 2021, 19 pgs. [cited by applicant]
Li et al., “A Diversity-Promoting Objective Function for Neural Conversation Models”, Proceedings of NAACL-HLT 2016, San Diego, California, Jun. 12-17, 2016, pp. 110-119. [cited by applicant]
Li et al., “Don't Say That! Making Inconsistent Dialogue Unlikely with Unlikelihood Training”, arXiv:1911.03860v2 [cs.CL], May 6, 2020, 15 pgs. [cited by applicant]
Liu et al., “How Not To Evaluate Your Dialogue System: An Empirical Study of Unsupervised Evaluation Metrics for Dialogue Response Generation”, Proceedings of the 2016 Conference on Empirical Methods in Natural Language… [cited by applicant]
Mazare et al., “Training Millions of Personalized Dialogue Agents”, arXiv:1809.01984v1 [cs.CL], Sep. 6, 2018, 5 pgs. [cited by applicant]
Papineni et al., “BLEU: a Method for Automatic Evaluation of Machine Translation”, Proceedings of the 40th Annual Meeting of the Association for Computational Linguistics (ACL), Philadelphia, Jul. 2002, pp. 311-318. [cited by applicant]
Roller et al., “Recipes for Building an Open-Domain Chatbot”, arXiv:2004.13637v2 [cs.CL], Apr. 30, 2020, 25 pgs. [cited by applicant]
Serban et al., “Multiresolution Recurrent Neural Networks: An Application to Dialogue Response Generation”, arXiv:1606.00776v2 [cs.CL], Jun. 14, 2016, 21 pgs. [cited by applicant]
Welleck et al., “Neural Text Degeneration with Unlikelihood Training”, arXiv:1908.04319v2 [cs.LG], Sep. 26, 2019, 17 pgs. [cited by applicant]
Weston et al., “Retrieve and Refine: Improved Sequence Generation Models For Dialogue”, arXiv:1808.04776v2 [cs.CL], Sep. 6, 2018, 6 pgs. [cited by applicant]
Wu et al., “Response Generation by Context-aware Prototype Editing”, arXiv:1806.07042v4 [cs.CL], Nov. 16, 2018, 9 pgs. [cited by applicant]
Yang et al., “A Hybrid Retrieval-Generation Neural Conversation Model”, arXiv:1904.09068v1 [cs.IR], Apr. 19, 2019, 11 pgs. [cited by applicant]
Zhang et al., “DIALOGPT: Large-Scale Generative Pre-training for Conversational Response Generation”, Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics, Jul. 5-10, 2020, pp. 270-278. [cited by applicant]
Zhang et al., “Dialogue Distillation: Open-Domain Dialogue Augmentation Using Unpaired Data”, Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing, Nov. 16-20, 2020, pp. 3449-3460. [cited by applicant]