IP Library Granted Patent US 12,417,344
Granted Patent B2
US 12,417,344 · App. 17/296,821 · Granted Sep 16, 2025

Training recommendation model based on topic model and word importance

Inventors: James Russell Geraci (Suwon-si, KR); Francisco Pena (Dublin, IE); Aonghus Lawlor (Dublin, IE); Barry Smyth (Dublin, IE); Ilias Tragos (Dublin, IE); Neil Hurley (Dublin, IE)
Assignee: Samsung Electronics Co., Ltd.
G06F40/216G06F40/284
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,417,344
App. No.
17/296,821
Granted
Sep 16, 2025
Kind
B2
Abstract

An electronic device and a controlling method thereof are provided. The electronic device includes a memory and a processor configured to obtain importance data representing an importance of each of a plurality of words included in the text data using text data related to a plurality of items written by a plurality of users, obtain a topic model representing a relationship between a topic and a word by applying a topic modelling algorithm to the obtained importance data, obtain preference data representing preference of each of the plurality of users for the topic and relationship data representing a relationship between the plurality of items and the topic, based on the obtained topic model, and train a recommendation model to output result data including an estimated preference for the plurality of items of the plurality of users, based on the preference data and the relationship data.

Claims (49)

1. An electronic device comprising:

a memory; and

a processor configured to:

obtain importance data representing an importance of each of a plurality of words included in text data using text data related to a plurality of items written by a plurality of users,

obtain a topic model representing a relationship between a topic and a word by applying a topic modelling algorithm to the obtained importance data,

obtain preference data representing preference of each of the plurality of users for the topic and relationship data representing a relationship between the plurality of items and the topic, based on the obtained topic model, and

train a recommendation model to output result data including an estimated preference for the plurality of items of the plurality of users, based on the preference data and the relationship data,

wherein the recommendation model comprises a first latent factor and a second latent factor, and

wherein the processor is further configured to:

set a numeral value included in the preference data as an initial value of the first latent factor,

set a numeral value included in the relationship data as an initial value of the second latent factor, and

train the recommendation model so that a difference value between a computation value of the initial value of the first latent factor and the initial value of the second latent factor and a pre-defined learning value is minimized.

2. The electronic device of claim 1 , wherein the processor is further configured to:

identify a noun word among the plurality of words included in the text data,

generate a bag of words including the identified noun word, and

obtain the importance data based on the generated bag of words.

3. The electronic device of claim 2 , wherein the processor is further configured to obtain the importance data by applying a term frequency-inverse document frequency (TF-IDF) algorithm to the generated bag of words.

4. The electronic device of claim 1 , wherein the processor is further configured to:

group the text data by the plurality of users,

obtain first frequency data representing a frequency of words used by each of the plurality of users using the grouped text data, and

obtain the preference data by performing a computation between the first frequency data and the topic model.

5. The electronic device of claim 1 , wherein the processor is further configured to:

group the text data by the plurality of items,

obtain second frequency data corresponding to the item using a frequency of a word included in the grouped text data, and

obtain the relationship data by performing computation between the second frequency data and the topic model.

6. The electronic device of claim 1 , wherein the recommendation model is further configured to, based on the difference value being minimized, perform computation between the first latent factor and the second latent factor and output the result data.

7. A controlling method of an electronic device, the controlling method comprising:

obtaining importance data representing an importance of each of a plurality of words included in text data using text data related to a plurality of items written by a plurality of users;

obtaining a topic model representing a relationship between a topic and a word by applying a topic modelling algorithm to the obtained importance data;

obtaining preference data representing preference of each of the plurality of users for the topic and relationship data representing a relationship between the plurality of items and the topic, based on the obtained topic model; and

training a recommendation model to output result data including an estimated preference for the plurality of items of the plurality of users, based on the preference data and the relationship data,

wherein the recommendation model comprises a first latent factor and a second latent factor, and,

wherein the training comprises:

setting a numeral value included in the preference data as an initial value of the first latent factor;

setting a numeral value included in the relationship data as an initial value of the second latent factor; and

training the recommendation model so that a difference value between a computation value of the initial value of the first latent factor and the initial value of the second latent factor and a pre-defined learning value is minimized.

8. The method of claim 7 , wherein the obtaining the importance data comprises:

identifying a noun word among the plurality of words included in the text data;

generating a bag of words including the identified noun word; and

obtaining the importance data based on the generated bag of words.

9. The method of claim 8 , wherein the obtaining the importance data comprises obtaining the importance data by applying a term frequency-inverse document frequency (TF-IDF) algorithm to the generated bag of words.

10. The method of claim 7 , wherein the obtaining the preference data comprises:

grouping the text data by the plurality of users;

obtaining first frequency data representing a frequency of words used by each of the plurality of users using the grouped text data; and

obtaining the preference data by performing a computation between the first frequency data and the topic model.

11. The method of claim 7 , wherein the obtaining the relationship data comprises:

grouping the text data by the plurality of items;

obtaining second frequency data corresponding to the item using frequency of a word included in the grouped text data; and

obtaining the relationship data by performing computation between the second frequency data and the topic model.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 25, 2021
From: GERACI, JAMES RUSSELL; PENA, FRANCISCO; LAWLOR, AONGHUS; SMYTH, BARRY; TRAGOS, ILIAS; HURLEY, NEIL
To: SAMSUNG ELECTRONICS CO., LTD.
Reel/Frame 056345/0601 →
Priority Claims (2)
KR 10-2020-0065387 · May 29, 2020 · national
KR 10-2021-0041217 · Mar 30, 2021 · national
Continuity (1)
Related Publication 20230259703A1 · Aug 17, 2023
References Cited (33)
US 6981040B1 · Konig · 2005 [cited by examiner]
US 8356044B2 · Stefik · 2013 [cited by examiner]
US 8463662B2 · Koren et al. · 2013 [cited by applicant]
US 9454528B2 · St. Jacques, Jr. et al. · 2016 [cited by applicant]
US 9477777B2 · Stankiewicz · 2016 [cited by examiner]
US 9552555B1 · Yee · 2017 [cited by examiner]
US 10332015B2 · Kawale et al. · 2019 [cited by applicant]
US 10552843B1 · Podgorny et al. · 2020 [cited by applicant]
US 10614365B2 · Sathish et al. · 2020 [cited by applicant]
US 10963941B2 · Garcia Duran et al. · 2021 [cited by applicant]
US 20070162272A1 · Koshinaka · 2007 [cited by examiner]
US 20130325942A1 · Chen · 2013 [cited by examiner]
US 20150262069A1 · Gabriel · 2015 [cited by examiner]
US 20150379610A1 · Stankiewicz et al. · 2015 [cited by applicant]
US 20160034483A1 · Ge · 2016 [cited by examiner]
US 20160041985A1 · Manterach · 2016 [cited by examiner]
US 20170091171A1 · Perez · 2017 [cited by examiner]
US 20170161628A1 · Chiba · 2017 [cited by examiner]
US 20180075137A1 · Lifar · 2018 [cited by examiner]
US 20180158164A1 · Srivastava · 2018 [cited by examiner]
US 20190080383A1 · Garcia Duran et al. · 2019 [cited by applicant]
US 20190180321A1 · Shen · 2019 [cited by examiner]
US 20190197121A1 · Jeon · 2019 [cited by examiner]
US 20200311114A1 · Sood · 2020 [cited by examiner]
US 20200349393A1 · Zhong · 2020 [cited by examiner]
US 20210089884A1 · Macready · 2021 [cited by examiner]
JP 5683758B1 · 2015 [cited by applicant]
JP 2019049980A · 2019 [cited by applicant]
KR 1020130092310A · 2013 [cited by applicant]
KR 1020190103505A · 2019 [cited by applicant]
KR 1020200027089A · 2020 [cited by applicant]
Peña et al., “Combining Rating and Review Data by Initializing Latent Factor Models with Topic Models for Top-N Recommendation ”, Fourteenth ACM Conference on Recommender ACM Conference on Recommender Systems (RecSys '2… [cited by examiner]
International Search Report dated Aug. 30, 2021, issued in International Patent Application No. PCT/ KR2021/095039. [cited by applicant]