IP Library › Granted Patent US 12,423,576
Granted Patent B2
US 12,423,576 · App. 17/444,687 · Granted Sep 23, 2025

Method and apparatus for updating parameter of multi-task model, and storage medium

Inventors: Wenhui Zhang (Beijing, CN); Dianhai Yu (Beijing, CN); Zhihua Wu (Beijing, CN)
Assignee: BEIJING BAIDU NETCOM SCIENCE AND TECHNOLOGY CO., LTD.
G06N3/08G06F8/65G06N3/045G06N3/063G06N5/02G06N20/00G06N20/20
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,423,576
App. No.
17/444,687
Granted
Sep 23, 2025
Kind
B2
Abstract

The invention discloses a method and an apparatus for updating parameters of a multi-task model. The method includes: obtaining a training sample set, in which the training sample set comprises a plurality of samples and a task to which each sample belongs; putting each sample into a corresponding sample queue sequentially according to the task to which each sample belongs; training a shared network layer in the multi-task model and a target sub-network layer of tasks associated with the sample queue with samples in the sample queue in case that the number of the samples in the sample queue reaches a training data requirement, so as to generate a model parameter update gradient corresponding to the tasks associated with the sample queue; and updating parameters of the shared network layer and the target sub-network layer in a parameter server according to the model parameter update gradient.

Claims (54)

1. A method for updating parameters of a multi-task model, wherein the multi-task model is an advertisement recommendation model, the advertisement recommendation model comprises two subtasks which are predicting a click-through rate and a conversion rate of the advertisement, the method comprising:

obtaining a training sample set, wherein the training sample set comprises a plurality of samples and a task to which each sample belongs;

putting each sample into a corresponding sample queue sequentially according to a task to which each sample belongs;

training a shared network layer in the multi-task model and a target sub-network layer of a task associated with a sample queue with only samples corresponding to the task in the sample queue;

determining a number of the samples in the sample queue reaches a training data requirement;

wherein in response to the determination the number of the samples in the sample queue reaches the training data requirement, determining a model parameter update gradient corresponding to the task associated with the sample queue; and

updating parameters of the shared network layer and the target sub-network layer in a parameter server according to the model parameter update gradient;

wherein the training is a supervised training scenario, wherein a sample comprises training data and annotation data corresponding to the training data, the training data is feature data of an advertisement comprising the type, the duration, or the label, and the annotation data comprises user operation data comprising a click-through rate, viewing time, a number of favorites, a number of reposts, a number of shares, a conversion rate;

wherein putting each sample into the corresponding sample queue sequentially according to a task to which each sample belongs comprises:

determining a task label corresponding to each sample according to a task to which each sample belongs; and

putting each sample sequentially into a sample queue corresponding to the task label according to the task label corresponding to each sample;

wherein updating the parameters of the shared network layer and the target sub-network layer in the parameter server according to the model parameter update gradient comprises:

determining weights of the task associated with the sample queue;

updating the parameters of the target sub-network layer according to the model parameter update gradient;

determining an update gradient of the shared network layer according to the weights and the model parameter update gradient; and

updating the parameters of the shared network layer according to the update gradient of the shared network layer.

2. The method of claim 1 , wherein putting each sample sequentially into the sample queue corresponding to the task label according to the task label corresponding to each sample comprises:

determining, in the case that a sample corresponds to the plurality of task labels, that the sample queue corresponding to each task label of a plurality of task labels contains the sample.

3. An apparatus for updating parameters of a multi-task model, wherein the multi-task model is an advertisement recommendation model, the advertisement recommendation model comprises two subtasks which are predicting a click-through rate and a conversion rate of the advertisement, the method comprising:

one or more processors;

a memory storing instructions executable by the one or more processors;

wherein the one or more processors are configured to:

obtain a training sample set, wherein the training sample set comprises a plurality of samples and a task to which each sample belongs;

put each sample into a corresponding sample queue sequentially according to a task to which each sample belongs;

train a shared network layer in the multi-task model and a target sub-network layer of a task associated with a sample queue with only samples corresponding to the task in the sample queue;

determine a number of the samples in the sample queue reaches a training data requirement;

wherein in response to the determination the number of the samples in the sample queue reaches the training data requirement, determine a model parameter update gradient corresponding to the task associated with the sample queue; and

update parameters of the shared network layer and the target sub-network layer in a parameter server according to the model parameter update gradient;

wherein the training is a supervised training scenario, wherein a sample comprises training data and annotation data corresponding to the training data, the training data is feature data of an advertisement comprising the type, the duration, or the label, and the annotation data comprises user operation data comprising a click-through rate, viewing time, a number of favorites, a number of reposts, a number of shares, a conversion rate;

wherein the one or more processors are configured to:

determine a task label corresponding to each sample according to a task to which each sample belongs; and

put each sample sequentially into a sample queue corresponding to the task label according to the task label corresponding to each sample;

determine weights of the task associated with the sample queue;

update the parameters of the target sub-network layer according to the model parameter update gradient;

determine an update gradient of the shared network layer according to the weights and the model parameter update gradient; and

update the parameters of the shared network layer according to the update gradient of the shared network layer.

4. The apparatus of claim 3 , wherein the one or more processors are configured to:

determine, in the case that a sample corresponds to the plurality of task labels, that the sample queue corresponding to each task label of a plurality of task labels contains the sample.

5. A non-transitory computer-readable storage medium storing computer instructions, wherein when the computer instructions are executed by a computer, the computer is caused to perform a method for updating parameters of a multi-task model, wherein the multi-task model is an advertisement recommendation model, the advertisement recommendation model comprises two subtasks which are predicting a click-through rate and a conversion rate of the advertisement, and the method comprises:

obtaining a training sample set, wherein the training sample set comprises a plurality of samples and a task to which each sample belongs;

putting each sample into a corresponding sample queue sequentially according to a task to which each sample belongs;

training a shared network layer in the multi-task model and a target sub-network layer of a task associated with a sample queue with only samples corresponding to the task in the sample queue;

determining a number of the samples in the sample queue reaches a training data requirement;

wherein in response to the determination the number of the samples in the sample queue reaches the training data requirement, determining a model parameter update gradient corresponding to the task associated with the sample queue; and

update parameters of the shared network layer and the target sub-network layer in a parameter server according to the model parameter update gradient;

wherein the training is a supervised training scenario, wherein a sample comprises training data and annotation data corresponding to the training data, the training data is feature data of an advertisement comprising the type, the duration, or the label, and the annotation data comprises user operation data comprising a click-through rate, viewing time, a number of favorites, a number of reposts, a number of shares, a conversion rate;

wherein putting each sample into the corresponding sample queue sequentially according to a task to which each sample belongs comprises:

determining a task label corresponding to each sample according to a task to which each sample belongs; and

putting each sample sequentially into a sample queue corresponding to the task label according to the task label corresponding to each sample;

wherein updating the parameters of the shared network layer and the target sub-network layer in the parameter server according to the model parameter update gradient comprises:

determining weights of the task associated with the sample queue;

updating the parameters of the target sub-network layer according to the model parameter update gradient;

determining an update gradient of the shared network layer according to the weights and the model parameter update gradient; and

updating the parameters of the shared network layer according to the update gradient of the shared network layer.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 25, 2022
From: ZHANG, WENHUI; YU, DIANHAI; WU, ZHIHUA
To: BEIJING BAIDU NETCOM SCIENCE AND TECHNOLOGY CO., LTD.
Reel/Frame 060605/0660 →
Priority Claims (1)
CN 202011474877.6 · Dec 14, 2020 · national
Continuity (1)
Related Publication 20210374542A1 · Dec 2, 2021
References Cited (19)
US 20120191631A1 · Breckenridge · 2012 [cited by examiner]
US 20130097011A1 · Wang · 2013 [cited by examiner]
US 20130346182A1 · Cheng · 2013 [cited by examiner]
US 20190130275A1 · Chen et al. · 2019 [cited by applicant]
US 20190286986A1 · Bin et al. · 2019 [cited by applicant]
US 20200104706A1 · Sandler et al. · 2020 [cited by applicant]
US 20200349464A1 · Lin · 2020 [cited by examiner]
US 20220405808A1 · Khoury · 2022 [cited by examiner]
CN 111881968A · 2020 [cited by applicant]
KR 20200078531A · 2020 [cited by applicant]
WO 2018047225A1 · 2018 [cited by applicant]
WO 2019198814A1 · 2019 [cited by applicant]
Zhao, Ying, et al. “Multi-task network anomaly detection using federated learning.” Proceedings of the 10th international symposium on information and communication technology. 2019. (Year: 2019). [cited by examiner]
Wen, Hong, et al. “Entire space multi-task modeling via post-click behavior decomposition for conversion rate prediction.” Proceedings of the 43rd International ACM SIGIR conference on research and development in Inform… [cited by examiner]
Baytas, Inci M., et al. “Asynchronous multi-task learning.” 2016 IEEE 16th International Conference on Data Mining (ICDM). IEEE, 2016. (Year: 2016). [cited by examiner]
Ma, Xiao, et al. “Entire space multi-task model: An effective approach for estimating post-click conversion rate.” The 41st International ACM SIGIR Conference on Research & Development in Information Retrieval. 2018. (Y… [cited by examiner]
European Patent Office Office Action EP21 191 012.0, mailed Jan. 5, 2023. [cited by applicant]
OA for KR application 10-2021-0109539, Aug. 26, 2024, 6 pgs. [cited by applicant]
English translation of OA for KR application 10-2021-0109539, Aug. 26, 2024, 7 pgs. [cited by applicant]