IP Library › Granted Patent US 12,645,744
Granted Patent B2
US 12,645,744 · App. 18/622,051 · Granted Jun 2, 2026

User-specific content generation using text-to-image machine-learned models

Inventor: Arash Sadr (Belmont, CA)
Assignee: GOOGLE LLC
G06F16/9535G06F3/04845G06F16/9538G06T11/00G06T2200/24
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,645,744
App. No.
18/622,051
Granted
Jun 2, 2026
Kind
B2
Abstract

Techniques for presenting a content item using text-to-image machine-learned models are presented. For example, a system can obtain user personalization data associated with a user and merchant assets data of a merchant. Additionally, the system can process the user personalization data and the merchant assets data with a text generation model to generate one or more model-generated terms. Moreover, the system can process the one or more model-generated terms with an image generation model to generate one or more model-generated images. Furthermore, the system can determine a content item based on the one or more model-generated images. Subsequently, the system can present, on a display of a user device of the user, a graphical user interface having the content item.

Claims (62)

1 . A computing system, the system comprising:

one or more processors; and

a memory storing instructions that when executed by the one or more processors cause the system to perform operations comprising:

obtaining user personalization data associated with a user;

obtaining asset data;

processing the user personalization data and the asset data with a text generation model to generate one or more model-generated terms;

processing the one or more model-generated terms with an image generation model to generate one or more model-generated images;

determining a content item based on the one or more model-generated images; and

causing a presentation, on a display of a user device of the user, of a graphical user interface having the content item.

2 . The system of claim 1 , wherein the operations further comprise:

receiving, from the user device, a request to modify a feature of the content item;

transmitting, to a content server, the request to modify the feature of the content item;

receiving, from the content server, an updated content item, the updated content item having a modification to the feature of the content item; and

causing a presentation, on the graphical user interface, of the updated content item.

3 . The system of claim 1 , wherein determining the content item based on the one or more model-generated images comprises:

providing the one or more model-generated images to a content item database; and

receiving the content item from the content item database, wherein the content item is similar to an image from the one or more model-generated images.

4 . The system of claim 1 , wherein determining the content item further comprises:

determining the content item based on the asset data.

5 . The system of claim 1 , wherein the content item includes a link associated with a purchase interface for a product sold by a merchant, wherein the operations further comprise:

receiving, from the user device, a request to purchase the product.

6 . The system of claim 1 , wherein the operations further comprise:

causing a presentation of the one or more model-generated terms in the graphical user interface;

receiving a user input modifying the one or more model-generated terms; and

generating an updated set of terms based on the user input.

7 . The system of claim 1 , wherein the operations further comprise:

causing a presentation of the one or more model-generated images in the graphical user interface;

receiving a user input modifying the one or more model-generated image; and

generating an updated set of images based on the user input.

8 . The system of claim 1 , wherein the one or more model-generated terms include a first term associated with a type of object and a second term associated with a particular descriptive feature, and wherein the one or more model-generated images are descriptive of a particular object of the type of object with the particular descriptive feature.

9 . The system of claim 1 , wherein the asset data includes a product that is sold by a merchant.

10 . The system of claim 9 , wherein the product includes a set of features that are modifiable.

11 . The system of claim 1 , wherein the user personalization data includes explicit personalization data that is received from the user device of the user.

12 . The system of claim 1 , wherein the user personalization data includes implicit personalization data that is derived based on history data of the user and location data of the user.

13 . The system of claim 1 , wherein the operations further comprise:

obtaining, from a search engine, fashion knowledge data; and

wherein the one or more model-generated terms are generated based at least in part on the fashion knowledge data.

14 . The system of claim 1 , wherein the operations further comprise:

obtaining, from a search engine, recent trend data; and

wherein the one or more model-generated terms are generated based at least in part on the recent trend data.

15 . The system of claim 1 , wherein the one or more model-generated terms are generated based at least in part on the user personalization data.

16 . The system of claim 1 , wherein the one or more model-generated terms are generated based at least in part on the asset data.

17 . The system of claim 1 , wherein the one or more model-generated images are generated based at least in part on the one or more model-generated terms.

18 . A computer-implemented method for presenting a content item, the method comprising:

obtaining user personalization data associated with a user;

obtaining asset data;

processing the user personalization data and the asset data with a text generation model to generate one or more model-generated terms;

processing the one or more model-generated terms with an image generation model to generate one or more model-generated images;

determining a content item based on the one or more model-generated images; and

causing a presentation, on a display of a user device of the user, of a graphical user interface having the content item.

19 . The method of claim 18 , the method further comprising:

receiving, from the user device, a request to modify a feature of the content item;

transmitting, to a content server, the request to modify the feature of the content item;

receiving, from the content server, an updated content item, the updated content item having a modification to the feature of the content item; and

causing a presentation, on the graphical user interface, of the updated content item.

20 . One or more non-transitory computer-readable media that collectively store instructions that, when executed by one or more computing devices, cause the one or more computing devices to perform operations, the operations comprising:

obtaining user personalization data associated with a user;

obtaining asset data;

processing the user personalization data and the asset data with a text generation model to generate one or more model-generated terms;

processing the one or more model-generated terms with an image generation model to generate one or more model-generated images;

determining a content item based on the one or more model-generated images; and

causing a presentation, on a display of a user device of the user, of a graphical user interface having the content item.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 2, 2024
From: SADR, ARASH
To: GOOGLE LLC
Reel/Frame 066972/0787 →
Continuity (2)
Provisional Application 63492842 · Mar 29, 2023
Related Publication 20240330381A1 · Oct 3, 2024
References Cited (11)
US 10713821B1 · Surya · 2020 [cited by examiner]
US 11816174B2 · Yu · 2023 [cited by examiner]
US 12456020B1 · Mishra · 2025 [cited by examiner]
US 20200134089A1 · Sankaran et al. · 2020 [cited by applicant]
US 20200356591A1 · Yada · 2020 [cited by examiner]
US 20220398651A1 · Perschk · 2022 [cited by examiner]
US 20250131605A1 · Mekel · 2025 [cited by examiner]
International Preliminary Report on Patentability for Application No. PCT/US2024/022320, mailed Oct. 9, 2025, 7 pages. [cited by applicant]
Deckers et al., “The Infinite Index: Information Retrieval on Generative Text-To-Image Models”, Jan. 21, 2023, arXiv:2212.07476v2, Jan. 21, 2023, 15 pages. [cited by applicant]
International Search Report and Written Opinion for PCT/US2024/022320, mailed on Jul. 4, 2024, 13 pages. [cited by applicant]
Liu et al., “Opal: Multimodal Image Generation for News Illustration”, arXiv:2204.09007v3, Aug. 16, 2022, 17 pages. [cited by applicant]