IP Library › Granted Patent US 12,566,914
Granted Patent B2
US 12,566,914 · App. 18/240,054 · Granted Mar 3, 2026

System and methods to facilitate content generation using generative artificial intelligence models

Inventors: Ning Xu (Irvine, CA); Jean-Yves Couleaud (Mission Viejo, CA); Cato Yang (San Jose, CA)
Assignee: Adeia Imaging LLC
G06F40/174G06F3/0482G06F3/04847G06F16/24578G06N3/0475
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,566,914
App. No.
18/240,054
Filed
Aug 30, 2023
Granted
Mar 3, 2026
Kind
B2
Art Unit
2179
USPC
715/780
Abstract

The present disclosure is directed to systems and methods to enhance the process of creating an artificial intelligence (AI) generated content or content items, such as images, text, video, sounds, etc., using a text or other suitable prompt, such as via voice input. The systems and methods disclosed provide streamlined content generation with, e.g., reduced processing power and computing time. In an embodiment the systems and methods receive a prompt for generating a first content item using a generative artificial intelligence (AI) model and retrieve, based on the prompt, a collection of matching content items. The systems and methods may then receive input selecting one of the content items from the collection and identify a prompt used to generate the selected content item. The systems and methods may then merge using a trained natural language processing model, the received prompt with the prompt of the selected content item to create a third prompt. In an embodiment the systems and methods may modify the third prompt based on additional input and, based on the modified third prompt, generate a second content item.

Claims (43)

1 . A method comprising:

receiving, via a user interface, a first prompt for generating a first content item using a first generative artificial intelligence (AI) model;

retrieving, based on the first prompt and from a database of stored AI-generated content items, a plurality of content items;

receiving input selecting a content item from the plurality of retrieved content items;

identifying a second prompt used to generate the selected content item;

merging, using a trained natural language processing model, the first prompt with the second prompt to create a third prompt;

modifying the third prompt based on second input received via the user interface; and

generating, using a second generative AI model and based on the modified third prompt, a second content item.

2 . The method of claim 1 , wherein merging, using the trained natural language processing model, the first prompt with the second prompt to create the third prompt further comprises segmenting the first and second prompts, and designating a main description and modifiers.

3 . The method of claim 1 , further comprising suggesting a suggested model based on the first prompt, and wherein generating the second content item is performed using the suggested model.

4 . The method of claim 1 , further comprising suggesting a suggested sampler based on the first prompt, and wherein generating the second content item is performed using the suggested sampler.

5 . The method of claim 1 , further comprising ranking the plurality of content items based on a similarity score.

6 . The method of claim 5 , wherein the similarity score of a respective content item is based on at least a similarity between an embedding vector of the first prompt and an embedding vector of the respective content item, a similarity between the first prompt and prompts used to generate the plurality of content items using their respective embedding vectors, and an image quality of the respective content item.

7 . The method of claim 1 , wherein the trained natural language processing model includes a sentence merging model.

8 . The method of claim 1 , further comprising displaying metadata of the selected content item upon receiving the input selecting the selected content item.

9 . The method of claim 1 , further comprising in response to the merging, determining generation parameters, and wherein the generating the second content item is further based on the generation parameters.

10 . The method of claim 1 , further comprising receiving a negative prompt, and wherein the merging comprises merging the first prompt, the second prompt, and the negative prompt.

11 . The method of claim 1 further comprising:

generating, based at least in part on a text description for the third prompt and one or more generation parameters for the second generative AI model, a fourth prompt, wherein the fourth prompt comprises one or more of:

the one or more generation parameters in-line with the text description and in a format configured for the second generative AI model; and

computer-readable instructions configured to populate one or more parameter fields of the second generative AI model with the one or more generation parameters and the text description.

12 . The method of claim 1 , wherein the first generative AI model and the second generative AI model are the same.

13 . A system comprising:

memory; and

processing circuitry configured to:

receive, via a user interface, a first prompt for generating a first content item using a first generative artificial intelligence (AI) model;

retrieve, based on the first prompt and from a database of stored AI-generated content items stored in memory, a plurality of content items;

receive input selecting a content item from the plurality of retrieved content items;

identify a second prompt used to generate the selected content item;

merge using a trained natural language processing model, the first prompt with the second prompt to create a third prompt;

modify the third prompt based on second input received via the user interface; and

generate, using a second generative AI model and based on the modified third prompt, a second content item.

14 . The system of claim 13 , wherein the processing circuitry is further configured to merge, using the trained natural language processing model, the first prompt with the second prompt to create the third prompt further by segmenting the first and second prompts, and designating a main description and modifiers.

15 . The system of claim 13 , wherein the processing circuitry is further configured to:

suggest a suggested model based on the first prompt, and

generate the second content item by using the suggested model.

16 . The system of claim 13 , wherein the processing circuitry is further configured to:

suggest a suggested sampler based on the first prompt, and

generate the second content item by using the suggested sampler.

17 . The system of claim 13 , wherein the processing circuitry further configured to rank the plurality of content items based on a similarity score.

18 . The system of claim 17 , wherein the similarity score of a respective content item is based on at least a similarity between an embedding vector of the first prompt and an embedding vector of the respective content item, a similarity between the first prompt and prompts used to generate the plurality of content items using their respective embedding vectors, and an image quality of the respective content item.

19 . The system of claim 13 , wherein the trained natural language processing model includes a sentence merging model.

20 . The system of claim 13 , wherein further the processing circuitry is further configured to display metadata of the selected content item upon receiving the input selecting the selected content item.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 9, 2023
From: XU, NING; COULEAUD, JEAN-YVES; YANG, CATO
To: ADEIA IMAGING LLC
Reel/Frame 065512/0520 →
Continuity (1)
Related Publication 20250077765A1 · Mar 6, 2025
References Cited (29)
US 20240193351A1 · Benedetto · 2024 [cited by examiner]
US 20240296276A1 · Hattangady · 2024 [cited by examiner]
US 20240354130A1 · Cadoni · 2024 [cited by examiner]
US 20250022100A1 · Wu · 2025 [cited by examiner]
“DALL-E 2,” OpenAI, https://openai.com/dall-e-2. [cited by applicant]
“DreamStudio,” stability.ai, https://beta.dreamstudio.ai/generate. [cited by applicant]
“Fréchet inception distance,” Wikipedia, https://en.wikipedia.org/wiki/Fr%C3%A9chet_inception_distance. [cited by applicant]
“Imagen: Text-to-Image Diffusion Models,” Google Research, https://imagen.research.google/. [cited by applicant]
“Lexica,” Lexica, https://lexica.art/. [cited by applicant]
“Models, Hugging Face,” https://huggingface.co/models?pipeline_tag=text-to-image&sort=downloads. [cited by applicant]
“NLTK: Natural Language Toolkit,” NLTK, https://www.nltk.org/. [cited by applicant]
“OpenArt,” OpenArt, https://openart.ai/discovery. [cited by applicant]
“Playground,” Playgound, https://playgroundai.com/. [cited by applicant]
“Prompt Generator for Textt2Img,” socialbu, https://socialbu.com/tools/generate-prompt-text2img. [cited by applicant]
“PromptHero,” PromptHero, https://prompthero.com/. [cited by applicant]
Zhang, et. al., “StackGAN: Text to Photo-realistic Image Synthesis with Stacked Generative Adversarial Networks,” arXiv preprint arXiv:1612.03242 (2016). [cited by applicant]
Yalalov, et. al., “Best 100+ Stable Diffusion Prompts: The Most Beautiful AI Text-to-Image Prompts,” Metaverse Post, https://mpost.io/best-100-stable-diffusion-prompts-the-most-beautiful-ai-text-to-image-prompts/. [cited by applicant]
“Sampler v Steps,” Reddit, https://www.reddit.com/r/StableDiffusion/comments/xe26ob/sampler_v_steps/. [cited by applicant]
“SD Steps vs. CFG vs. Sampling Method,” parrot zone, https://proximacentaurib.notion.site/SD-Steps-vs-CFG-vs-Sampling-Method-e8765704d8a6457ca3f66058466fe43a. [cited by applicant]
“Stable Diffusion Launch Announcement,” stability.ai, https://stability.ai/blog/stable-diffusion-announcement. [cited by applicant]
“Stable Diffusion Samplers,” NightCafe, https://nightcafe.studio/blogs/info/stable-diffusion-samplers (2023). [cited by applicant]
Kaddoura, Ahmad, “Stable Diffusion—Samplers,” ArtStation, https://www.artstation.com/blogs/kaddoura/pBPo/stable-diffusion-samplers. [cited by applicant]
RupertAvery “Diffusion Toolkit Beta v0.9,” GitHub, https://github.com/RupertAvery/DiffusionToolkit/releases/tag/beta_v0.9. [cited by applicant]
Karras, et. al., “Elucidating the Design Space of Diffusion-Based Generative Models,” 36th Conference on Neural Information Processing Systems (2022). [cited by applicant]
Radford, et. al., “Learning Transferable Visual Models From Natural Language Supervision,” arXiv preprint arXiv:2103.00020 (2021). [cited by applicant]
Ramesh, et. al., “Hierarchical Text-Conditional Image Generation with CLIP Latents,” arXiv preprint arXiv:2204.06125 (2022). [cited by applicant]
Reed, et al., “Generative Adversarial Text to Image Synthesis,” 33rd ICML (2016). [cited by applicant]
Rombach, Robin , et al., “High-Resolution Image Synthesis with Latent Diffusion Models”, Rombach, et. al., “High-Resolution Image Synthesis with Latent Diffusion Models,” Proceedings of the IEEE/CVF Conference on Comput… [cited by applicant]
Saharia, et. al., “Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding,” arXiv preprint arXiv:2205.11487 (2022). [cited by applicant]