IP Library Granted Patent US 12,705,427
Granted Patent B2
US 12,705,427 · App. 18/387,663 · Granted Aug 11, 2026

Providing diverse visual contents based on prompts

Inventors: Yair Adato (Kfar Ben Nun, IL); Michael Feinstein (Tel Aviv, IL); Efrat Taig (Beer Sheva, IL); Dvir Yerushalmi (Kfar Saba, IL); Ori Liberman (Netanya, IL)
Assignee: BRIA ARTIFICIAL INTELLIGENCE LTD.
G06F40/30G06F40/279G06F40/40G06T7/194G06T7/70G06T11/10G06V10/764G06V10/774G10L15/063G10L15/18H04N5/272G06T2207/30196G10L2015/0631
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,705,427
App. No.
18/387,663
Filed
Nov 7, 2023
Granted
Aug 11, 2026
Kind
B2
Art Unit
2612
USPC
345/593
Abstract

Systems, methods and non-transitory computer readable media for providing diverse visual contents based on prompts are provided. A textual input in a natural language indicative of a desire of an individual to receive at least one visual content of an inanimate object of a particular category may be received. Further, a demographic requirement may be obtained. For example, the textual input may be analyzed to determine a demographic requirement. Further, a visual content may be obtained based on the demographic requirement and the textual input. The visual content may include a depiction of at least one inanimate object of the particular category and a depiction of one or more persons matching the demographic requirement. Further, a presentation of the visual content to the individual may be caused.

Claims (71)

1 . A non-transitory computer readable medium storing a software program comprising data and computer implementable instructions that when executed by at least one processor cause the at least one processor to perform operations for providing diverse visual contents based on prompts, the operations comprising:

receiving a textual input in a natural language indicative of a desire of an individual to receive at least one visual content of an inanimate object of a particular category;

obtaining a demographic requirement;

obtaining a visual content based on the demographic requirement and the textual input, the visual content includes a depiction of at least one inanimate object of the particular category and a depiction of one or more persons matching the demographic requirement; and

causing a presentation of the visual content to the individual,

wherein the textual input includes a noun and an adjective adjacent to the noun, wherein the determination of the demographic requirement is based on the adjective, and wherein the particular category is based on the noun.

2 . The non-transitory computer readable medium of claim 1 , wherein the operations further comprise using a machine learning model to analyze the textual input to determine the demographic requirement.

3 . The non-transitory computer readable medium of claim 1 , wherein the operations further comprise:

analyzing the textual input to identify a first mathematical object in a mathematical space, wherein the first mathematical object corresponds to a first word of the textual input;

analyzing the textual input to identify a second mathematical object in a mathematical space, wherein the second mathematical object corresponds to a second word of the textual input;

calculating a function of the first mathematical object and the second mathematical object to identify a third mathematical object in the mathematical space; and

determining the demographic requirement based on the third mathematical object.

4 . The non-transitory computer readable medium of claim 1 , wherein the textual input is indicative of a geographical region, and the determination of the demographic requirement is based on the geographical region.

5 . The non-transitory computer readable medium of claim 1 , wherein the textual input is indicative of a specific category different from the particular category, wherein the visual content further includes a depiction of at least one inanimate object of the specific category, and wherein the determination of the demographic requirement is based on the specific category.

6 . The non-transitory computer readable medium of claim 1 , wherein the textual input includes a verb and an adverb adjacent to the verb, wherein the determination of the demographic requirement is based on the adverb.

7 . The non-transitory computer readable medium of claim 1 , wherein the operations further comprise using an inference model to generate the visual content based on the demographic requirement and the textual input.

8 . The non-transitory computer readable medium of claim 1 , wherein the operations further comprise selecting the visual content of a plurality of alternative visual contents based on the demographic requirement and the textual input.

9 . The non-transitory computer readable medium of claim 1 , wherein the operations further comprise:

identifying a first mathematical object in a mathematical space, wherein the first mathematical object corresponds to a word of the textual input;

identifying a second mathematical object in the mathematical space, wherein the second mathematical object corresponds to the demographic requirement;

calculating a function of the first mathematical object and the second mathematical object to identify a third mathematical object in the mathematical space; and

using the third mathematical object to generate the visual content.

10 . The non-transitory computer readable medium of claim 1 , wherein the visual content includes a depiction of a specific person matching the demographic requirement performing an action associated with a specific inanimate object of the particular category, and wherein the action is selected based on the demographic requirement.

11 . The non-transitory computer readable medium of claim 1 , wherein the demographic requirement comply with a diversity requirement.

12 . The non-transitory computer readable medium of claim 1 , wherein a characteristic of the at least one inanimate object of the particular category in the visual content is selected based on the demographic requirement.

13 . The non-transitory computer readable medium of claim 12 , wherein the characteristic is a quantity of the at least one inanimate object.

14 . The non-transitory computer readable medium of claim 1 , wherein the demographic requirement is indicative of an age group, and a size of the at least one inanimate object in the visual content is selected based on the age group.

15 . The non-transitory computer readable medium of claim 1 , wherein the at least one inanimate object includes an amulet, and a color of the amulet in the visual content is selected based on the demographic requirement.

16 . The non-transitory computer readable medium of claim 1 , wherein a spatial relation between the at least one inanimate object and the one or more persons in the visual content is selected based on the demographic requirement.

17 . The non-transitory computer readable medium of claim 1 , wherein the visual content includes a depiction of a first person matching the demographic requirement performing a first action associated with an inanimate object of the particular category, and a depiction of a second person matching the demographic requirement performing a second action associated with the inanimate object of the particular category, and wherein a temporal relation between the first action and the second action in the visual content is selected based on the demographic requirement.

18 . A system for providing diverse visual contents based on prompts, the system comprising:

at least one processor configured to perform the operations of:

receiving a textual input in a natural language indicative of a desire of an individual to receive at least one visual content of an inanimate object of a particular category;

obtaining a demographic requirement;

obtaining a visual content based on the demographic requirement and the textual input, the visual content includes a depiction of at least one inanimate object of the particular category and a depiction of one or more persons matching the demographic requirement; and

causing a presentation of the visual content to the individual,

wherein the textual input includes a noun and an adjective adjacent to the noun, wherein the determination of the demographic requirement is based on the adjective, and wherein the particular category is based on the noun.

19 . A non-transitory computer readable medium storing a software program comprising data and computer implementable instructions that when executed by at least one processor cause the at least one processor to perform operations for providing diverse visual contents based on prompts, the operations comprising:

receiving a textual input in a natural language indicative of a desire of an individual to receive at least one visual content of an inanimate object of a particular category;

obtaining a demographic requirement;

obtaining a visual content based on the demographic requirement and the textual input, the visual content includes a depiction of at least one inanimate object of the particular category and a depiction of one or more persons matching the demographic requirement; and

causing a presentation of the visual content to the individual,

wherein the operations further comprise:

analyzing the textual input to identify a first mathematical object in a mathematical space, wherein the first mathematical object corresponds to a first word of the textual input;

analyzing the textual input to identify a second mathematical object in a mathematical space, wherein the second mathematical object corresponds to a second word of the textual input;

calculating a function of the first mathematical object and the second mathematical object to identify a third mathematical object in the mathematical space; and

determining the demographic requirement based on the third mathematical object.

20 . A non-transitory computer readable medium storing a software program comprising data and computer implementable instructions that when executed by at least one processor cause the at least one processor to perform operations for providing diverse visual contents based on prompts, the operations comprising:

receiving a textual input in a natural language indicative of a desire of an individual to receive at least one visual content of an inanimate object of a particular category;

obtaining a demographic requirement;

obtaining a visual content based on the demographic requirement and the textual input, the visual content includes a depiction of at least one inanimate object of the particular category and a depiction of one or more persons matching the demographic requirement; and

causing a presentation of the visual content to the individual,

wherein the textual input is indicative of a specific category different from the particular category, wherein the visual content further includes a depiction of at least one inanimate object of the specific category, and wherein the determination of the demographic requirement is based on the specific category.

21 . A non-transitory computer readable medium storing a software program comprising data and computer implementable instructions that when executed by at least one processor cause the at least one processor to perform operations for providing diverse visual contents based on prompts, the operations comprising:

receiving a textual input in a natural language indicative of a desire of an individual to receive at least one visual content of an inanimate object of a particular category;

obtaining a demographic requirement;

obtaining a visual content based on the demographic requirement and the textual input, the visual content includes a depiction of at least one inanimate object of the particular category and a depiction of one or more persons matching the demographic requirement; and

causing a presentation of the visual content to the individual,

wherein the textual input includes a verb and an adverb adjacent to the verb, wherein the determination of the demographic requirement is based on the adverb.

22 . A non-transitory computer readable medium storing a software program comprising data and computer implementable instructions that when executed by at least one processor cause the at least one processor to perform operations for providing diverse visual contents based on prompts, the operations comprising:

receiving a textual input in a natural language indicative of a desire of an individual to receive at least one visual content of an inanimate object of a particular category;

obtaining a demographic requirement;

obtaining a visual content based on the demographic requirement and the textual input, the visual content includes a depiction of at least one inanimate object of the particular category and a depiction of one or more persons matching the demographic requirement; and

causing a presentation of the visual content to the individual,

wherein the demographic requirement is indicative of an age group, and a size of the at least one inanimate object in the visual content is selected based on the age group.

23 . A non-transitory computer readable medium storing a software program comprising data and computer implementable instructions that when executed by at least one processor cause the at least one processor to perform operations for providing diverse visual contents based on prompts, the operations comprising:

receiving a textual input in a natural language indicative of a desire of an individual to receive at least one visual content of an inanimate object of a particular category;

obtaining a demographic requirement;

obtaining a visual content based on the demographic requirement and the textual input, the visual content includes a depiction of at least one inanimate object of the particular category and a depiction of one or more persons matching the demographic requirement; and

causing a presentation of the visual content to the individual,

wherein the at least one inanimate object includes an amulet, and a color of the amulet in the visual content is selected based on the demographic requirement.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 30, 2024
From: ADATO, YAIR; FEINSTEIN, MICHAEL; TAIG, EFRAT; YERUSHALMI, DVIR; LIBERMAN, ORI
To: BRIA ARTIFICIAL INTELLIGENCE LTD.
Reel/Frame 067265/0060 →
Continuity (4)
Continuation PCTIL2023051132 · Nov 5, 2023
Provisional Application 63525754 · Jul 10, 2023
Provisional Application 63444805 · Feb 10, 2023
Related Publication 20240273782A1 · Aug 15, 2024
References Cited (48)
US 9519461B2 · Gabel et al. · 2016 [cited by applicant]
US 10083009B2 · Gabel et al. · 2018 [cited by applicant]
US 10091140B2 · Galley et al. · 2018 [cited by applicant]
US 10282672B1 · Mishra et al. · 2019 [cited by applicant]
US 10825227B2 · Amer et al. · 2020 [cited by applicant]
US 11082487B1 · Jain et al. · 2021 [cited by applicant]
US 11551440B1 · Namballa et al. · 2023 [cited by applicant]
US 11880917B2 · Adato et al. · 2024 [cited by applicant]
US 11881958B2 · Rey et al. · 2024 [cited by applicant]
US 11900052B2 · Li et al. · 2024 [cited by applicant]
US 11922317B2 · Takehara · 2024 [cited by applicant]
US 11934792B1 · Adato et al. · 2024 [cited by applicant]
US 11947922B1 · Adato et al. · 2024 [cited by applicant]
US 12073605B1 · Adato et al. · 2024 [cited by applicant]
US 12080277B1 · Adato et al. · 2024 [cited by applicant]
US 20130195361A1 · Deng et al. · 2013 [cited by applicant]
US 20190080205A1 · Kaufhold et al. · 2019 [cited by applicant]
US 20200066025A1 · Peebler et al. · 2020 [cited by applicant]
US 20200265153A1 · Li et al. · 2020 [cited by applicant]
US 20200356591A1 · Yada et al. · 2020 [cited by applicant]
US 20210019541A1 · Wang et al. · 2021 [cited by applicant]
US 20210064825A1 · Sadamasa et al. · 2021 [cited by applicant]
US 20210350604A1 · Pejsa et al. · 2021 [cited by applicant]
US 20220060438A1 · Prasad · 2022 [cited by examiner]
US 20220130078A1 · Maheshwari et al. · 2022 [cited by applicant]
US 20220269980A1 · Keren et al. · 2022 [cited by applicant]
US 20220335203A1 · Van Dyke et al. · 2022 [cited by applicant]
US 20220405314A1 · Du et al. · 2022 [cited by applicant]
US 20230034089A1 · Patel · 2023 [cited by examiner]
US 20230138780A1 · Manamohan et al. · 2023 [cited by applicant]
US 20230306504A1 · Liila et al. · 2023 [cited by applicant]
US 20230351807A1 · Ren et al. · 2023 [cited by applicant]
US 20230360364A1 · Wu et al. · 2023 [cited by applicant]
US 20230409298A1 · Ciminelli et al. · 2023 [cited by applicant]
US 20240104697A1 · Adato et al. · 2024 [cited by applicant]
US 20240112394A1 · Neal · 2024 [cited by examiner]
US 20240144565A1 · Alkalay · 2024 [cited by examiner]
US 20240242428A1 · Ackerman · 2024 [cited by examiner]
US 20240362830A1 · Zhang · 2024 [cited by examiner]
US 20250200114A1 · Parasnis · 2025 [cited by examiner]
CN 111414736A · 2020 [cited by applicant]
Wolfe et al., “American == White in Multimodal Language-and-Image AI,” AIES '22: Proceedings of the 2022 AAAI/ACM Conference on AI, Ethics, and Society (Jul. 2022) hps://doi.org/10.1145/3514094.3534136 ISBN: 97814503924… [cited by examiner]
PCT International Search Report for International Application No. PCT/IL2022/051189, mailed Feb. 9, 2023, 5pp. [cited by applicant]
PCT Written Opinion for International Application No. PCT/IL2022/051189, mailed Feb. 9, 2023, 7pp. [cited by applicant]
PCT International Search Report for International Application No. PCT/IL2023/051132, mailed Mar. 31, 2024, 5pp. [cited by applicant]
PCT Written Opinion for International Application No. PCT/IL2023/051132, mailed Mar. 31, 2024, 5pp. [cited by applicant]
Pang B. Lee L. Opinion mining and sentiment ana lysis. Foundations and Trends® in information retrieva I. Jul. 6, 2008;2(1-2): 1-35. <http://www.cs.cornell.edu/home/llee/omsa/omsa-published.pdf> (Jul. 6, 2008), 94pp. [cited by applicant]
Tzeng E. Hoffman J. Darrell T. Saenko K. Simultaneous deep transfer a cross domains and tasks. InProceedings of the IEEE international conference on computer vision 2015 (pp. 4068-4076). <https:/ /openaccess.thecvf.com/… [cited by applicant]