IP Library › Granted Patent US 10,769,492
Granted Patent B2
US 10,769,492 · App. 16/050,072 · Granted Sep 8, 2020

Unsupervised visual attribute transfer through reconfigurable image translation

Inventors: Taeksoo Kim (Seoul, KR); Byoungjip Kim (Seoul, KR); Jiwon Kim (Seoul, KR); Moonsu Cha (Seoul, KR)
Assignee: SK TELECOM CO., LTD.
G06K9/6234G06N3/0454G06N3/088G06T5/50
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,769,492
App. No.
16/050,072
Granted
Sep 8, 2020
Kind
B2
Abstract

The present disclosure relates to unsupervised visual attribute transfer through reconfigurable image translation. One aspect of the present disclosure provides a system for learning the transfer of visual attributes, including an encoder, converter and generator. The encoder encodes an original source image to generate a plurality of attribute values that specify the original source image, and to encode an original reference image to generate a plurality of attribute values that specify the original reference image. The converter replaces at least one attribute value of an attribute that is target attribute of the attribute values of the original source image with at least one corresponding attribute value of the original reference image, to obtain a plurality of attribute values that specify a target image of interest. The generator generates a target image based on the attribute values of the target image of interest.

Claims (28)

1. A system for learning an attribute transfer that conveys at least one attribute value of a reference image to a source image, the system comprising:

an encoder configured to encode an original source image to generate a plurality of attribute values that specify the original source image, and to encode an original reference image to generate a plurality of attribute values that specify the original reference image;

a converter configured to replace at least one attribute value of a target attribute of the attribute values of the original source image with at least one corresponding attribute value of the original reference image, to obtain a plurality of attribute values that specify a target image of interest; and

a generator configured to generate a target image based on the attribute values of the target image of interest,

wherein during training the system,

the encoder encodes the generated target image to generate a plurality of attribute values that specify the generated target image,

the converter replaces at least one attribute value corresponding to the target attribute among the attribute values of the generated target image with at least one attribute value corresponding to the target attribute of the original source image, to generate a plurality of attribute values for a source image reconstruction,

the generator generates a reconstructed source image based on the attribute values of the source image reconstruction,

the converter replaces at least one attribute value corresponding to the target attribute among the attribute values of the original reference image with at least one attribute value corresponding to the target attribute of the target image, to generate a plurality of attribute values for a reference image reconstruction, and

the generator generates a reconstructed reference image based on the attribute values for the reference image reconstruction,

wherein, during training the system, parameters of the encoder and the generator are updated by using,

a reconstruction loss that represents a difference between the reconstructed source image and the original source image,

a reconstruction loss that represents a difference between the reconstructed reference image and the original reference image, and

a generative adversarial loss of the generated target image.

2. The system of claim 1 , further comprising a discriminator configured to learn a discrimination model for discriminating a generated target image from the original source image.

3. A computer-implemented method of training one or more artificial neural networks to learn a visual attribute transfer that conveys a target attribute of a reference image to a source image, for causing the artificial neural networks to perform the computer-implemented method comprising:

encoding an original source image to generate a plurality of attribute values that specify the original source image, and encoding an original reference image to generate a plurality of attribute values that specify the original reference image;

in order to generate a plurality of attribute values that specify a target image of interest, replacing at least one attribute value of a target attribute from among the attribute values of the original source image with at least one corresponding attribute value of the original reference image;

generating a target image based on the attribute values of the target image of interest;

encoding the generated target image to generate a plurality of attribute values that specify a generated target image;

in order to generate a plurality of attribute values for a source image reconstruction, replacing at least one attribute value corresponding to the target attribute from among attribute values of the generated target image with at least one attribute value corresponding to the target attribute of the original source image;

in order to generate a plurality of attribute values for a reference image reconstruction, replacing at least one attribute value corresponding to the target attribute from among the attribute values of the original reference image with at least one attribute value corresponding to the target attribute of the target image;

generating a reconstructed source image based on the attribute values for the source image reconstruction;

generating a reconstructed reference image based on the attribute values for the reference image reconstruction; and

updating parameters of the artificial neural networks by using:

a reconstruction loss that represents a difference between the reconstructed source image and the original source image,

a reconstruction loss that represents a difference between the reconstructed reference image and the original reference image, and

a generative adversarial loss of the generated target image.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 24, 2018
From: KIM, TAEKSOO; KIM, BYOUNGJIP; KIM, JIWON; CHA, MOONSU
To: SK TELECOM CO., LTD.
Reel/Frame 046954/0504 →
Priority Claims (1)
KR 10-2017-0099141 · Aug 4, 2017 · national
Continuity (1)
Related Publication 20190042882A1 · Feb 7, 2019
Cited By (1)
US 12,456,226