Wonwoong Cho scite author profile

Recently, unsupervised exemplar-based image-to-image translation, conditioned on a given exemplar without the paired data, has accomplished substantial advancements.In order to transfer the information from an exemplar to an input image, existing methods often use a normalization technique, e.g., adaptive instance normalization, that controls the channel-wise statistics of an input activation map at a particular layer, such as the mean and the variance. Meanwhile, style transfer approaches similar task to image translation by nature, demonstrated superior performance by using the higher-order statistics such as covariance among channels in representing a style. In detail, it works via whitening (given a zero-mean input feature, transforming its covariance matrix into the identity). followed by coloring (changing the covariance matrix of the whitened feature to those of the style feature). However, applying this approach in image translation is computationally intensive and error-prone due to the expensive time complexity and its non-trivial backpropagation. In response, this paper proposes an end-to-end approach tailored for image translation that efficiently approximates this transformation with our novel regularization methods. We further extend our approach to a group-wise form for memory and time efficiency as well as image quality. Extensive qualitative and quantitative experiments demonstrate that our proposed method is fast, both in training and inference, and highly effective in reflecting the style of an exemplar. Finally, our code is available at https://github.com/ WonwoongCho/GDWCT.

show abstract

Coloring with Words: Guiding Image Colorization Through Text-Based Palette Generation

Bahng

Yoo

Cho

et al. 2018

View full text Add to dashboard Cite

This paper proposes a novel approach to generate multiple color palettes that reflect the semantics of input text and then colorize a given grayscale image according to the generated color palette. In contrast to existing approaches, our model can understand rich text, whether it is a single word, a phrase, or a sentence, and generate multiple possible palettes from it. For this task, we introduce our manually curated dataset called Palette-and-Text (PAT). Our proposed model called Text2Colors consists of two conditional generative adversarial networks: the text-topalette generation networks and the palette-based colorization networks. The former captures the semantics of the text input and produce relevant color palettes. The latter colorizes a grayscale image using the generated color palette. Our evaluation results show that people preferred our generated palettes over ground truth palettes and that our model can effectively reflect the given palette when colorizing an image.

show abstract

StyleUV: Diverse and High-fidelity UV Map Generative Model

Lee¹,

Cho²,

Kim³

et al. 2020

Preprint

View full text Add to dashboard Cite

show abstract

A Comparison of the Effects of Data Imputation Methods on Model Performance

Kim

Cho

Choi

et al. 2019

View full text Add to dashboard Cite

Towards Enhanced Controllability of Diffusion Models

Cho¹,

Ravi²,

Harikumar³

et al. 2023

Preprint

View full text Add to dashboard Cite

Figure 1. The proposed framework enhances editability in diffusion models by conditioning the generation on two latent spaces, i.e., content and style. The latent codes are effectively combined to generate novel images. Our proposed sampling technique and timestep scheduling further improve controllability. (a) Magnitude of style can be controlled to translate semantic information from the style image. (b) The learned style space supports smooth interpolations while (c) PCA on the learned latent space gives disentangled attribute specific manipulation directions. Details are provided in sections D.1 and D.2 in appendix. More results can be found in Fig. 25 in appendix.

show abstract

scite is a Brooklyn-based organization that helps researchers better discover and understand research articles through Smart Citations–citations that display the context of the citation and describe whether the article provides supporting or contrasting evidence. scite is used by students and researchers from around the world and is funded in part by the National Science Foundation and the National Institute on Drug Abuse of the National Institutes of Health.

Contact Info

customersupport@researchsolutions.com

10624 S. Eastern Ave., Ste. A-614

Henderson, NV 89052, USA

This site is protected by reCAPTCHA and the Google Privacy Policy and Terms of Service apply.

Blog Terms and Conditions API Terms Privacy Policy Contact Cookie Preferences Do Not Sell or Share My Personal Information

Made with 💙 for researchers

Part of the Research Solutions Family.

Wonwoong Cho

Image-To-Image Translation via Group-Wise Deep Whitening-And-Coloring Transformation

Coloring with Words: Guiding Image Colorization Through Text-Based Palette Generation

StyleUV: Diverse and High-fidelity UV Map Generative Model

A Comparison of the Effects of Data Imputation Methods on Model Performance

Towards Enhanced Controllability of Diffusion Models

Contact Info

Product

Resources

About