DiffStyler: Controllable Dual Diffusion for Text-Driven Image Stylization.

Nisha Huang,Haibin Huang,Weiming Dong,Fan Tang,Yuxin Zhang,Changsheng Xu,Chongyang Ma

doi:10.1109/tnnls.2023.3342645

Abstract

Despite the impressive results of arbitrary image-guided style transfer methods, text-driven image stylization has recently been proposed for transferring a natural image into a stylized one according to textual descriptions of the target style provided by the user. Unlike the previous image-to-image transfer approaches, text-guided stylization progress provides users with a more precise and intuitive way to express the desired style. However, the huge discrepancy between cross-modal inputs/outputs makes it challenging to conduct text-driven image stylization in a typical feed-forward CNN pipeline. In this article, we present DiffStyler, a dual diffusion processing architecture to control the balance between the content and style of the diffused results. The cross-modal style information can be easily integrated as guidance during the diffusion process step-by-step. Furthermore, we propose a content image-based learnable noise on which the reverse denoising process is based, enabling the stylization results to better preserve the structure information of the content image. We validate the proposed DiffStyler beyond the baseline methods through extensive qualitative and quantitative experiments. The code is available at https://github.com/haha-lisa/Diffstyler.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

DiffStyler: Controllable Dual Diffusion for Text-Driven Image Stylization.

Abstract

Talk to us

Similar Papers

More From: IEEE Transactions on Neural Networks and Learning Systems

Lead the way for us

Journal: IEEE Transactions on Neural Networks and Learning Systems	Publication Date: Jan 1, 2024
Citations: 7

Similar Papers

Convolutional autoencoder-based reconstruction of vascular structures in photoacoustic images
Israr U. Haq ... Yoshinobu Kawahara
-
Israr U. Haq, et. al.Israr U. Haq ... Yoshinobu Kawahara
01 Apr 2020
01 Apr 2020

2DPCANet: Dayside Aurora Classification Based on Deep Learning
Zhonghua Jia ... Bing Han
-
Zhonghua Jia, et. al.Zhonghua Jia ... Bing Han
01 Jan 2015
01 Jan 2015

Gradient domain statistical image-importance model for content-aware image resizing
Chanho Jung
Optical Engineering | VOL. 50
Chanho JungChanho Jung
01 Dec 2011
Optical Engineering | VOL. 50

Representation of structural information in images using fuzzy set theory
I Bloch
-
I BlochI Bloch
11 Oct 1998
11 Oct 1998

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

DiffStyler: Controllable Dual Diffusion for Text-Driven Image Stylization.

Abstract

Talk to us

Similar Papers

More From: IEEE Transactions on Neural Networks and Learning Systems