High-fidelity instructional fashion image editing

Yinglin Zheng,Ting Zhang,Jianmin Bao,Dong Chen,Ming Zeng

doi:10.1016/j.gmod.2024.101223

Abstract

Instructional image editing has received a significant surge of attention recently. In this work, we are interested in the challenging problem of instructional image editing within the particular fashion realm, a domain with significant potential demand in both commercial and personal contexts. This specific domain presents heightened challenges owing to the stringent quality requirements. It necessitates not only the creation of vivid details in alignment with instructions, but also the preservation of precise attributes unrelated to the text guidance. Naive extensions of existing image editing methods produce noticeable artifacts. In order to achieve high-fidelity fashion editing, we propose a novel framework, leveraging the generative prior of a pre-trained human generator and performing edit in the latent space. In addition, we introduce a novel CLIP-based loss to better align the generated target with the instruction. Extensive experiments demonstrate that our approach outperforms prior works including GAN-based editing as well as diffusion-based editing by a large margin, showing impressive visual quality.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

High-fidelity instructional fashion image editing

Abstract

Talk to us

Similar Papers

More From: Graphical Models

Lead the way for us

Similar Papers

GarTemFormer: Temporal transformer-based for optimizing virtual garment animation
Jiazhe Miao ... Li Li
Graphical Models | VOL. 136
Jiazhe Miao, et. al.Jiazhe Miao ... Li Li
11 Oct 2024
Graphical Models | VOL. 136

Building semantic segmentation from large-scale point clouds via primitive recognition
Chiara Romanengo ... Michela Mortara
Graphical Models | VOL. 136
Chiara Romanengo, et. al.Chiara Romanengo ... Michela Mortara
10 Oct 2024
Graphical Models | VOL. 136

Deep-learning-based point cloud completion methods: A review
Kun Zhang ... Weisong Li
Graphical Models | VOL. 136
Kun Zhang, et. al.Kun Zhang ... Weisong Li
03 Oct 2024
Graphical Models | VOL. 136

Sketch-2-4D: Sketch driven dynamic 3D scene generation
Guo-Wei Yang ... Tai-Jiang Mu
Graphical Models | VOL. 136
Guo-Wei Yang, et. al.Guo-Wei Yang ... Tai-Jiang Mu
16 Sep 2024
Graphical Models | VOL. 136

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

High-fidelity instructional fashion image editing

Abstract

Talk to us

Similar Papers

More From: Graphical Models