Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:StylePrompter: All Styles Need Is Attention

Jul 30, 2023

Chenyi Zhuang, Pan Gao, Aljosa Smolic

Figure 1 for StylePrompter: All Styles Need Is Attention

Figure 2 for StylePrompter: All Styles Need Is Attention

Figure 3 for StylePrompter: All Styles Need Is Attention

Figure 4 for StylePrompter: All Styles Need Is Attention

Share this with someone who'll enjoy it:

Abstract:GAN inversion aims at inverting given images into corresponding latent codes for Generative Adversarial Networks (GANs), especially StyleGAN where exists a disentangled latent space that allows attribute-based image manipulation at latent level. As most inversion methods build upon Convolutional Neural Networks (CNNs), we transfer a hierarchical vision Transformer backbone innovatively to predict $\mathcal{W^+}$ latent codes at token level. We further apply a Style-driven Multi-scale Adaptive Refinement Transformer (SMART) in $\mathcal{F}$ space to refine the intermediate style features of the generator. By treating style features as queries to retrieve lost identity information from the encoder's feature maps, SMART can not only produce high-quality inverted images but also surprisingly adapt to editing tasks. We then prove that StylePrompter lies in a more disentangled $\mathcal{W^+}$ and show the controllability of SMART. Finally, quantitative and qualitative experiments demonstrate that StylePrompter can achieve desirable performance in balancing reconstruction quality and editability, and is "smart" enough to fit into most edits, outperforming other $\mathcal{F}$-involved inversion methods.

* Some figures in the appendix are compressed for the reason of arXiv submission constrict

View paper on

Share this with someone who'll enjoy it:

Title:StylePrompter: All Styles Need Is Attention

Paper and Code