Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:GlyphControl: Glyph Conditional Control for Visual Text Generation

May 29, 2023

Yukang Yang, Dongnan Gui, Yuhui Yuan, Haisong Ding, Han Hu, Kai Chen

Figure 1 for GlyphControl: Glyph Conditional Control for Visual Text Generation

Figure 2 for GlyphControl: Glyph Conditional Control for Visual Text Generation

Figure 3 for GlyphControl: Glyph Conditional Control for Visual Text Generation

Figure 4 for GlyphControl: Glyph Conditional Control for Visual Text Generation

Share this with someone who'll enjoy it:

Abstract:Recently, there has been a growing interest in developing diffusion-based text-to-image generative models capable of generating coherent and well-formed visual text. In this paper, we propose a novel and efficient approach called GlyphControl to address this task. Unlike existing methods that rely on character-aware text encoders like ByT5 and require retraining of text-to-image models, our approach leverages additional glyph conditional information to enhance the performance of the off-the-shelf Stable-Diffusion model in generating accurate visual text. By incorporating glyph instructions, users can customize the content, location, and size of the generated text according to their specific requirements. To facilitate further research in visual text generation, we construct a training benchmark dataset called LAION-Glyph. We evaluate the effectiveness of our approach by measuring OCR-based metrics and CLIP scores of the generated visual text. Our empirical evaluations demonstrate that GlyphControl outperforms the recent DeepFloyd IF approach in terms of OCR accuracy and CLIP scores, highlighting the efficacy of our method.

* Technical report. The codes will be released at https://github.com/AIGText/GlyphControl-release

View paper on

Share this with someone who'll enjoy it:

Title:GlyphControl: Glyph Conditional Control for Visual Text Generation

Paper and Code