Our new X account is live! Follow @wizwand_team for updates
WorkDL logo mark

GlyphControl: Glyph Conditional Control for Visual Text Generation

About

Recently, there has been an increasing interest in developing diffusion-based text-to-image generative models capable of generating coherent and well-formed visual text. In this paper, we propose a novel and efficient approach called GlyphControl to address this task. Unlike existing methods that rely on character-aware text encoders like ByT5 and require retraining of text-to-image models, our approach leverages additional glyph conditional information to enhance the performance of the off-the-shelf Stable-Diffusion model in generating accurate visual text. By incorporating glyph instructions, users can customize the content, location, and size of the generated text according to their specific requirements. To facilitate further research in visual text generation, we construct a training benchmark dataset called LAION-Glyph. We evaluate the effectiveness of our approach by measuring OCR-based metrics, CLIP score, and FID of the generated visual text. Our empirical evaluations demonstrate that GlyphControl outperforms the recent DeepFloyd IF approach in terms of OCR accuracy, CLIP score, and FID, highlighting the efficacy of our method.

Yukang Yang, Dongnan Gui, Yuhui Yuan, Weicong Liang, Haisong Ding, Han Hu, Kai Chen• 2023

Related benchmarks

TaskDatasetResultRank
Text-to-Image GenerationMARIO-Eval
CLIPScore34.56
25
Visual Text GenerationSimpleBench
Accuracy42
8
Visual Text GenerationCreativeBench
Accuracy28
8
Visual Text GenerationSimpleBench 1.0 (test)
CLIP Score33.9
8
Visual Text GenerationCreativeBench 1.0 (test)
CLIP Score36.2
8
Visual Text GenerationLAION-Glyph 10K samples
FID-10K22.04
8
OCR-based Text RecognitionOpenLibrary (test)
AP47.34
7
OCR-based Text RecognitionTMDB (test)
AP36.35
7
OCR-based Text RecognitionMARIO-7M (test)
AP50.75
7
Text GenerationAnyText benchmark English (test)
Sentence Accuracy52.62
6
Showing 10 of 12 rows

Other info

Code

Follow for update