Our new X account is live! Follow @wizwand_team for updates
WorkDL logo mark

Controlling Your Image via Simplified Vector Graphics

About

Recent advances in image generation have achieved remarkable visual quality, while a fundamental challenge remains: Can image generation be controlled at the element level, enabling intuitive modifications such as adjusting shapes, altering colors, or adding and removing objects? In this work, we address this challenge by introducing layer-wise controllable generation through simplified vector graphics (VGs). Our approach first efficiently parses images into hierarchical VG representations that are semantic-aligned and structurally coherent. Building on this representation, we design a novel image synthesis framework guided by VGs, allowing users to freely modify elements and seamlessly translate these edits into photorealistic outputs. By leveraging the structural and semantic features of VGs in conjunction with noise prediction, our method provides precise control over geometry, color, and object semantics. Extensive experiments demonstrate the effectiveness of our approach in diverse applications, including image editing, object-level manipulation, and fine-grained content creation, establishing a new paradigm for controllable image generation. Project page: https://guolanqing.github.io/Vec2Pix/

Lanqing Guo, Xi Liu, Yufei Wang, Zhihao Li, Siyu Huang• 2026

Related benchmarks

TaskDatasetResultRank
Controllable Image Generation5,000-image Reconstruction (evaluation)
FID15.52
6
Controllable Image GenerationEditing 5,000-image (evaluation)
FID17.84
6
Showing 2 of 2 rows

Other info

Follow for update