Share your thoughts, 1 month free Claude Pro on usSee more
WorkDL logo mark

Accelerating Discrete Diffusion Models with Parallel-In-Time Sampling

About

Discrete diffusion models are widely used for learning and generating discrete distributions. As the generation process is inherently sequential, the acceleration of sampling is of significant importance. In this work, we parallelize the mainstream $\tau$-leaping algorithm for absorbing discrete diffusion in a Continuous-Time Markov Chain (CTMC) framework. By leveraging the continuous-time stochastic integral form of the $\tau$-leaping algorithm and the Picard iteration method, we achieve parallel-in-time sampling acceleration and provide a proof of exponential-factorial convergence for our algorithm. We improve the overall time complexity of $\tau$-leaping under absorbing settings from ${\mathcal{O}}(d \log S)$ to ${\mathcal{O}}(\log (d\log S)\cdot \log d)$ with respect to NFE. Empirically, our method shows consistent acceleration across synthetic and real-data settings. The new sampler achieves at most $7$--$9\times$ runtime speedup for synthetic distribution, and maintains the same quality with $50\%$ fewer NFE and $1.45$--$1.86\times$ runtime speedups in image/text tasks on a single GPU. Our research expands the potential of discrete diffusion models for efficient parallel inference, with broader implications for applications such as molecular structure and language generation.

Yu Yao, Huanjian Zhou, Andi Han, Wei Huang, Masashi Sugiyama• 2026

Related benchmarks

TaskDatasetResultRank
Image GenerationImageNet
FID6.82
106
Text GenerationOpenWebText
Generative Perplexity22.099
30
Text GenerationGPT-2 Large
Perplexity25.876
5
Discrete Diffusion SamplingDiscrete Diffusion Theoretical bounds
Time Complexity (Big O)2
4
Showing 4 of 4 rows

Other info

Follow for update