Accelerating Discrete Diffusion Models with Parallel-In-Time Sampling

arXiv:2607.0077317.4
Predicted impact top 6% in LG · last 90 daysOriginality Incremental advance
AI Analysis

For practitioners using discrete diffusion models, this provides a practical speedup for sampling without quality loss, though the improvement is incremental over existing methods.

This work accelerates discrete diffusion model sampling by parallelizing the τ-leaping algorithm, achieving up to 7-9× runtime speedup on synthetic data and 1.45-1.86× on image/text tasks with 50% fewer NFE.

Discrete diffusion models are widely used for learning and generating discrete distributions. As the generation process is inherently sequential, the acceleration of sampling is of significant importance. In this work, we parallelize the mainstream $τ$-leaping algorithm for absorbing discrete diffusion in a Continuous-Time Markov Chain (CTMC) framework. By leveraging the continuous-time stochastic integral form of the $τ$-leaping algorithm and the Picard iteration method, we achieve parallel-in-time sampling acceleration and provide a proof of exponential-factorial convergence for our algorithm. We improve the overall time complexity of $τ$-leaping under absorbing settings from ${\mathcal{O}}(d \log S)$ to ${\mathcal{O}}(\log (d\log S)\cdot \log d)$ with respect to NFE. Empirically, our method shows consistent acceleration across synthetic and real-data settings. The new sampler achieves at most $7$--$9\times$ runtime speedup for synthetic distribution, and maintains the same quality with $50\%$ fewer NFE and $1.45$--$1.86\times$ runtime speedups in image/text tasks on a single GPU. Our research expands the potential of discrete diffusion models for efficient parallel inference, with broader implications for applications such as molecular structure and language generation.

Foundations

The foundational work for this paper's niche, ranked by how specifically the neighbourhood builds on it — not by global fame.

Your Notes