DRiffusion: Draft-and-Refine Process Parallelizes Diffusion Models with Ease
Fuente:
arXiv
Saved in:
| Main Authors: | Bai, Runsheng, Zhang, Chengyu, Deng, Yangdong |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Refining Adaptive Zeroth-Order Optimization at Ease
by: Shu, Yao, et al.
Published: (2025)
by: Shu, Yao, et al.
Published: (2025)
Free Draft-and-Verification: Toward Lossless Parallel Decoding for Diffusion Large Language Models
by: Wu, Shutong, et al.
Published: (2025)
by: Wu, Shutong, et al.
Published: (2025)
A Multi-Level Framework for Accelerating Training Transformer Models
by: Zou, Longwei, et al.
Published: (2024)
by: Zou, Longwei, et al.
Published: (2024)
DEER: Draft with Diffusion, Verify with Autoregressive Models
by: Cheng, Zicong, et al.
Published: (2025)
by: Cheng, Zicong, et al.
Published: (2025)
Quantifying the Ease of Reproducing Training Data in Unconditional Diffusion Models
by: Hasegawa, Masaya, et al.
Published: (2025)
by: Hasegawa, Masaya, et al.
Published: (2025)
DREAM-S: Speculative Decoding with Searchable Drafting and Target-Aware Refinement for Multimodal Generation
by: Liu, Zining, et al.
Published: (2026)
by: Liu, Zining, et al.
Published: (2026)
P-EAGLE: Parallel-Drafting EAGLE with Scalable Training
by: Hui, Mude, et al.
Published: (2026)
by: Hui, Mude, et al.
Published: (2026)
AlignedKV: Reducing Memory Access of KV-Cache with Precision-Aligned Quantization
by: Tan, Yifan, et al.
Published: (2024)
by: Tan, Yifan, et al.
Published: (2024)
Exploring and Improving Drafts in Blockwise Parallel Decoding
by: Kim, Taehyeon, et al.
Published: (2024)
by: Kim, Taehyeon, et al.
Published: (2024)
SKIM: Any-bit Quantization Pushing The Limits of Post-Training Quantization
by: Bai, Runsheng, et al.
Published: (2024)
by: Bai, Runsheng, et al.
Published: (2024)
Entropy-MCMC: Sampling from Flat Basins with Ease
by: Li, Bolian, et al.
Published: (2023)
by: Li, Bolian, et al.
Published: (2023)
When Latent Geometry Is Not Enough: Draft-Conditioned Latent Refinement for Non-Autoregressive Text Generation
by: Zhang, De Shuai
Published: (2026)
by: Zhang, De Shuai
Published: (2026)
Self-Refining Diffusion Samplers: Enabling Parallelization via Parareal Iterations
by: Selvam, Nikil Roashan, et al.
Published: (2024)
by: Selvam, Nikil Roashan, et al.
Published: (2024)
PARD: Accelerating LLM Inference with Low-Cost PARallel Draft Model Adaptation
by: An, Zihao, et al.
Published: (2025)
by: An, Zihao, et al.
Published: (2025)
Provably Learning Diffusion Models under the Manifold Hypothesis: Collapse and Refine
by: Huang, Wei, et al.
Published: (2026)
by: Huang, Wei, et al.
Published: (2026)
D-PACE: Dynamic Position-Aware Cross-Entropy for Parallel Speculative Drafting
by: Wu, Tianyu, et al.
Published: (2026)
by: Wu, Tianyu, et al.
Published: (2026)
Refiner: Data Refining against Gradient Leakage Attacks in Federated Learning
by: Fan, Mingyuan, et al.
Published: (2022)
by: Fan, Mingyuan, et al.
Published: (2022)
Learning to Refine: Self-Refinement of Parallel Reasoning in LLMs
by: Wang, Qibin, et al.
Published: (2025)
by: Wang, Qibin, et al.
Published: (2025)
Weak-to-Strong Elicitation via Mismatched Wrong Drafts
by: Deng, Wei
Published: (2026)
by: Deng, Wei
Published: (2026)
Accelerating Parallel Sampling of Diffusion Models
by: Tang, Zhiwei, et al.
Published: (2024)
by: Tang, Zhiwei, et al.
Published: (2024)
Reversible Unfolding Network for Concealed Visual Perception with Generative Refinement
by: He, Chunming, et al.
Published: (2025)
by: He, Chunming, et al.
Published: (2025)
Easing Optimization Paths: a Circuit Perspective
by: Odonnat, Ambroise, et al.
Published: (2025)
by: Odonnat, Ambroise, et al.
Published: (2025)
Parallel Sampling of Diffusion Models on $SO(3)$
by: Chen, Yan-Ting, et al.
Published: (2025)
by: Chen, Yan-Ting, et al.
Published: (2025)
ProPD: Dynamic Token Tree Pruning and Generation for LLM Parallel Decoding
by: Zhong, Shuzhang, et al.
Published: (2024)
by: Zhong, Shuzhang, et al.
Published: (2024)
MineDraft: A Framework for Batch Parallel Speculative Decoding
by: Tang, Zhenwei, et al.
Published: (2026)
by: Tang, Zhenwei, et al.
Published: (2026)
ParallelBench: Understanding the Trade-offs of Parallel Decoding in Diffusion LLMs
by: Kang, Wonjun, et al.
Published: (2025)
by: Kang, Wonjun, et al.
Published: (2025)
EAGLE-2: Faster Inference of Language Models with Dynamic Draft Trees
by: Li, Yuhui, et al.
Published: (2024)
by: Li, Yuhui, et al.
Published: (2024)
ADiff4TPP: Asynchronous Diffusion Models for Temporal Point Processes
by: Mukherjee, Amartya, et al.
Published: (2025)
by: Mukherjee, Amartya, et al.
Published: (2025)
Refining Alignment Framework for Diffusion Models with Intermediate-Step Preference Ranking
by: Ren, Jie, et al.
Published: (2025)
by: Ren, Jie, et al.
Published: (2025)
Generation Order and Parallel Decoding in Masked Diffusion Models: An Information-Theoretic Perspective
by: Zhang, Shaorong, et al.
Published: (2026)
by: Zhang, Shaorong, et al.
Published: (2026)
Towards Cold-Start Drafting and Continual Refining: A Value-Driven Memory Approach with Application to NPU Kernel Synthesis
by: Zheng, Yujie, et al.
Published: (2026)
by: Zheng, Yujie, et al.
Published: (2026)
DRDT3: Diffusion-Refined Decision Test-Time Training Model
by: Huang, Xingshuai, et al.
Published: (2025)
by: Huang, Xingshuai, et al.
Published: (2025)
Inference-Time Scaling of Diffusion Language Models via Trajectory Refinement
by: Dang, Meihua, et al.
Published: (2025)
by: Dang, Meihua, et al.
Published: (2025)
Orchestrating Dual-Boundaries: An Arithmetic Intensity Inspired Acceleration Framework for Diffusion Language Models
by: Wei, Linye, et al.
Published: (2025)
by: Wei, Linye, et al.
Published: (2025)
Parallel Complex Diffusion for Scalable Time Series Generation
by: Cai, Rongyao, et al.
Published: (2026)
by: Cai, Rongyao, et al.
Published: (2026)
Tailed Low-Rank Matrix Factorization for Similarity Matrix Completion
by: Ma, Changyi, et al.
Published: (2024)
by: Ma, Changyi, et al.
Published: (2024)
Review, Remask, Refine (R3): Process-Guided Block Diffusion for Text Generation
by: Mounier, Nikita, et al.
Published: (2025)
by: Mounier, Nikita, et al.
Published: (2025)
Drifting Objectives for Refining Discrete Diffusion Language Models
by: Oba, Daisuke, et al.
Published: (2026)
by: Oba, Daisuke, et al.
Published: (2026)
Efficiently Aligning Draft Models via Parameter- and Data-Efficient Adaptation
by: Lin, Luxi, et al.
Published: (2026)
by: Lin, Luxi, et al.
Published: (2026)
Differentiable Information Bottleneck for Deterministic Multi-view Clustering
by: Yan, Xiaoqiang, et al.
Published: (2024)
by: Yan, Xiaoqiang, et al.
Published: (2024)
Similar Items
-
Refining Adaptive Zeroth-Order Optimization at Ease
by: Shu, Yao, et al.
Published: (2025) -
Free Draft-and-Verification: Toward Lossless Parallel Decoding for Diffusion Large Language Models
by: Wu, Shutong, et al.
Published: (2025) -
A Multi-Level Framework for Accelerating Training Transformer Models
by: Zou, Longwei, et al.
Published: (2024) -
DEER: Draft with Diffusion, Verify with Autoregressive Models
by: Cheng, Zicong, et al.
Published: (2025) -
Quantifying the Ease of Reproducing Training Data in Unconditional Diffusion Models
by: Hasegawa, Masaya, et al.
Published: (2025)