DRiffusion: Draft-and-Refine Process Parallelizes Diffusion Models with Ease
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bai, Runsheng, Zhang, Chengyu, Deng, Yangdong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Refining Adaptive Zeroth-Order Optimization at Ease
von: Shu, Yao, et al.
Veröffentlicht: (2025)
von: Shu, Yao, et al.
Veröffentlicht: (2025)
Free Draft-and-Verification: Toward Lossless Parallel Decoding for Diffusion Large Language Models
von: Wu, Shutong, et al.
Veröffentlicht: (2025)
von: Wu, Shutong, et al.
Veröffentlicht: (2025)
A Multi-Level Framework for Accelerating Training Transformer Models
von: Zou, Longwei, et al.
Veröffentlicht: (2024)
von: Zou, Longwei, et al.
Veröffentlicht: (2024)
DEER: Draft with Diffusion, Verify with Autoregressive Models
von: Cheng, Zicong, et al.
Veröffentlicht: (2025)
von: Cheng, Zicong, et al.
Veröffentlicht: (2025)
Quantifying the Ease of Reproducing Training Data in Unconditional Diffusion Models
von: Hasegawa, Masaya, et al.
Veröffentlicht: (2025)
von: Hasegawa, Masaya, et al.
Veröffentlicht: (2025)
DREAM-S: Speculative Decoding with Searchable Drafting and Target-Aware Refinement for Multimodal Generation
von: Liu, Zining, et al.
Veröffentlicht: (2026)
von: Liu, Zining, et al.
Veröffentlicht: (2026)
P-EAGLE: Parallel-Drafting EAGLE with Scalable Training
von: Hui, Mude, et al.
Veröffentlicht: (2026)
von: Hui, Mude, et al.
Veröffentlicht: (2026)
AlignedKV: Reducing Memory Access of KV-Cache with Precision-Aligned Quantization
von: Tan, Yifan, et al.
Veröffentlicht: (2024)
von: Tan, Yifan, et al.
Veröffentlicht: (2024)
Exploring and Improving Drafts in Blockwise Parallel Decoding
von: Kim, Taehyeon, et al.
Veröffentlicht: (2024)
von: Kim, Taehyeon, et al.
Veröffentlicht: (2024)
SKIM: Any-bit Quantization Pushing The Limits of Post-Training Quantization
von: Bai, Runsheng, et al.
Veröffentlicht: (2024)
von: Bai, Runsheng, et al.
Veröffentlicht: (2024)
Entropy-MCMC: Sampling from Flat Basins with Ease
von: Li, Bolian, et al.
Veröffentlicht: (2023)
von: Li, Bolian, et al.
Veröffentlicht: (2023)
When Latent Geometry Is Not Enough: Draft-Conditioned Latent Refinement for Non-Autoregressive Text Generation
von: Zhang, De Shuai
Veröffentlicht: (2026)
von: Zhang, De Shuai
Veröffentlicht: (2026)
Self-Refining Diffusion Samplers: Enabling Parallelization via Parareal Iterations
von: Selvam, Nikil Roashan, et al.
Veröffentlicht: (2024)
von: Selvam, Nikil Roashan, et al.
Veröffentlicht: (2024)
PARD: Accelerating LLM Inference with Low-Cost PARallel Draft Model Adaptation
von: An, Zihao, et al.
Veröffentlicht: (2025)
von: An, Zihao, et al.
Veröffentlicht: (2025)
Provably Learning Diffusion Models under the Manifold Hypothesis: Collapse and Refine
von: Huang, Wei, et al.
Veröffentlicht: (2026)
von: Huang, Wei, et al.
Veröffentlicht: (2026)
D-PACE: Dynamic Position-Aware Cross-Entropy for Parallel Speculative Drafting
von: Wu, Tianyu, et al.
Veröffentlicht: (2026)
von: Wu, Tianyu, et al.
Veröffentlicht: (2026)
Refiner: Data Refining against Gradient Leakage Attacks in Federated Learning
von: Fan, Mingyuan, et al.
Veröffentlicht: (2022)
von: Fan, Mingyuan, et al.
Veröffentlicht: (2022)
Learning to Refine: Self-Refinement of Parallel Reasoning in LLMs
von: Wang, Qibin, et al.
Veröffentlicht: (2025)
von: Wang, Qibin, et al.
Veröffentlicht: (2025)
Weak-to-Strong Elicitation via Mismatched Wrong Drafts
von: Deng, Wei
Veröffentlicht: (2026)
von: Deng, Wei
Veröffentlicht: (2026)
Accelerating Parallel Sampling of Diffusion Models
von: Tang, Zhiwei, et al.
Veröffentlicht: (2024)
von: Tang, Zhiwei, et al.
Veröffentlicht: (2024)
Reversible Unfolding Network for Concealed Visual Perception with Generative Refinement
von: He, Chunming, et al.
Veröffentlicht: (2025)
von: He, Chunming, et al.
Veröffentlicht: (2025)
Easing Optimization Paths: a Circuit Perspective
von: Odonnat, Ambroise, et al.
Veröffentlicht: (2025)
von: Odonnat, Ambroise, et al.
Veröffentlicht: (2025)
Parallel Sampling of Diffusion Models on $SO(3)$
von: Chen, Yan-Ting, et al.
Veröffentlicht: (2025)
von: Chen, Yan-Ting, et al.
Veröffentlicht: (2025)
ProPD: Dynamic Token Tree Pruning and Generation for LLM Parallel Decoding
von: Zhong, Shuzhang, et al.
Veröffentlicht: (2024)
von: Zhong, Shuzhang, et al.
Veröffentlicht: (2024)
MineDraft: A Framework for Batch Parallel Speculative Decoding
von: Tang, Zhenwei, et al.
Veröffentlicht: (2026)
von: Tang, Zhenwei, et al.
Veröffentlicht: (2026)
ParallelBench: Understanding the Trade-offs of Parallel Decoding in Diffusion LLMs
von: Kang, Wonjun, et al.
Veröffentlicht: (2025)
von: Kang, Wonjun, et al.
Veröffentlicht: (2025)
EAGLE-2: Faster Inference of Language Models with Dynamic Draft Trees
von: Li, Yuhui, et al.
Veröffentlicht: (2024)
von: Li, Yuhui, et al.
Veröffentlicht: (2024)
ADiff4TPP: Asynchronous Diffusion Models for Temporal Point Processes
von: Mukherjee, Amartya, et al.
Veröffentlicht: (2025)
von: Mukherjee, Amartya, et al.
Veröffentlicht: (2025)
Refining Alignment Framework for Diffusion Models with Intermediate-Step Preference Ranking
von: Ren, Jie, et al.
Veröffentlicht: (2025)
von: Ren, Jie, et al.
Veröffentlicht: (2025)
Generation Order and Parallel Decoding in Masked Diffusion Models: An Information-Theoretic Perspective
von: Zhang, Shaorong, et al.
Veröffentlicht: (2026)
von: Zhang, Shaorong, et al.
Veröffentlicht: (2026)
Towards Cold-Start Drafting and Continual Refining: A Value-Driven Memory Approach with Application to NPU Kernel Synthesis
von: Zheng, Yujie, et al.
Veröffentlicht: (2026)
von: Zheng, Yujie, et al.
Veröffentlicht: (2026)
DRDT3: Diffusion-Refined Decision Test-Time Training Model
von: Huang, Xingshuai, et al.
Veröffentlicht: (2025)
von: Huang, Xingshuai, et al.
Veröffentlicht: (2025)
Inference-Time Scaling of Diffusion Language Models via Trajectory Refinement
von: Dang, Meihua, et al.
Veröffentlicht: (2025)
von: Dang, Meihua, et al.
Veröffentlicht: (2025)
Orchestrating Dual-Boundaries: An Arithmetic Intensity Inspired Acceleration Framework for Diffusion Language Models
von: Wei, Linye, et al.
Veröffentlicht: (2025)
von: Wei, Linye, et al.
Veröffentlicht: (2025)
Parallel Complex Diffusion for Scalable Time Series Generation
von: Cai, Rongyao, et al.
Veröffentlicht: (2026)
von: Cai, Rongyao, et al.
Veröffentlicht: (2026)
Tailed Low-Rank Matrix Factorization for Similarity Matrix Completion
von: Ma, Changyi, et al.
Veröffentlicht: (2024)
von: Ma, Changyi, et al.
Veröffentlicht: (2024)
Review, Remask, Refine (R3): Process-Guided Block Diffusion for Text Generation
von: Mounier, Nikita, et al.
Veröffentlicht: (2025)
von: Mounier, Nikita, et al.
Veröffentlicht: (2025)
Drifting Objectives for Refining Discrete Diffusion Language Models
von: Oba, Daisuke, et al.
Veröffentlicht: (2026)
von: Oba, Daisuke, et al.
Veröffentlicht: (2026)
Efficiently Aligning Draft Models via Parameter- and Data-Efficient Adaptation
von: Lin, Luxi, et al.
Veröffentlicht: (2026)
von: Lin, Luxi, et al.
Veröffentlicht: (2026)
Differentiable Information Bottleneck for Deterministic Multi-view Clustering
von: Yan, Xiaoqiang, et al.
Veröffentlicht: (2024)
von: Yan, Xiaoqiang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Refining Adaptive Zeroth-Order Optimization at Ease
von: Shu, Yao, et al.
Veröffentlicht: (2025) -
Free Draft-and-Verification: Toward Lossless Parallel Decoding for Diffusion Large Language Models
von: Wu, Shutong, et al.
Veröffentlicht: (2025) -
A Multi-Level Framework for Accelerating Training Transformer Models
von: Zou, Longwei, et al.
Veröffentlicht: (2024) -
DEER: Draft with Diffusion, Verify with Autoregressive Models
von: Cheng, Zicong, et al.
Veröffentlicht: (2025) -
Quantifying the Ease of Reproducing Training Data in Unconditional Diffusion Models
von: Hasegawa, Masaya, et al.
Veröffentlicht: (2025)