dParallel: Learnable Parallel Decoding for dLLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Zigeng, Fang, Gongfan, Ma, Xinyin, Yu, Ruonan, Wang, Xinchao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DMax: Aggressive Parallel Decoding for dLLMs
von: Chen, Zigeng, et al.
Veröffentlicht: (2026)
von: Chen, Zigeng, et al.
Veröffentlicht: (2026)
dMoE: dLLMs with Learnable Block Experts
von: Feng, Sicheng, et al.
Veröffentlicht: (2026)
von: Feng, Sicheng, et al.
Veröffentlicht: (2026)
dVoting: Fast Voting for dLLMs
von: Feng, Sicheng, et al.
Veröffentlicht: (2026)
von: Feng, Sicheng, et al.
Veröffentlicht: (2026)
Mix-Quant: Quantized Prefilling, Precise Decoding for Agentic LLMs
von: Lu, Haiquan, et al.
Veröffentlicht: (2026)
von: Lu, Haiquan, et al.
Veröffentlicht: (2026)
dKV-Cache: The Cache for Diffusion Language Models
von: Ma, Xinyin, et al.
Veröffentlicht: (2025)
von: Ma, Xinyin, et al.
Veröffentlicht: (2025)
AsyncDiff: Parallelizing Diffusion Models by Asynchronous Denoising
von: Chen, Zigeng, et al.
Veröffentlicht: (2024)
von: Chen, Zigeng, et al.
Veröffentlicht: (2024)
Collaborative Decoding Makes Visual Auto-Regressive Modeling Efficient
von: Chen, Zigeng, et al.
Veröffentlicht: (2024)
von: Chen, Zigeng, et al.
Veröffentlicht: (2024)
Every Step Counts: Decoding Trajectories as Authorship Fingerprints of dLLMs
von: Li, Qi, et al.
Veröffentlicht: (2025)
von: Li, Qi, et al.
Veröffentlicht: (2025)
Thinkless: LLM Learns When to Think
von: Fang, Gongfan, et al.
Veröffentlicht: (2025)
von: Fang, Gongfan, et al.
Veröffentlicht: (2025)
SlimSAM: 0.1% Data Makes Segment Anything Slim
von: Chen, Zigeng, et al.
Veröffentlicht: (2023)
von: Chen, Zigeng, et al.
Veröffentlicht: (2023)
VeriThinker: Learning to Verify Makes Reasoning Model Efficient
von: Chen, Zigeng, et al.
Veröffentlicht: (2025)
von: Chen, Zigeng, et al.
Veröffentlicht: (2025)
Efficient Reasoning Models: A Survey
von: Feng, Sicheng, et al.
Veröffentlicht: (2025)
von: Feng, Sicheng, et al.
Veröffentlicht: (2025)
ConciseHint: Boosting Efficient Reasoning via Continuous Concise Hints during Generation
von: Tang, Siao, et al.
Veröffentlicht: (2025)
von: Tang, Siao, et al.
Veröffentlicht: (2025)
STDec: Spatio-Temporal Stability Guided Decoding for dLLMs
von: Chen, Yuzhe, et al.
Veröffentlicht: (2026)
von: Chen, Yuzhe, et al.
Veröffentlicht: (2026)
SparseD: Sparse Attention for Diffusion Language Models
von: Wang, Zeqing, et al.
Veröffentlicht: (2025)
von: Wang, Zeqing, et al.
Veröffentlicht: (2025)
CoT-Valve: Length-Compressible Chain-of-Thought Tuning
von: Ma, Xinyin, et al.
Veröffentlicht: (2025)
von: Ma, Xinyin, et al.
Veröffentlicht: (2025)
Dimple: Discrete Diffusion Multimodal Large Language Model with Parallel Decoding
von: Yu, Runpeng, et al.
Veröffentlicht: (2025)
von: Yu, Runpeng, et al.
Veröffentlicht: (2025)
MixReasoning: Switching Modes to Think
von: Lu, Haiquan, et al.
Veröffentlicht: (2025)
von: Lu, Haiquan, et al.
Veröffentlicht: (2025)
In-Video Instructions: Visual Signals as Generative Control
von: Fang, Gongfan, et al.
Veröffentlicht: (2025)
von: Fang, Gongfan, et al.
Veröffentlicht: (2025)
Remix-DiT: Mixing Diffusion Transformers for Multi-Expert Denoising
von: Fang, Gongfan, et al.
Veröffentlicht: (2024)
von: Fang, Gongfan, et al.
Veröffentlicht: (2024)
Learning to Parallel: Accelerating Diffusion Large Language Models via Learnable Parallel Decoding
von: Bao, Wenrui, et al.
Veröffentlicht: (2025)
von: Bao, Wenrui, et al.
Veröffentlicht: (2025)
TinyFusion: Diffusion Transformers Learned Shallow
von: Fang, Gongfan, et al.
Veröffentlicht: (2024)
von: Fang, Gongfan, et al.
Veröffentlicht: (2024)
Quantization Meets dLLMs: A Systematic Study of Post-training Quantization for Diffusion LLMs
von: Lin, Haokun, et al.
Veröffentlicht: (2025)
von: Lin, Haokun, et al.
Veröffentlicht: (2025)
Learning-to-Cache: Accelerating Diffusion Transformer via Layer Caching
von: Ma, Xinyin, et al.
Veröffentlicht: (2024)
von: Ma, Xinyin, et al.
Veröffentlicht: (2024)
Isomorphic Pruning for Vision Models
von: Fang, Gongfan, et al.
Veröffentlicht: (2024)
von: Fang, Gongfan, et al.
Veröffentlicht: (2024)
Q-ARVD: Quantizing Autoregressive Video Diffusion Models
von: Tang, Siao, et al.
Veröffentlicht: (2026)
von: Tang, Siao, et al.
Veröffentlicht: (2026)
ASPD: Unlocking Adaptive Serial-Parallel Decoding by Exploring Intrinsic Parallelism in LLMs
von: Chen, Keyu, et al.
Veröffentlicht: (2025)
von: Chen, Keyu, et al.
Veröffentlicht: (2025)
MaskLLM: Learnable Semi-Structured Sparsity for Large Language Models
von: Fang, Gongfan, et al.
Veröffentlicht: (2024)
von: Fang, Gongfan, et al.
Veröffentlicht: (2024)
LoPA: Scaling dLLM Inference via Lookahead Parallel Decoding
von: Xu, Chenkai, et al.
Veröffentlicht: (2025)
von: Xu, Chenkai, et al.
Veröffentlicht: (2025)
PSD: Pushing the Pareto Frontier of Diffusion LLMs via Parallel Speculative Decoding
von: Sun, Shengyin, et al.
Veröffentlicht: (2026)
von: Sun, Shengyin, et al.
Veröffentlicht: (2026)
Invisible Safety Threat: Malicious Finetuning for LLM via Steganography
von: Wan, Guangnian, et al.
Veröffentlicht: (2026)
von: Wan, Guangnian, et al.
Veröffentlicht: (2026)
LiteFocus: Accelerated Diffusion Inference for Long Audio Synthesis
von: Tan, Zhenxiong, et al.
Veröffentlicht: (2024)
von: Tan, Zhenxiong, et al.
Veröffentlicht: (2024)
Decoding at the Speed of Thought: Harnessing Parallel Decoding of Lexical Units for LLMs
von: Sun, Chenxi, et al.
Veröffentlicht: (2024)
von: Sun, Chenxi, et al.
Veröffentlicht: (2024)
AdaDecode: Accelerating LLM Decoding with Adaptive Layer Parallelism
von: Wei, Zhepei, et al.
Veröffentlicht: (2025)
von: Wei, Zhepei, et al.
Veröffentlicht: (2025)
ParallelSpec: Parallel Drafter for Efficient Speculative Decoding
von: Xiao, Zilin, et al.
Veröffentlicht: (2024)
von: Xiao, Zilin, et al.
Veröffentlicht: (2024)
APAR: LLMs Can Do Auto-Parallel Auto-Regressive Decoding
von: Liu, Mingdao, et al.
Veröffentlicht: (2024)
von: Liu, Mingdao, et al.
Veröffentlicht: (2024)
Fast-dLLM: Training-free Acceleration of Diffusion LLM by Enabling KV Cache and Parallel Decoding
von: Wu, Chengyue, et al.
Veröffentlicht: (2025)
von: Wu, Chengyue, et al.
Veröffentlicht: (2025)
Heavy Labels Out! Dataset Distillation with Label Space Lightening
von: Yu, Ruonan, et al.
Veröffentlicht: (2024)
von: Yu, Ruonan, et al.
Veröffentlicht: (2024)
ViFeEdit: A Video-Free Tuner of Your Video Diffusion Transformer
von: Yu, Ruonan, et al.
Veröffentlicht: (2026)
von: Yu, Ruonan, et al.
Veröffentlicht: (2026)
Parallel Decoder Transformer: Planner-Seeded Latent Coordination for Synchronized Parallel Decoding
von: Robbins, Logan
Veröffentlicht: (2025)
von: Robbins, Logan
Veröffentlicht: (2025)
Ähnliche Einträge
-
DMax: Aggressive Parallel Decoding for dLLMs
von: Chen, Zigeng, et al.
Veröffentlicht: (2026) -
dMoE: dLLMs with Learnable Block Experts
von: Feng, Sicheng, et al.
Veröffentlicht: (2026) -
dVoting: Fast Voting for dLLMs
von: Feng, Sicheng, et al.
Veröffentlicht: (2026) -
Mix-Quant: Quantized Prefilling, Precise Decoding for Agentic LLMs
von: Lu, Haiquan, et al.
Veröffentlicht: (2026) -
dKV-Cache: The Cache for Diffusion Language Models
von: Ma, Xinyin, et al.
Veröffentlicht: (2025)