DAPD: Dependency-Aware Parallel Decoding via Attention for Diffusion LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Kim, Bumjun, Jeon, Dongjae, Jeon, Moongyu, No, Albert |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Information-Theoretic Discrete Diffusion
by: Jeon, Moongyu, et al.
Published: (2025)
by: Jeon, Moongyu, et al.
Published: (2025)
Preserve-Then-Quantize: Balancing Rank Budgets for Quantization Error Reconstruction in LLMs
by: Cho, Yoonjun, et al.
Published: (2026)
by: Cho, Yoonjun, et al.
Published: (2026)
Understanding and Mitigating Memorization in Generative Models via Sharpness of Probability Landscapes
by: Jeon, Dongjae, et al.
Published: (2024)
by: Jeon, Dongjae, et al.
Published: (2024)
Rainbow Padding: Mitigating Early Termination in Instruction-Tuned Diffusion LLMs
by: Kim, Bumjun, et al.
Published: (2025)
by: Kim, Bumjun, et al.
Published: (2025)
Memorization In Stable Diffusion Is Unexpectedly Driven by CLIP Embeddings
by: Kim, Bumjun, et al.
Published: (2026)
by: Kim, Bumjun, et al.
Published: (2026)
An Information Theoretic Evaluation Metric For Strong Unlearning
by: Jeon, Dongjae, et al.
Published: (2024)
by: Jeon, Dongjae, et al.
Published: (2024)
Assigning Distinct Roles to Quantized and Low-Rank Matrices Toward Optimal Weight Decomposition
by: Cho, Yoonjun, et al.
Published: (2025)
by: Cho, Yoonjun, et al.
Published: (2025)
ParallelBench: Understanding the Trade-offs of Parallel Decoding in Diffusion LLMs
by: Kang, Wonjun, et al.
Published: (2025)
by: Kang, Wonjun, et al.
Published: (2025)
Two Stage Wireless Federated LoRA Fine-Tuning with Sparsified Orthogonal Updates
by: Kim, Bumjun, et al.
Published: (2025)
by: Kim, Bumjun, et al.
Published: (2025)
Accelerating Diffusion LLMs via Adaptive Parallel Decoding
by: Israel, Daniel, et al.
Published: (2025)
by: Israel, Daniel, et al.
Published: (2025)
TTA-DAME: Test-Time Adaptation with Domain Augmentation and Model Ensemble for Dynamic Driving Conditions
by: Jeon, Dongjae, et al.
Published: (2025)
by: Jeon, Dongjae, et al.
Published: (2025)
A Theoretical Analysis of Why Masked Diffusion Models Mitigate the Reversal Curse
by: Jeon, Moongyu, et al.
Published: (2026)
by: Jeon, Moongyu, et al.
Published: (2026)
Multi-Level Knowledge Distillation and Dynamic Self-Supervised Learning for Continual Learning
by: Kim, Taeheon, et al.
Published: (2025)
by: Kim, Taeheon, et al.
Published: (2025)
Bifurcated Attention: Accelerating Massively Parallel Decoding with Shared Prefixes in LLMs
by: Athiwaratkun, Ben, et al.
Published: (2024)
by: Athiwaratkun, Ben, et al.
Published: (2024)
L4Q: Parameter Efficient Quantization-Aware Fine-Tuning on Large Language Models
by: Jeon, Hyesung, et al.
Published: (2024)
by: Jeon, Hyesung, et al.
Published: (2024)
Efficient Diffusion Models under Nonconvex Equality and Inequality constraints via Landing
by: Jeon, Kijung, et al.
Published: (2026)
by: Jeon, Kijung, et al.
Published: (2026)
ST-LINK: Spatially-Aware Large Language Models for Spatio-Temporal Forecasting
by: Jeon, Hyotaek, et al.
Published: (2025)
by: Jeon, Hyotaek, et al.
Published: (2025)
MOGAM: A Multimodal Object-oriented Graph Attention Model for Depression Detection
by: Cha, Junyeop, et al.
Published: (2024)
by: Cha, Junyeop, et al.
Published: (2024)
AdaEDL: Early Draft Stopping for Speculative Decoding of Large Language Models via an Entropy-based Lower Bound on Token Acceptance Probability
by: Agrawal, Sudhanshu, et al.
Published: (2024)
by: Agrawal, Sudhanshu, et al.
Published: (2024)
SPI-GAN: Denoising Diffusion GANs with Straight-Path Interpolations
by: Jeon, Jinsung, et al.
Published: (2022)
by: Jeon, Jinsung, et al.
Published: (2022)
An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning
by: Jang, Wonseo, et al.
Published: (2025)
by: Jang, Wonseo, et al.
Published: (2025)
Whisfusion: Parallel ASR Decoding via a Diffusion Transformer
by: Kwon, Taeyoun, et al.
Published: (2025)
by: Kwon, Taeyoun, et al.
Published: (2025)
The Active Discoverer Framework: Towards Autonomous Physics Reasoning through Neuro-Symbolic LaTeX Synthesis
by: Jeon, Hyunjun
Published: (2026)
by: Jeon, Hyunjun
Published: (2026)
Self-Refining Language Model Anonymizers via Adversarial Distillation
by: Kim, Kyuyoung, et al.
Published: (2025)
by: Kim, Kyuyoung, et al.
Published: (2025)
Gradient-free Decoder Inversion in Latent Diffusion Models
by: Hong, Seongmin, et al.
Published: (2024)
by: Hong, Seongmin, et al.
Published: (2024)
Two-Stage Grid Optimization for Group-wise Quantization of LLMs
by: Kim, Junhan, et al.
Published: (2026)
by: Kim, Junhan, et al.
Published: (2026)
Direct Alignment of Draft Model for Speculative Decoding with Chat-Fine-Tuned LLMs
by: Goel, Raghavv, et al.
Published: (2024)
by: Goel, Raghavv, et al.
Published: (2024)
DMax: Aggressive Parallel Decoding for dLLMs
by: Chen, Zigeng, et al.
Published: (2026)
by: Chen, Zigeng, et al.
Published: (2026)
Interpretable Water Level Forecaster with Spatiotemporal Causal Attention Mechanisms
by: Hong, Sungchul, et al.
Published: (2023)
by: Hong, Sungchul, et al.
Published: (2023)
Quantization-Aware Imitation-Learning for Resource-Efficient Robotic Control
by: Park, Seongmin, et al.
Published: (2024)
by: Park, Seongmin, et al.
Published: (2024)
Geometry-Aware Attention Guidance for Diffusion Models via Modern Hopfield Dynamics
by: Kim, Kwanyoung
Published: (2026)
by: Kim, Kwanyoung
Published: (2026)
BoA: Attention-aware Post-training Quantization without Backpropagation
by: Kim, Junhan, et al.
Published: (2024)
by: Kim, Junhan, et al.
Published: (2024)
Crash-Consistent Checkpointing for AI Training on macOS/APFS
by: Jeon, Juha
Published: (2025)
by: Jeon, Juha
Published: (2025)
TurboBoA: Faster and Exact Attention-aware Quantization without Backpropagation
by: Kim, Junhan, et al.
Published: (2026)
by: Kim, Junhan, et al.
Published: (2026)
Recursive Speculative Decoding: Accelerating LLM Inference via Sampling Without Replacement
by: Jeon, Wonseok, et al.
Published: (2024)
by: Jeon, Wonseok, et al.
Published: (2024)
Estimating Subgraph Importance with Structural Prior Domain Knowledge
by: Kim, Changhyun, et al.
Published: (2026)
by: Kim, Changhyun, et al.
Published: (2026)
Physics-Guided Geometric Diffusion for Macro Placement Generation
by: Yoon, Jongho, et al.
Published: (2026)
by: Yoon, Jongho, et al.
Published: (2026)
Can Blindfolded LLMs Still Trade? An Anonymization-First Framework for Portfolio Optimization
by: Jeon, Joohyoung, et al.
Published: (2026)
by: Jeon, Joohyoung, et al.
Published: (2026)
REP: Resource-Efficient Prompting for Rehearsal-Free Continual Learning
by: Jeon, Sungho, et al.
Published: (2024)
by: Jeon, Sungho, et al.
Published: (2024)
Dimensionality Reduction Considered Harmful (Some of the Time)
by: Jeon, Hyeon
Published: (2025)
by: Jeon, Hyeon
Published: (2025)
Similar Items
-
Information-Theoretic Discrete Diffusion
by: Jeon, Moongyu, et al.
Published: (2025) -
Preserve-Then-Quantize: Balancing Rank Budgets for Quantization Error Reconstruction in LLMs
by: Cho, Yoonjun, et al.
Published: (2026) -
Understanding and Mitigating Memorization in Generative Models via Sharpness of Probability Landscapes
by: Jeon, Dongjae, et al.
Published: (2024) -
Rainbow Padding: Mitigating Early Termination in Instruction-Tuned Diffusion LLMs
by: Kim, Bumjun, et al.
Published: (2025) -
Memorization In Stable Diffusion Is Unexpectedly Driven by CLIP Embeddings
by: Kim, Bumjun, et al.
Published: (2026)