Optimal Inference Schedules for Masked Diffusion Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Sitan, Cong, Kevin, Li, Jerry |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Optimal high-precision shadow estimation
von: Chen, Sitan, et al.
Veröffentlicht: (2024)
von: Chen, Sitan, et al.
Veröffentlicht: (2024)
Train for the Worst, Plan for the Best: Understanding Token Ordering in Masked Diffusions
von: Kim, Jaeyeon, et al.
Veröffentlicht: (2025)
von: Kim, Jaeyeon, et al.
Veröffentlicht: (2025)
S4S: Solving for a Diffusion Model Solver
von: Frankel, Eric, et al.
Veröffentlicht: (2025)
von: Frankel, Eric, et al.
Veröffentlicht: (2025)
The Cosine Schedule is Fisher-Rao-Optimal for Masked Discrete Diffusion Models
von: Zhang, Leo, et al.
Veröffentlicht: (2025)
von: Zhang, Leo, et al.
Veröffentlicht: (2025)
Stop Training for the Worst: Progressive Unmasking Accelerates Masked Diffusion Training
von: Kim, Jaeyeon, et al.
Veröffentlicht: (2026)
von: Kim, Jaeyeon, et al.
Veröffentlicht: (2026)
An optimal tradeoff between entanglement and copy complexity for state tomography
von: Chen, Sitan, et al.
Veröffentlicht: (2024)
von: Chen, Sitan, et al.
Veröffentlicht: (2024)
Any-Order Flexible Length Masked Diffusion
von: Kim, Jaeyeon, et al.
Veröffentlicht: (2025)
von: Kim, Jaeyeon, et al.
Veröffentlicht: (2025)
Error Bounds and Optimal Schedules for Masked Diffusions with Factorized Approximations
von: Lavenant, Hugo, et al.
Veröffentlicht: (2025)
von: Lavenant, Hugo, et al.
Veröffentlicht: (2025)
Selective Underfitting in Diffusion Models
von: Song, Kiwhan, et al.
Veröffentlicht: (2025)
von: Song, Kiwhan, et al.
Veröffentlicht: (2025)
Score-Optimal Diffusion Schedules
von: Williams, Christopher, et al.
Veröffentlicht: (2024)
von: Williams, Christopher, et al.
Veröffentlicht: (2024)
Provably learning a multi-head attention layer
von: Chen, Sitan, et al.
Veröffentlicht: (2024)
von: Chen, Sitan, et al.
Veröffentlicht: (2024)
KLASS: KL-Guided Fast Inference in Masked Diffusion Models
von: Kim, Seo Hyun, et al.
Veröffentlicht: (2025)
von: Kim, Seo Hyun, et al.
Veröffentlicht: (2025)
Why Masking Diffusion Works: Condition on the Jump Schedule for Improved Discrete Diffusion
von: Amin, Alan N., et al.
Veröffentlicht: (2025)
von: Amin, Alan N., et al.
Veröffentlicht: (2025)
Critical windows: non-asymptotic theory for feature emergence in diffusion models
von: Li, Marvin, et al.
Veröffentlicht: (2024)
von: Li, Marvin, et al.
Veröffentlicht: (2024)
Blink of an eye: a simple theory for feature localization in generative models
von: Li, Marvin, et al.
Veröffentlicht: (2025)
von: Li, Marvin, et al.
Veröffentlicht: (2025)
Faster Diffusion Sampling with Randomized Midpoints: Sequential and Parallel
von: Gupta, Shivam, et al.
Veröffentlicht: (2024)
von: Gupta, Shivam, et al.
Veröffentlicht: (2024)
Noise Schedule Design for Diffusion Models: An Optimal Control Perspective
von: Kong, Seo Taek, et al.
Veröffentlicht: (2026)
von: Kong, Seo Taek, et al.
Veröffentlicht: (2026)
MDPO: Overcoming the Training-Inference Divide of Masked Diffusion Language Models
von: He, Haoyu, et al.
Veröffentlicht: (2025)
von: He, Haoyu, et al.
Veröffentlicht: (2025)
Throughput-Optimal Scheduling Algorithms for LLM Inference and AI Agents
von: Dai, J. G., et al.
Veröffentlicht: (2025)
von: Dai, J. G., et al.
Veröffentlicht: (2025)
Not All Denoising Steps Are Equal: Model Scheduling for Faster Masked Diffusion Language Models
von: Sedykh, Ivan, et al.
Veröffentlicht: (2026)
von: Sedykh, Ivan, et al.
Veröffentlicht: (2026)
What Exactly Does Guidance Do in Masked Discrete Diffusion Models
von: Ye, He, et al.
Veröffentlicht: (2025)
von: Ye, He, et al.
Veröffentlicht: (2025)
Learning Generation Orders for Masked Discrete Diffusion Models via Variational Inference
von: Fox, David, et al.
Veröffentlicht: (2026)
von: Fox, David, et al.
Veröffentlicht: (2026)
ReGuidance: A Simple Diffusion Wrapper for Boosting Sample Quality on Hard Inverse Problems
von: Karan, Aayush, et al.
Veröffentlicht: (2025)
von: Karan, Aayush, et al.
Veröffentlicht: (2025)
Gradient dynamics for low-rank fine-tuning beyond kernels
von: Dayi, Arif Kerem, et al.
Veröffentlicht: (2024)
von: Dayi, Arif Kerem, et al.
Veröffentlicht: (2024)
Beyond Masked and Unmasked: Discrete Diffusion Models via Partial Masking
von: Chao, Chen-Hao, et al.
Veröffentlicht: (2025)
von: Chao, Chen-Hao, et al.
Veröffentlicht: (2025)
Sublinear iterations can suffice even for DDPMs
von: Zhang, Matthew S., et al.
Veröffentlicht: (2025)
von: Zhang, Matthew S., et al.
Veröffentlicht: (2025)
High-accuracy and dimension-free sampling with diffusions
von: Gatmiry, Khashayar, et al.
Veröffentlicht: (2026)
von: Gatmiry, Khashayar, et al.
Veröffentlicht: (2026)
Optimal Scheduling Algorithms for LLM Inference: Theory and Practice
von: Bari, Agrim, et al.
Veröffentlicht: (2025)
von: Bari, Agrim, et al.
Veröffentlicht: (2025)
Inference-Time Scaling of Discrete Diffusion Models via Importance Weighting and Optimal Proposal Design
von: Ou, Zijing, et al.
Veröffentlicht: (2025)
von: Ou, Zijing, et al.
Veröffentlicht: (2025)
Masked Diffusion Models as Energy Minimization
von: Chen, Sitong, et al.
Veröffentlicht: (2025)
von: Chen, Sitong, et al.
Veröffentlicht: (2025)
ILRR: Inference-Time Steering Method for Masked Diffusion Language Models
von: Avrahami, Eden, et al.
Veröffentlicht: (2026)
von: Avrahami, Eden, et al.
Veröffentlicht: (2026)
Auto-Regressive Masked Diffusion Models
von: Karami, Mahdi, et al.
Veröffentlicht: (2026)
von: Karami, Mahdi, et al.
Veröffentlicht: (2026)
Backdooring Masked Diffusion Language Models
von: Cao, Daniel Yiming, et al.
Veröffentlicht: (2026)
von: Cao, Daniel Yiming, et al.
Veröffentlicht: (2026)
Learning general Gaussian mixtures with efficient score matching
von: Chen, Sitan, et al.
Veröffentlicht: (2024)
von: Chen, Sitan, et al.
Veröffentlicht: (2024)
MDNS: Masked Diffusion Neural Sampler via Stochastic Optimal Control
von: Zhu, Yuchen, et al.
Veröffentlicht: (2025)
von: Zhu, Yuchen, et al.
Veröffentlicht: (2025)
Efficient Pauli channel estimation with logarithmic quantum memory
von: Chen, Sitan, et al.
Veröffentlicht: (2023)
von: Chen, Sitan, et al.
Veröffentlicht: (2023)
Mamba Meets Scheduling: Learning to Solve Flexible Job Shop Scheduling with Efficient Sequence Modeling
von: Cao, Zhi, et al.
Veröffentlicht: (2026)
von: Cao, Zhi, et al.
Veröffentlicht: (2026)
Urban In-Context Learning: Bridging Pretraining and Inference through Masked Diffusion for Urban Profiling
von: Zhang, Ruixing, et al.
Veröffentlicht: (2025)
von: Zhang, Ruixing, et al.
Veröffentlicht: (2025)
Improving Text Style Transfer using Masked Diffusion Language Models with Inference-time Scaling
von: Padole, Tejomay Kishor, et al.
Veröffentlicht: (2025)
von: Padole, Tejomay Kishor, et al.
Veröffentlicht: (2025)
Sharp Convergence Rates for Masked Diffusion Models
von: Liang, Yuchen, et al.
Veröffentlicht: (2026)
von: Liang, Yuchen, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Optimal high-precision shadow estimation
von: Chen, Sitan, et al.
Veröffentlicht: (2024) -
Train for the Worst, Plan for the Best: Understanding Token Ordering in Masked Diffusions
von: Kim, Jaeyeon, et al.
Veröffentlicht: (2025) -
S4S: Solving for a Diffusion Model Solver
von: Frankel, Eric, et al.
Veröffentlicht: (2025) -
The Cosine Schedule is Fisher-Rao-Optimal for Masked Discrete Diffusion Models
von: Zhang, Leo, et al.
Veröffentlicht: (2025) -
Stop Training for the Worst: Progressive Unmasking Accelerates Masked Diffusion Training
von: Kim, Jaeyeon, et al.
Veröffentlicht: (2026)