Saved in:
| Main Authors: | Chen, Sitan, Cong, Kevin, Li, Jerry |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2511.04647 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Optimal high-precision shadow estimation
by: Chen, Sitan, et al.
Published: (2024)
by: Chen, Sitan, et al.
Published: (2024)
S4S: Solving for a Diffusion Model Solver
by: Frankel, Eric, et al.
Published: (2025)
by: Frankel, Eric, et al.
Published: (2025)
Train for the Worst, Plan for the Best: Understanding Token Ordering in Masked Diffusions
by: Kim, Jaeyeon, et al.
Published: (2025)
by: Kim, Jaeyeon, et al.
Published: (2025)
Stop Training for the Worst: Progressive Unmasking Accelerates Masked Diffusion Training
by: Kim, Jaeyeon, et al.
Published: (2026)
by: Kim, Jaeyeon, et al.
Published: (2026)
An optimal tradeoff between entanglement and copy complexity for state tomography
by: Chen, Sitan, et al.
Published: (2024)
by: Chen, Sitan, et al.
Published: (2024)
The Cosine Schedule is Fisher-Rao-Optimal for Masked Discrete Diffusion Models
by: Zhang, Leo, et al.
Published: (2025)
by: Zhang, Leo, et al.
Published: (2025)
Any-Order Flexible Length Masked Diffusion
by: Kim, Jaeyeon, et al.
Published: (2025)
by: Kim, Jaeyeon, et al.
Published: (2025)
Error Bounds and Optimal Schedules for Masked Diffusions with Factorized Approximations
by: Lavenant, Hugo, et al.
Published: (2025)
by: Lavenant, Hugo, et al.
Published: (2025)
Selective Underfitting in Diffusion Models
by: Song, Kiwhan, et al.
Published: (2025)
by: Song, Kiwhan, et al.
Published: (2025)
Provably learning a multi-head attention layer
by: Chen, Sitan, et al.
Published: (2024)
by: Chen, Sitan, et al.
Published: (2024)
Faster Diffusion Sampling with Randomized Midpoints: Sequential and Parallel
by: Gupta, Shivam, et al.
Published: (2024)
by: Gupta, Shivam, et al.
Published: (2024)
Critical windows: non-asymptotic theory for feature emergence in diffusion models
by: Li, Marvin, et al.
Published: (2024)
by: Li, Marvin, et al.
Published: (2024)
Blink of an eye: a simple theory for feature localization in generative models
by: Li, Marvin, et al.
Published: (2025)
by: Li, Marvin, et al.
Published: (2025)
Score-Optimal Diffusion Schedules
by: Williams, Christopher, et al.
Published: (2024)
by: Williams, Christopher, et al.
Published: (2024)
ReGuidance: A Simple Diffusion Wrapper for Boosting Sample Quality on Hard Inverse Problems
by: Karan, Aayush, et al.
Published: (2025)
by: Karan, Aayush, et al.
Published: (2025)
Why Masking Diffusion Works: Condition on the Jump Schedule for Improved Discrete Diffusion
by: Amin, Alan N., et al.
Published: (2025)
by: Amin, Alan N., et al.
Published: (2025)
KLASS: KL-Guided Fast Inference in Masked Diffusion Models
by: Kim, Seo Hyun, et al.
Published: (2025)
by: Kim, Seo Hyun, et al.
Published: (2025)
Gradient dynamics for low-rank fine-tuning beyond kernels
by: Dayi, Arif Kerem, et al.
Published: (2024)
by: Dayi, Arif Kerem, et al.
Published: (2024)
Noise Schedule Design for Diffusion Models: An Optimal Control Perspective
by: Kong, Seo Taek, et al.
Published: (2026)
by: Kong, Seo Taek, et al.
Published: (2026)
Sublinear iterations can suffice even for DDPMs
by: Zhang, Matthew S., et al.
Published: (2025)
by: Zhang, Matthew S., et al.
Published: (2025)
Throughput-Optimal Scheduling Algorithms for LLM Inference and AI Agents
by: Dai, J. G., et al.
Published: (2025)
by: Dai, J. G., et al.
Published: (2025)
Not All Denoising Steps Are Equal: Model Scheduling for Faster Masked Diffusion Language Models
by: Sedykh, Ivan, et al.
Published: (2026)
by: Sedykh, Ivan, et al.
Published: (2026)
MDPO: Overcoming the Training-Inference Divide of Masked Diffusion Language Models
by: He, Haoyu, et al.
Published: (2025)
by: He, Haoyu, et al.
Published: (2025)
What Exactly Does Guidance Do in Masked Discrete Diffusion Models
by: Ye, He, et al.
Published: (2025)
by: Ye, He, et al.
Published: (2025)
High-accuracy and dimension-free sampling with diffusions
by: Gatmiry, Khashayar, et al.
Published: (2026)
by: Gatmiry, Khashayar, et al.
Published: (2026)
Learning Generation Orders for Masked Discrete Diffusion Models via Variational Inference
by: Fox, David, et al.
Published: (2026)
by: Fox, David, et al.
Published: (2026)
Efficient Pauli channel estimation with logarithmic quantum memory
by: Chen, Sitan, et al.
Published: (2023)
by: Chen, Sitan, et al.
Published: (2023)
Learning general Gaussian mixtures with efficient score matching
by: Chen, Sitan, et al.
Published: (2024)
by: Chen, Sitan, et al.
Published: (2024)
Beyond Masked and Unmasked: Discrete Diffusion Models via Partial Masking
by: Chao, Chen-Hao, et al.
Published: (2025)
by: Chao, Chen-Hao, et al.
Published: (2025)
Adaptivity can help exponentially for shadow tomography
by: Chen, Sitan, et al.
Published: (2024)
by: Chen, Sitan, et al.
Published: (2024)
Optimal Scheduling Algorithms for LLM Inference: Theory and Practice
by: Bari, Agrim, et al.
Published: (2025)
by: Bari, Agrim, et al.
Published: (2025)
Masked Diffusion Models as Energy Minimization
by: Chen, Sitong, et al.
Published: (2025)
by: Chen, Sitong, et al.
Published: (2025)
ILRR: Inference-Time Steering Method for Masked Diffusion Language Models
by: Avrahami, Eden, et al.
Published: (2026)
by: Avrahami, Eden, et al.
Published: (2026)
Predicting quantum channels over general product distributions
by: Chen, Sitan, et al.
Published: (2024)
by: Chen, Sitan, et al.
Published: (2024)
Backdooring Masked Diffusion Language Models
by: Cao, Daniel Yiming, et al.
Published: (2026)
by: Cao, Daniel Yiming, et al.
Published: (2026)
MDNS: Masked Diffusion Neural Sampler via Stochastic Optimal Control
by: Zhu, Yuchen, et al.
Published: (2025)
by: Zhu, Yuchen, et al.
Published: (2025)
Inference-Time Scaling of Discrete Diffusion Models via Importance Weighting and Optimal Proposal Design
by: Ou, Zijing, et al.
Published: (2025)
by: Ou, Zijing, et al.
Published: (2025)
Auto-Regressive Masked Diffusion Models
by: Karami, Mahdi, et al.
Published: (2026)
by: Karami, Mahdi, et al.
Published: (2026)
Mamba Meets Scheduling: Learning to Solve Flexible Job Shop Scheduling with Efficient Sequence Modeling
by: Cao, Zhi, et al.
Published: (2026)
by: Cao, Zhi, et al.
Published: (2026)
Improving Text Style Transfer using Masked Diffusion Language Models with Inference-time Scaling
by: Padole, Tejomay Kishor, et al.
Published: (2025)
by: Padole, Tejomay Kishor, et al.
Published: (2025)
Similar Items
-
Optimal high-precision shadow estimation
by: Chen, Sitan, et al.
Published: (2024) -
S4S: Solving for a Diffusion Model Solver
by: Frankel, Eric, et al.
Published: (2025) -
Train for the Worst, Plan for the Best: Understanding Token Ordering in Masked Diffusions
by: Kim, Jaeyeon, et al.
Published: (2025) -
Stop Training for the Worst: Progressive Unmasking Accelerates Masked Diffusion Training
by: Kim, Jaeyeon, et al.
Published: (2026) -
An optimal tradeoff between entanglement and copy complexity for state tomography
by: Chen, Sitan, et al.
Published: (2024)