Why Masking Diffusion Works: Condition on the Jump Schedule for Improved Discrete Diffusion
Fuente:
arXiv
Saved in:
| Main Authors: | Amin, Alan N., Gruver, Nate, Wilson, Andrew Gordon |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Large Language Models Are Zero-Shot Time Series Forecasters
by: Gruver, Nate, et al.
Published: (2023)
by: Gruver, Nate, et al.
Published: (2023)
The Lie Derivative for Measuring Learned Equivariance
by: Gruver, Nate, et al.
Published: (2022)
by: Gruver, Nate, et al.
Published: (2022)
A Unification of Discrete, Gaussian, and Simplicial Diffusion
by: Chandra, Nuria Alina, et al.
Published: (2025)
by: Chandra, Nuria Alina, et al.
Published: (2025)
Fine-Tuned Language Models Generate Stable Inorganic Materials as Text
by: Gruver, Nate, et al.
Published: (2024)
by: Gruver, Nate, et al.
Published: (2024)
Bayesian Optimization of Antibodies Informed by a Generative Model of Evolving Sequences
by: Amin, Alan Nawzad, et al.
Published: (2024)
by: Amin, Alan Nawzad, et al.
Published: (2024)
Improved Sampling Schedules for Discrete Diffusion Models
by: Foresti, Alberto, et al.
Published: (2026)
by: Foresti, Alberto, et al.
Published: (2026)
A Diffusion Model to Shrink Proteins While Maintaining Their Function
by: Baron, Ethan, et al.
Published: (2025)
by: Baron, Ethan, et al.
Published: (2025)
The Cosine Schedule is Fisher-Rao-Optimal for Masked Discrete Diffusion Models
by: Zhang, Leo, et al.
Published: (2025)
by: Zhang, Leo, et al.
Published: (2025)
Scalable and Flexible Causal Discovery with an Efficient Test for Adjacency
by: Amin, Alan Nawzad, et al.
Published: (2024)
by: Amin, Alan Nawzad, et al.
Published: (2024)
$\textit{Jump Your Steps}$: Optimizing Sampling Schedule of Discrete Diffusion Models
by: Park, Yong-Hyun, et al.
Published: (2024)
by: Park, Yong-Hyun, et al.
Published: (2024)
Optimal Inference Schedules for Masked Diffusion Models
by: Chen, Sitan, et al.
Published: (2025)
by: Chen, Sitan, et al.
Published: (2025)
Training Flexible Models of Genetic Variant Effects from Functional Annotations using Accelerated Linear Algebra
by: Amin, Alan N., et al.
Published: (2025)
by: Amin, Alan N., et al.
Published: (2025)
Simplified and Generalized Masked Diffusion for Discrete Data
by: Shi, Jiaxin, et al.
Published: (2024)
by: Shi, Jiaxin, et al.
Published: (2024)
Beyond Masked and Unmasked: Discrete Diffusion Models via Partial Masking
by: Chao, Chen-Hao, et al.
Published: (2025)
by: Chao, Chen-Hao, et al.
Published: (2025)
Score-Optimal Diffusion Schedules
by: Williams, Christopher, et al.
Published: (2024)
by: Williams, Christopher, et al.
Published: (2024)
Small Batch Size Training for Language Models: When Vanilla SGD Works, and Why Gradient Accumulation Is Wasteful
by: Marek, Martin, et al.
Published: (2025)
by: Marek, Martin, et al.
Published: (2025)
Discrete Diffusion with Sample-Efficient Estimators for Conditionals
by: Elamvazhuthi, Karthik, et al.
Published: (2026)
by: Elamvazhuthi, Karthik, et al.
Published: (2026)
Self-Speculative Masked Diffusions
by: Campbell, Andrew, et al.
Published: (2025)
by: Campbell, Andrew, et al.
Published: (2025)
Latent Shadows: The Gaussian-Discrete Duality in Masked Diffusion
by: Chen, Guinan, et al.
Published: (2026)
by: Chen, Guinan, et al.
Published: (2026)
What Exactly Does Guidance Do in Masked Discrete Diffusion Models
by: Ye, He, et al.
Published: (2025)
by: Ye, He, et al.
Published: (2025)
Steering Masked Discrete Diffusion Models via Discrete Denoising Posterior Prediction
by: Rector-Brooks, Jarrid, et al.
Published: (2024)
by: Rector-Brooks, Jarrid, et al.
Published: (2024)
Large Language Models Must Be Taught to Know What They Don't Know
by: Kapoor, Sanyam, et al.
Published: (2024)
by: Kapoor, Sanyam, et al.
Published: (2024)
Error Bounds and Optimal Schedules for Masked Diffusions with Factorized Approximations
by: Lavenant, Hugo, et al.
Published: (2025)
by: Lavenant, Hugo, et al.
Published: (2025)
Neural Continuous-Time Markov Chain: Discrete Diffusion via Decoupled Jump Timing and Direction
by: Li, Jingyuan, et al.
Published: (2026)
by: Li, Jingyuan, et al.
Published: (2026)
Distillation of Discrete Diffusion by Exact Conditional Distribution Matching
by: Gao, Yansong, et al.
Published: (2025)
by: Gao, Yansong, et al.
Published: (2025)
Learning Generation Orders for Masked Discrete Diffusion Models via Variational Inference
by: Fox, David, et al.
Published: (2026)
by: Fox, David, et al.
Published: (2026)
Simple Self-Conditioning Adaptation for Masked Diffusion Models
by: Cardei, Michael, et al.
Published: (2026)
by: Cardei, Michael, et al.
Published: (2026)
Discrete Modeling via Boundary Conditional Diffusion Processes
by: Gu, Yuxuan, et al.
Published: (2024)
by: Gu, Yuxuan, et al.
Published: (2024)
WavefrontDiffusion: Dynamic Decoding Schedule for Improved Reasoning
by: Yang, Haojin, et al.
Published: (2025)
by: Yang, Haojin, et al.
Published: (2025)
Improving Discrete Diffusion Models via Structured Preferential Generation
by: Rissanen, Severi, et al.
Published: (2024)
by: Rissanen, Severi, et al.
Published: (2024)
Not All Denoising Steps Are Equal: Model Scheduling for Faster Masked Diffusion Language Models
by: Sedykh, Ivan, et al.
Published: (2026)
by: Sedykh, Ivan, et al.
Published: (2026)
Discrete Copula Diffusion
by: Liu, Anji, et al.
Published: (2024)
by: Liu, Anji, et al.
Published: (2024)
Robust Reinforcement Learning under Diffusion Models for Data with Jumps
by: Jiang, Chenyang, et al.
Published: (2024)
by: Jiang, Chenyang, et al.
Published: (2024)
Deep Learning is Not So Mysterious or Different
by: Wilson, Andrew Gordon
Published: (2025)
by: Wilson, Andrew Gordon
Published: (2025)
Generative Representation Learning on Hyper-relational Knowledge Graphs via Masked Discrete Diffusion
by: Lee, Jaejun, et al.
Published: (2026)
by: Lee, Jaejun, et al.
Published: (2026)
Symbol-Aware Reasoning with Masked Discrete Diffusion for Handwritten Mathematical Expression Recognition
by: Kawakatsu, Takaya, et al.
Published: (2026)
by: Kawakatsu, Takaya, et al.
Published: (2026)
Guided Star-Shaped Masked Diffusion
by: Meshchaninov, Viacheslav, et al.
Published: (2025)
by: Meshchaninov, Viacheslav, et al.
Published: (2025)
Parallel Sampling from Masked Diffusion Models via Conditional Independence Testing
by: Azangulov, Iskander, et al.
Published: (2025)
by: Azangulov, Iskander, et al.
Published: (2025)
Auto-Regressive Masked Diffusion Models
by: Karami, Mahdi, et al.
Published: (2026)
by: Karami, Mahdi, et al.
Published: (2026)
Reinforcement Learning for Jump-Diffusions, with Financial Applications
by: Gao, Xuefeng, et al.
Published: (2024)
by: Gao, Xuefeng, et al.
Published: (2024)
Similar Items
-
Large Language Models Are Zero-Shot Time Series Forecasters
by: Gruver, Nate, et al.
Published: (2023) -
The Lie Derivative for Measuring Learned Equivariance
by: Gruver, Nate, et al.
Published: (2022) -
A Unification of Discrete, Gaussian, and Simplicial Diffusion
by: Chandra, Nuria Alina, et al.
Published: (2025) -
Fine-Tuned Language Models Generate Stable Inorganic Materials as Text
by: Gruver, Nate, et al.
Published: (2024) -
Bayesian Optimization of Antibodies Informed by a Generative Model of Evolving Sequences
by: Amin, Alan Nawzad, et al.
Published: (2024)