Training-Free Self-Correction for Multimodal Masked Diffusion Models
Fuente:
arXiv
Saved in:
| Main Authors: | Ouyang, Yidong, Hu, Panwen, Wan, Zhengyan, Wang, Zhe, Xie, Liyan, Bespalov, Dmitriy, Wu, Ying Nian, Cheng, Guang, Zha, Hongyuan, Sun, Qiang |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Corrected Samplers for Discrete Flow Models
by: Wan, Zhengyan, et al.
Published: (2026)
by: Wan, Zhengyan, et al.
Published: (2026)
Transfer Learning for Diffusion Models
by: Ouyang, Yidong, et al.
Published: (2024)
by: Ouyang, Yidong, et al.
Published: (2024)
Alignment of Diffusion Model and Flow Matching for Text-to-Image Generation
by: Ouyang, Yidong, et al.
Published: (2026)
by: Ouyang, Yidong, et al.
Published: (2026)
Discrete Guidance Matching: Exact Guidance for Discrete Flow Matching
by: Wan, Zhengyan, et al.
Published: (2025)
by: Wan, Zhengyan, et al.
Published: (2025)
Error Analysis of Discrete Flow with Generator Matching
by: Wan, Zhengyan, et al.
Published: (2025)
by: Wan, Zhengyan, et al.
Published: (2025)
dFlowGRPO: Rate-Aware Policy Optimization for Discrete Flow Models
by: Wan, Zhengyan, et al.
Published: (2026)
by: Wan, Zhengyan, et al.
Published: (2026)
DirectSwap: Mask-Free Cross-Identity Training and Benchmarking for Expression-Consistent Video Head Swapping
by: Wang, Yanan, et al.
Published: (2025)
by: Wang, Yanan, et al.
Published: (2025)
Masked Diffusion Modeling for Anomaly Detection
by: Zhang, Lixing, et al.
Published: (2026)
by: Zhang, Lixing, et al.
Published: (2026)
Graph of Attacks with Pruning: Optimizing Stealthy Jailbreak Prompt Generation for Enhanced LLM Content Moderation
by: Schwartz, Daniel, et al.
Published: (2025)
by: Schwartz, Daniel, et al.
Published: (2025)
TurboFuzzLLM: Turbocharging Mutation-based Fuzzing for Effectively Jailbreaking Large Language Models in Practice
by: Goel, Aman, et al.
Published: (2025)
by: Goel, Aman, et al.
Published: (2025)
A multi-purpose automatic editing system based on lecture semantics for remote education
by: Hu, Panwen, et al.
Published: (2024)
by: Hu, Panwen, et al.
Published: (2024)
BridgeIV: Bridging Customized Image and Video Generation through Test-Time Autoregressive Identity Propagation
by: Hu, Panwen, et al.
Published: (2025)
by: Hu, Panwen, et al.
Published: (2025)
Fast T2T: Optimization Consistency Speeds Up Diffusion-Based Training-to-Testing Solving for Combinatorial Optimization
by: Li, Yang, et al.
Published: (2025)
by: Li, Yang, et al.
Published: (2025)
Mixture of Noise for Pre-Trained Model-Based Class-Incremental Learning
by: Jiang, Kai, et al.
Published: (2025)
by: Jiang, Kai, et al.
Published: (2025)
Towards Building a Robust Toxicity Predictor
by: Bespalov, Dmitriy, et al.
Published: (2024)
by: Bespalov, Dmitriy, et al.
Published: (2024)
TaeBench: Improving Quality of Toxic Adversarial Examples
by: Zhu, Xuan, et al.
Published: (2024)
by: Zhu, Xuan, et al.
Published: (2024)
Revise, Don't Freeze: Sampler-Matched Training for Self-Correcting Masked Diffusion Language Models
by: Yu, Longxuan, et al.
Published: (2026)
by: Yu, Longxuan, et al.
Published: (2026)
AdaCorrection: Adaptive Offset Cache Correction for Accurate Diffusion Transformers
by: Liu, Dong, et al.
Published: (2026)
by: Liu, Dong, et al.
Published: (2026)
A Variational Autoencoder for Neural Temporal Point Processes with Dynamic Latent Graphs
by: Yang, Sikun, et al.
Published: (2023)
by: Yang, Sikun, et al.
Published: (2023)
From infant separation distress to preschool behavior problems: The mediating roles of maternal self‐efficacy and positive parenting
by: Qiang Wang, et al.
Published: (2025)
by: Qiang Wang, et al.
Published: (2025)
ViewMask-1-to-3: Multi-View Consistent Image Generation via Multimodal Diffusion Models
by: Zhu, Ruishu, et al.
Published: (2025)
by: Zhu, Ruishu, et al.
Published: (2025)
Optimising Language Models for Downstream Tasks: A Post-Training Perspective
by: Shi, Zhengyan
Published: (2025)
by: Shi, Zhengyan
Published: (2025)
FreeVPS: Repurposing Training-Free SAM2 for Generalizable Video Polyp Segmentation
by: Hu, Qiang, et al.
Published: (2025)
by: Hu, Qiang, et al.
Published: (2025)
Generalized Discrete Diffusion with Self-Correction
by: Wang, Linxuan, et al.
Published: (2026)
by: Wang, Linxuan, et al.
Published: (2026)
LaRS: Latent Reasoning Skills for Chain-of-Thought Reasoning
by: Xu, Zifan, et al.
Published: (2023)
by: Xu, Zifan, et al.
Published: (2023)
An adaptive multimesh rational approximation scheme for the spectral fractional Laplacian
by: Bespalov, Alex, et al.
Published: (2025)
by: Bespalov, Alex, et al.
Published: (2025)
Convergence analysis of the adaptive stochastic collocation finite element method
by: Bespalov, Alex, et al.
Published: (2024)
by: Bespalov, Alex, et al.
Published: (2024)
Learn from Your Mistakes: Self-Correcting Masked Diffusion Models
by: Schiff, Yair, et al.
Published: (2026)
by: Schiff, Yair, et al.
Published: (2026)
Mitigating Backdoor Attack by Injecting Proactive Defensive Backdoor
by: Wei, Shaokui, et al.
Published: (2024)
by: Wei, Shaokui, et al.
Published: (2024)
Backdoor Mitigation by Distance-Driven Detoxification
by: Wei, Shaokui, et al.
Published: (2024)
by: Wei, Shaokui, et al.
Published: (2024)
Quantum Corrections to Randall-Sundrum Model from JT Gravity
by: Chen, Ying-Jian, et al.
Published: (2025)
by: Chen, Ying-Jian, et al.
Published: (2025)
Exploring Semantic Masked Autoencoder for Self-supervised Point Cloud Understanding
by: Zha, Yixin, et al.
Published: (2025)
by: Zha, Yixin, et al.
Published: (2025)
On extended 1-perfect bitrades
by: Bespalov, Evgeny A., et al.
Published: (2020)
by: Bespalov, Evgeny A., et al.
Published: (2020)
Self-Speculative Masked Diffusions
by: Campbell, Andrew, et al.
Published: (2025)
by: Campbell, Andrew, et al.
Published: (2025)
LaVieID: Local Autoregressive Diffusion Transformers for Identity-Preserving Video Creation
by: Song, Wenhui, et al.
Published: (2025)
by: Song, Wenhui, et al.
Published: (2025)
Overcoming Dynamics-Blindness: Training-Free Pace-and-Path Correction for VLA Models
by: Zhang, Yanyan, et al.
Published: (2026)
by: Zhang, Yanyan, et al.
Published: (2026)
DGMO: Training-Free Audio Source Separation through Diffusion-Guided Mask Optimization
by: Lee, Geonyoung, et al.
Published: (2025)
by: Lee, Geonyoung, et al.
Published: (2025)
Enhancing Prompt Following with Visual Control Through Training-Free Mask-Guided Diffusion
by: Chen, Hongyu, et al.
Published: (2024)
by: Chen, Hongyu, et al.
Published: (2024)
Lavida-O: Elastic Large Masked Diffusion Models for Unified Multimodal Understanding and Generation
by: Li, Shufan, et al.
Published: (2025)
by: Li, Shufan, et al.
Published: (2025)
Self-Attention Decomposition For Training Free Diffusion Editing
by: Anand, Tharun, et al.
Published: (2025)
by: Anand, Tharun, et al.
Published: (2025)
Similar Items
-
Corrected Samplers for Discrete Flow Models
by: Wan, Zhengyan, et al.
Published: (2026) -
Transfer Learning for Diffusion Models
by: Ouyang, Yidong, et al.
Published: (2024) -
Alignment of Diffusion Model and Flow Matching for Text-to-Image Generation
by: Ouyang, Yidong, et al.
Published: (2026) -
Discrete Guidance Matching: Exact Guidance for Discrete Flow Matching
by: Wan, Zhengyan, et al.
Published: (2025) -
Error Analysis of Discrete Flow with Generator Matching
by: Wan, Zhengyan, et al.
Published: (2025)