Taming Sampling Perturbations with Variance Expansion Loss for Latent Diffusion Models
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Qifan, Zhou, Xingyu, Zhang, Jinhua, You, Weiyi, Gu, Shuhang |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Uncertainty-guided Perturbation for Image Super-Resolution Diffusion Model
by: Zhang, Leheng, et al.
Published: (2025)
by: Zhang, Leheng, et al.
Published: (2025)
MVAR: Visual Autoregressive Modeling with Scale and Spatial Markovian Conditioning
by: Zhang, Jinhua, et al.
Published: (2025)
by: Zhang, Jinhua, et al.
Published: (2025)
Texture Vector-Quantization and Reconstruction Aware Prediction for Generative Super-Resolution
by: Li, Qifan, et al.
Published: (2025)
by: Li, Qifan, et al.
Published: (2025)
Guiding a Diffusion Transformer with the Internal Dynamics of Itself
by: Zhou, Xingyu, et al.
Published: (2025)
by: Zhou, Xingyu, et al.
Published: (2025)
Consistency Trajectory Matching for One-Step Generative Super-Resolution
by: You, Weiyi, et al.
Published: (2025)
by: You, Weiyi, et al.
Published: (2025)
Generative Image Compression by Estimating Gradients of the Rate-variable Feature Distribution
by: Han, Minghao, et al.
Published: (2025)
by: Han, Minghao, et al.
Published: (2025)
Small Clips, Big Gains: Learning Long-Range Refocused Temporal Information for Video Super-Resolution
by: Zhou, Xingyu, et al.
Published: (2025)
by: Zhou, Xingyu, et al.
Published: (2025)
Improved Implicit Neural Representation with Fourier Reparameterized Training
by: Shi, Kexuan, et al.
Published: (2024)
by: Shi, Kexuan, et al.
Published: (2024)
Progressive Focused Transformer for Single Image Super-Resolution
by: Long, Wei, et al.
Published: (2025)
by: Long, Wei, et al.
Published: (2025)
Learned Image Compression with Dictionary-based Entropy Model
by: Lu, Jingbo, et al.
Published: (2025)
by: Lu, Jingbo, et al.
Published: (2025)
Transcending the Limit of Local Window: Advanced Super-Resolution Transformer with Adaptive Token Dictionary
by: Zhang, Leheng, et al.
Published: (2024)
by: Zhang, Leheng, et al.
Published: (2024)
PerLDiff: Controllable Street View Synthesis Using Perspective-Layout Diffusion Models
by: Zhang, Jinhua, et al.
Published: (2024)
by: Zhang, Jinhua, et al.
Published: (2024)
LatentSync: Taming Audio-Conditioned Latent Diffusion Models for Lip Sync with SyncNet Supervision
by: Li, Chunyu, et al.
Published: (2024)
by: Li, Chunyu, et al.
Published: (2024)
ATD: Improved Transformer with Adaptive Token Dictionary for Image Restoration
by: Zhang, Leheng, et al.
Published: (2026)
by: Zhang, Leheng, et al.
Published: (2026)
Video Super-Resolution Transformer with Masked Inter&Intra-Frame Attention
by: Zhou, Xingyu, et al.
Published: (2024)
by: Zhou, Xingyu, et al.
Published: (2024)
DOME: Taming Diffusion Model into High-Fidelity Controllable Occupancy World Model
by: Gu, Songen, et al.
Published: (2024)
by: Gu, Songen, et al.
Published: (2024)
ChronosObserver: Taming 4D World with Hyperspace Diffusion Sampling
by: Wang, Qisen, et al.
Published: (2025)
by: Wang, Qisen, et al.
Published: (2025)
Reconstruction vs. Generation: Taming Optimization Dilemma in Latent Diffusion Models
by: Yao, Jingfeng, et al.
Published: (2025)
by: Yao, Jingfeng, et al.
Published: (2025)
DepthMaster: Taming Diffusion Models for Monocular Depth Estimation
by: Song, Ziyang, et al.
Published: (2025)
by: Song, Ziyang, et al.
Published: (2025)
Taming Latent Diffusion Model for Neural Radiance Field Inpainting
by: Lin, Chieh Hubert, et al.
Published: (2024)
by: Lin, Chieh Hubert, et al.
Published: (2024)
Unified Directly Denoising for Both Variance Preserving and Variance Exploding Diffusion Models
by: Wang, Jingjing, et al.
Published: (2024)
by: Wang, Jingjing, et al.
Published: (2024)
Kaleido Diffusion: Improving Conditional Diffusion Models with Autoregressive Latent Modeling
by: Gu, Jiatao, et al.
Published: (2024)
by: Gu, Jiatao, et al.
Published: (2024)
IDESplat: Iterative Depth Probability Estimation for Generalizable 3D Gaussian Splatting
by: Long, Wei, et al.
Published: (2026)
by: Long, Wei, et al.
Published: (2026)
Taming Feed-forward Reconstruction Models as Latent Encoders for 3D Generative Models
by: Wizadwongsa, Suttisak, et al.
Published: (2024)
by: Wizadwongsa, Suttisak, et al.
Published: (2024)
DiP: Taming Diffusion Models in Pixel Space
by: Chen, Zhennan, et al.
Published: (2025)
by: Chen, Zhennan, et al.
Published: (2025)
Ultrasound Image Enhancement with the Variance of Diffusion Models
by: Zhang, Yuxin, et al.
Published: (2024)
by: Zhang, Yuxin, et al.
Published: (2024)
UniAnimate: Taming Unified Video Diffusion Models for Consistent Human Image Animation
by: Wang, Xiang, et al.
Published: (2024)
by: Wang, Xiang, et al.
Published: (2024)
CAT: Contrastive Adversarial Training for Evaluating the Robustness of Protective Perturbations in Latent Diffusion Models
by: Peng, Sen, et al.
Published: (2025)
by: Peng, Sen, et al.
Published: (2025)
HoloTime: Taming Video Diffusion Models for Panoramic 4D Scene Generation
by: Zhou, Haiyang, et al.
Published: (2025)
by: Zhou, Haiyang, et al.
Published: (2025)
An Efficient Watermarking Method for Latent Diffusion Models via Low-Rank Adaptation and Dynamic Loss Weighting
by: Lin, Dongdong, et al.
Published: (2024)
by: Lin, Dongdong, et al.
Published: (2024)
Taming Diffusion Models for Image Restoration: A Review
by: Luo, Ziwei, et al.
Published: (2024)
by: Luo, Ziwei, et al.
Published: (2024)
LatentINDIGO: An INN-Guided Latent Diffusion Algorithm for Image Restoration
by: You, Di, et al.
Published: (2025)
by: You, Di, et al.
Published: (2025)
Zero-1-to-G: Taming Pretrained 2D Diffusion Model for Direct 3D Generation
by: Meng, Xuyi, et al.
Published: (2025)
by: Meng, Xuyi, et al.
Published: (2025)
From Local Windows to Adaptive Candidates via Individualized Exploratory: Rethinking Attention for Image Super-Resolution
by: Meng, Chunyu, et al.
Published: (2026)
by: Meng, Chunyu, et al.
Published: (2026)
Taming the Long Tail: Rebalancing Adversarial Training via Adaptive Perturbation
by: Zhang, Lilin, et al.
Published: (2026)
by: Zhang, Lilin, et al.
Published: (2026)
VSDiffusion: Taming Ill-Posed Shadow Generation via Visibility-Constrained Diffusion
by: Li, Jing, et al.
Published: (2026)
by: Li, Jing, et al.
Published: (2026)
FairDiffusion: Enhancing Equity in Latent Diffusion Models via Fair Bayesian Perturbation
by: Luo, Yan, et al.
Published: (2024)
by: Luo, Yan, et al.
Published: (2024)
ControlSR: Taming Diffusion Models for Consistent Real-World Image Super Resolution
by: Wan, Yuhao, et al.
Published: (2024)
by: Wan, Yuhao, et al.
Published: (2024)
Alias-Free Latent Diffusion Models: Improving Fractional Shift Equivariance of Diffusion Latent Space
by: Zhou, Yifan, et al.
Published: (2025)
by: Zhou, Yifan, et al.
Published: (2025)
DiffSim: Taming Diffusion Models for Evaluating Visual Similarity
by: Song, Yiren, et al.
Published: (2024)
by: Song, Yiren, et al.
Published: (2024)
Similar Items
-
Uncertainty-guided Perturbation for Image Super-Resolution Diffusion Model
by: Zhang, Leheng, et al.
Published: (2025) -
MVAR: Visual Autoregressive Modeling with Scale and Spatial Markovian Conditioning
by: Zhang, Jinhua, et al.
Published: (2025) -
Texture Vector-Quantization and Reconstruction Aware Prediction for Generative Super-Resolution
by: Li, Qifan, et al.
Published: (2025) -
Guiding a Diffusion Transformer with the Internal Dynamics of Itself
by: Zhou, Xingyu, et al.
Published: (2025) -
Consistency Trajectory Matching for One-Step Generative Super-Resolution
by: You, Weiyi, et al.
Published: (2025)