Reconstruction vs. Generation: Taming Optimization Dilemma in Latent Diffusion Models
Fuente:
arXiv
Saved in:
| Main Authors: | Yao, Jingfeng, Yang, Bin, Wang, Xinggang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Taming Latent Diffusion Model for Neural Radiance Field Inpainting
by: Lin, Chieh Hubert, et al.
Published: (2024)
by: Lin, Chieh Hubert, et al.
Published: (2024)
Taming Feed-forward Reconstruction Models as Latent Encoders for 3D Generative Models
by: Wizadwongsa, Suttisak, et al.
Published: (2024)
by: Wizadwongsa, Suttisak, et al.
Published: (2024)
Taming Generative Diffusion Prior for Universal Blind Image Restoration
by: Tu, Siwei, et al.
Published: (2024)
by: Tu, Siwei, et al.
Published: (2024)
FasterDiT: Towards Faster Diffusion Transformers Training without Architecture Modification
by: Yao, Jingfeng, et al.
Published: (2024)
by: Yao, Jingfeng, et al.
Published: (2024)
Iterative CT Reconstruction via Latent Variable Optimization of Shallow Diffusion Models
by: Ozaki, Sho, et al.
Published: (2024)
by: Ozaki, Sho, et al.
Published: (2024)
Privacy-Preserving Low-Rank Adaptation against Membership Inference Attacks for Latent Diffusion Models
by: Luo, Zihao, et al.
Published: (2024)
by: Luo, Zihao, et al.
Published: (2024)
DiffusionVL: Translating Any Autoregressive Models into Diffusion Vision Language Models
by: Zeng, Lunbin, et al.
Published: (2025)
by: Zeng, Lunbin, et al.
Published: (2025)
Towards Scalable Pre-training of Visual Tokenizers for Generation
by: Yao, Jingfeng, et al.
Published: (2025)
by: Yao, Jingfeng, et al.
Published: (2025)
Corruption-Aware Training of Latent Video Diffusion Models for Robust Text-to-Video Generation
by: Maduabuchi, Chika, et al.
Published: (2025)
by: Maduabuchi, Chika, et al.
Published: (2025)
Matte Anything: Interactive Natural Image Matting with Segment Anything Models
by: Yao, Jingfeng, et al.
Published: (2023)
by: Yao, Jingfeng, et al.
Published: (2023)
Multimodal Latent Diffusion Model for Complex Sewing Pattern Generation
by: Liu, Shengqi, et al.
Published: (2024)
by: Liu, Shengqi, et al.
Published: (2024)
What's in a Latent? Leveraging Diffusion Latent Space for Domain Generalization
by: Thomas, Xavier, et al.
Published: (2025)
by: Thomas, Xavier, et al.
Published: (2025)
BrepGen: A B-rep Generative Diffusion Model with Structured Latent Geometry
by: Xu, Xiang, et al.
Published: (2024)
by: Xu, Xiang, et al.
Published: (2024)
Taming Outlier Tokens in Diffusion Transformers
by: Wu, Xiaoyu, et al.
Published: (2026)
by: Wu, Xiaoyu, et al.
Published: (2026)
JLT: Clean-Latent Prediction in Latent Diffusion Transformers
by: Fu, Funing, et al.
Published: (2026)
by: Fu, Funing, et al.
Published: (2026)
LatentMan: Generating Consistent Animated Characters using Image Diffusion Models
by: Eldesokey, Abdelrahman, et al.
Published: (2023)
by: Eldesokey, Abdelrahman, et al.
Published: (2023)
LaVR: Scene Latent Conditioned Generative Video Trajectory Re-Rendering using Large 4D Reconstruction Models
by: Xie, Mingyang, et al.
Published: (2026)
by: Xie, Mingyang, et al.
Published: (2026)
Taming Mode Collapse in Score Distillation for Text-to-3D Generation
by: Wang, Peihao, et al.
Published: (2023)
by: Wang, Peihao, et al.
Published: (2023)
Generative Latent Diffusion for Efficient Spatiotemporal Data Reduction
by: Li, Xiao, et al.
Published: (2025)
by: Li, Xiao, et al.
Published: (2025)
CASL: Concept-Aligned Sparse Latents for Interpreting Diffusion Models
by: He, Zhenghao, et al.
Published: (2026)
by: He, Zhenghao, et al.
Published: (2026)
Canonical Latent Representations in Conditional Diffusion Models
by: Xu, Yitao, et al.
Published: (2025)
by: Xu, Yitao, et al.
Published: (2025)
Single Mesh Diffusion Models with Field Latents for Texture Generation
by: Mitchel, Thomas W., et al.
Published: (2023)
by: Mitchel, Thomas W., et al.
Published: (2023)
Taming Diffusion for Dataset Distillation with High Representativeness
by: Zhao, Lin, et al.
Published: (2025)
by: Zhao, Lin, et al.
Published: (2025)
Segment Anything without Supervision
by: Wang, XuDong, et al.
Published: (2024)
by: Wang, XuDong, et al.
Published: (2024)
Making Reconstruction FID Predictive of Diffusion Generation FID
by: Xu, Tongda, et al.
Published: (2026)
by: Xu, Tongda, et al.
Published: (2026)
LURE: Latent Space Unblocking for Multi-Concept Reawakening in Diffusion Models
by: Sun, Mengyu, et al.
Published: (2026)
by: Sun, Mengyu, et al.
Published: (2026)
FairDiffusion: Enhancing Equity in Latent Diffusion Models via Fair Bayesian Perturbation
by: Luo, Yan, et al.
Published: (2024)
by: Luo, Yan, et al.
Published: (2024)
Orchid: Image Latent Diffusion for Joint Appearance and Geometry Generation
by: Krishnan, Akshay, et al.
Published: (2025)
by: Krishnan, Akshay, et al.
Published: (2025)
DP-aware AdaLN-Zero: Taming Conditioning-Induced Heavy-Tailed Gradients in Differentially Private Diffusion
by: Huang, Tao, et al.
Published: (2026)
by: Huang, Tao, et al.
Published: (2026)
Vision Mamba: Efficient Visual Representation Learning with Bidirectional State Space Model
by: Zhu, Lianghui, et al.
Published: (2024)
by: Zhu, Lianghui, et al.
Published: (2024)
RenderDiffusion: Image Diffusion for 3D Reconstruction, Inpainting and Generation
by: Anciukevičius, Titas, et al.
Published: (2022)
by: Anciukevičius, Titas, et al.
Published: (2022)
Learning Latent Space Hierarchical EBM Diffusion Models
by: Cui, Jiali, et al.
Published: (2024)
by: Cui, Jiali, et al.
Published: (2024)
Gradient-free Decoder Inversion in Latent Diffusion Models
by: Hong, Seongmin, et al.
Published: (2024)
by: Hong, Seongmin, et al.
Published: (2024)
Bayesian Diffusion Models for 3D Shape Reconstruction
by: Xu, Haiyang, et al.
Published: (2024)
by: Xu, Haiyang, et al.
Published: (2024)
Latent Diffusion Inversion Requires Understanding the Latent Space
by: Rao, Mingxing, et al.
Published: (2025)
by: Rao, Mingxing, et al.
Published: (2025)
AdvDiff: Generating Unrestricted Adversarial Examples using Diffusion Models
by: Dai, Xuelong, et al.
Published: (2023)
by: Dai, Xuelong, et al.
Published: (2023)
Multimodal Latent Language Modeling with Next-Token Diffusion
by: Sun, Yutao, et al.
Published: (2024)
by: Sun, Yutao, et al.
Published: (2024)
Latent Forcing: Reordering the Diffusion Trajectory for Pixel-Space Image Generation
by: Baade, Alan, et al.
Published: (2026)
by: Baade, Alan, et al.
Published: (2026)
Optimizing Diffusion Priors in Image Reconstruction from a Single Observation
by: Wang, Frederic, et al.
Published: (2026)
by: Wang, Frederic, et al.
Published: (2026)
Discovering Interpretable Directions in the Semantic Latent Space of Diffusion Models
by: Haas, René, et al.
Published: (2023)
by: Haas, René, et al.
Published: (2023)
Similar Items
-
Taming Latent Diffusion Model for Neural Radiance Field Inpainting
by: Lin, Chieh Hubert, et al.
Published: (2024) -
Taming Feed-forward Reconstruction Models as Latent Encoders for 3D Generative Models
by: Wizadwongsa, Suttisak, et al.
Published: (2024) -
Taming Generative Diffusion Prior for Universal Blind Image Restoration
by: Tu, Siwei, et al.
Published: (2024) -
FasterDiT: Towards Faster Diffusion Transformers Training without Architecture Modification
by: Yao, Jingfeng, et al.
Published: (2024) -
Iterative CT Reconstruction via Latent Variable Optimization of Shallow Diffusion Models
by: Ozaki, Sho, et al.
Published: (2024)