Reconstruction vs. Generation: Taming Optimization Dilemma in Latent Diffusion Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Yao, Jingfeng, Yang, Bin, Wang, Xinggang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Taming Latent Diffusion Model for Neural Radiance Field Inpainting
di: Lin, Chieh Hubert, et al.
Pubblicazione: (2024)
di: Lin, Chieh Hubert, et al.
Pubblicazione: (2024)
Taming Feed-forward Reconstruction Models as Latent Encoders for 3D Generative Models
di: Wizadwongsa, Suttisak, et al.
Pubblicazione: (2024)
di: Wizadwongsa, Suttisak, et al.
Pubblicazione: (2024)
Taming Generative Diffusion Prior for Universal Blind Image Restoration
di: Tu, Siwei, et al.
Pubblicazione: (2024)
di: Tu, Siwei, et al.
Pubblicazione: (2024)
FasterDiT: Towards Faster Diffusion Transformers Training without Architecture Modification
di: Yao, Jingfeng, et al.
Pubblicazione: (2024)
di: Yao, Jingfeng, et al.
Pubblicazione: (2024)
Iterative CT Reconstruction via Latent Variable Optimization of Shallow Diffusion Models
di: Ozaki, Sho, et al.
Pubblicazione: (2024)
di: Ozaki, Sho, et al.
Pubblicazione: (2024)
Privacy-Preserving Low-Rank Adaptation against Membership Inference Attacks for Latent Diffusion Models
di: Luo, Zihao, et al.
Pubblicazione: (2024)
di: Luo, Zihao, et al.
Pubblicazione: (2024)
DiffusionVL: Translating Any Autoregressive Models into Diffusion Vision Language Models
di: Zeng, Lunbin, et al.
Pubblicazione: (2025)
di: Zeng, Lunbin, et al.
Pubblicazione: (2025)
Towards Scalable Pre-training of Visual Tokenizers for Generation
di: Yao, Jingfeng, et al.
Pubblicazione: (2025)
di: Yao, Jingfeng, et al.
Pubblicazione: (2025)
Corruption-Aware Training of Latent Video Diffusion Models for Robust Text-to-Video Generation
di: Maduabuchi, Chika, et al.
Pubblicazione: (2025)
di: Maduabuchi, Chika, et al.
Pubblicazione: (2025)
Matte Anything: Interactive Natural Image Matting with Segment Anything Models
di: Yao, Jingfeng, et al.
Pubblicazione: (2023)
di: Yao, Jingfeng, et al.
Pubblicazione: (2023)
Multimodal Latent Diffusion Model for Complex Sewing Pattern Generation
di: Liu, Shengqi, et al.
Pubblicazione: (2024)
di: Liu, Shengqi, et al.
Pubblicazione: (2024)
What's in a Latent? Leveraging Diffusion Latent Space for Domain Generalization
di: Thomas, Xavier, et al.
Pubblicazione: (2025)
di: Thomas, Xavier, et al.
Pubblicazione: (2025)
BrepGen: A B-rep Generative Diffusion Model with Structured Latent Geometry
di: Xu, Xiang, et al.
Pubblicazione: (2024)
di: Xu, Xiang, et al.
Pubblicazione: (2024)
Taming Outlier Tokens in Diffusion Transformers
di: Wu, Xiaoyu, et al.
Pubblicazione: (2026)
di: Wu, Xiaoyu, et al.
Pubblicazione: (2026)
JLT: Clean-Latent Prediction in Latent Diffusion Transformers
di: Fu, Funing, et al.
Pubblicazione: (2026)
di: Fu, Funing, et al.
Pubblicazione: (2026)
LatentMan: Generating Consistent Animated Characters using Image Diffusion Models
di: Eldesokey, Abdelrahman, et al.
Pubblicazione: (2023)
di: Eldesokey, Abdelrahman, et al.
Pubblicazione: (2023)
LaVR: Scene Latent Conditioned Generative Video Trajectory Re-Rendering using Large 4D Reconstruction Models
di: Xie, Mingyang, et al.
Pubblicazione: (2026)
di: Xie, Mingyang, et al.
Pubblicazione: (2026)
Taming Mode Collapse in Score Distillation for Text-to-3D Generation
di: Wang, Peihao, et al.
Pubblicazione: (2023)
di: Wang, Peihao, et al.
Pubblicazione: (2023)
Generative Latent Diffusion for Efficient Spatiotemporal Data Reduction
di: Li, Xiao, et al.
Pubblicazione: (2025)
di: Li, Xiao, et al.
Pubblicazione: (2025)
CASL: Concept-Aligned Sparse Latents for Interpreting Diffusion Models
di: He, Zhenghao, et al.
Pubblicazione: (2026)
di: He, Zhenghao, et al.
Pubblicazione: (2026)
Canonical Latent Representations in Conditional Diffusion Models
di: Xu, Yitao, et al.
Pubblicazione: (2025)
di: Xu, Yitao, et al.
Pubblicazione: (2025)
Single Mesh Diffusion Models with Field Latents for Texture Generation
di: Mitchel, Thomas W., et al.
Pubblicazione: (2023)
di: Mitchel, Thomas W., et al.
Pubblicazione: (2023)
Taming Diffusion for Dataset Distillation with High Representativeness
di: Zhao, Lin, et al.
Pubblicazione: (2025)
di: Zhao, Lin, et al.
Pubblicazione: (2025)
Segment Anything without Supervision
di: Wang, XuDong, et al.
Pubblicazione: (2024)
di: Wang, XuDong, et al.
Pubblicazione: (2024)
Making Reconstruction FID Predictive of Diffusion Generation FID
di: Xu, Tongda, et al.
Pubblicazione: (2026)
di: Xu, Tongda, et al.
Pubblicazione: (2026)
LURE: Latent Space Unblocking for Multi-Concept Reawakening in Diffusion Models
di: Sun, Mengyu, et al.
Pubblicazione: (2026)
di: Sun, Mengyu, et al.
Pubblicazione: (2026)
FairDiffusion: Enhancing Equity in Latent Diffusion Models via Fair Bayesian Perturbation
di: Luo, Yan, et al.
Pubblicazione: (2024)
di: Luo, Yan, et al.
Pubblicazione: (2024)
Orchid: Image Latent Diffusion for Joint Appearance and Geometry Generation
di: Krishnan, Akshay, et al.
Pubblicazione: (2025)
di: Krishnan, Akshay, et al.
Pubblicazione: (2025)
DP-aware AdaLN-Zero: Taming Conditioning-Induced Heavy-Tailed Gradients in Differentially Private Diffusion
di: Huang, Tao, et al.
Pubblicazione: (2026)
di: Huang, Tao, et al.
Pubblicazione: (2026)
Vision Mamba: Efficient Visual Representation Learning with Bidirectional State Space Model
di: Zhu, Lianghui, et al.
Pubblicazione: (2024)
di: Zhu, Lianghui, et al.
Pubblicazione: (2024)
RenderDiffusion: Image Diffusion for 3D Reconstruction, Inpainting and Generation
di: Anciukevičius, Titas, et al.
Pubblicazione: (2022)
di: Anciukevičius, Titas, et al.
Pubblicazione: (2022)
Learning Latent Space Hierarchical EBM Diffusion Models
di: Cui, Jiali, et al.
Pubblicazione: (2024)
di: Cui, Jiali, et al.
Pubblicazione: (2024)
Gradient-free Decoder Inversion in Latent Diffusion Models
di: Hong, Seongmin, et al.
Pubblicazione: (2024)
di: Hong, Seongmin, et al.
Pubblicazione: (2024)
Bayesian Diffusion Models for 3D Shape Reconstruction
di: Xu, Haiyang, et al.
Pubblicazione: (2024)
di: Xu, Haiyang, et al.
Pubblicazione: (2024)
Latent Diffusion Inversion Requires Understanding the Latent Space
di: Rao, Mingxing, et al.
Pubblicazione: (2025)
di: Rao, Mingxing, et al.
Pubblicazione: (2025)
AdvDiff: Generating Unrestricted Adversarial Examples using Diffusion Models
di: Dai, Xuelong, et al.
Pubblicazione: (2023)
di: Dai, Xuelong, et al.
Pubblicazione: (2023)
Multimodal Latent Language Modeling with Next-Token Diffusion
di: Sun, Yutao, et al.
Pubblicazione: (2024)
di: Sun, Yutao, et al.
Pubblicazione: (2024)
Latent Forcing: Reordering the Diffusion Trajectory for Pixel-Space Image Generation
di: Baade, Alan, et al.
Pubblicazione: (2026)
di: Baade, Alan, et al.
Pubblicazione: (2026)
Optimizing Diffusion Priors in Image Reconstruction from a Single Observation
di: Wang, Frederic, et al.
Pubblicazione: (2026)
di: Wang, Frederic, et al.
Pubblicazione: (2026)
Discovering Interpretable Directions in the Semantic Latent Space of Diffusion Models
di: Haas, René, et al.
Pubblicazione: (2023)
di: Haas, René, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Taming Latent Diffusion Model for Neural Radiance Field Inpainting
di: Lin, Chieh Hubert, et al.
Pubblicazione: (2024) -
Taming Feed-forward Reconstruction Models as Latent Encoders for 3D Generative Models
di: Wizadwongsa, Suttisak, et al.
Pubblicazione: (2024) -
Taming Generative Diffusion Prior for Universal Blind Image Restoration
di: Tu, Siwei, et al.
Pubblicazione: (2024) -
FasterDiT: Towards Faster Diffusion Transformers Training without Architecture Modification
di: Yao, Jingfeng, et al.
Pubblicazione: (2024) -
Iterative CT Reconstruction via Latent Variable Optimization of Shallow Diffusion Models
di: Ozaki, Sho, et al.
Pubblicazione: (2024)