Geometric Autoencoder for Diffusion Models
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Hangyu, Wang, Jianyong, Sun, Yutao |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multimodal Latent Language Modeling with Next-Token Diffusion
by: Sun, Yutao, et al.
Published: (2024)
by: Sun, Yutao, et al.
Published: (2024)
Adaptive 1D Video Diffusion Autoencoder
by: Teng, Yao, et al.
Published: (2026)
by: Teng, Yao, et al.
Published: (2026)
GeoGPT4V: Towards Geometric Multi-modal Large Language Models with Geometric Image Generation
by: Cai, Shihao, et al.
Published: (2024)
by: Cai, Shihao, et al.
Published: (2024)
What Matters for Diffusion-Friendly Latent Manifold? Prior-Aligned Autoencoders for Latent Diffusion
by: Yue, Zhengrong, et al.
Published: (2026)
by: Yue, Zhengrong, et al.
Published: (2026)
Latent Diffusion Model without Variational Autoencoder
by: Shi, Minglei, et al.
Published: (2025)
by: Shi, Minglei, et al.
Published: (2025)
Training-free Geometric Image Editing on Diffusion Models
by: Zhu, Hanshen, et al.
Published: (2025)
by: Zhu, Hanshen, et al.
Published: (2025)
Supervise-assisted Multi-modality Fusion Diffusion Model for PET Restoration
by: Zhang, Yingkai, et al.
Published: (2026)
by: Zhang, Yingkai, et al.
Published: (2026)
Repurposing Geometric Foundation Models for Multi-view Diffusion
by: Jang, Wooseok, et al.
Published: (2026)
by: Jang, Wooseok, et al.
Published: (2026)
Laminating Representation Autoencoders for Efficient Diffusion
by: Calvo-González, Ramón, et al.
Published: (2026)
by: Calvo-González, Ramón, et al.
Published: (2026)
Efficient Conditional Diffusion Model with Probability Flow Sampling for Image Super-resolution
by: Yuan, Yutao, et al.
Published: (2024)
by: Yuan, Yutao, et al.
Published: (2024)
SRA 2: Variational Autoencoder Self-Representation Alignment for Efficient Diffusion Training
by: Wang, Mengmeng, et al.
Published: (2026)
by: Wang, Mengmeng, et al.
Published: (2026)
Variational Autoencoding Discrete Diffusion with Enhanced Dimensional Correlations Modeling
by: Xie, Tianyu, et al.
Published: (2025)
by: Xie, Tianyu, et al.
Published: (2025)
Landscape-Awareness for Geometric View Diffusion Model
by: Chen, Yan-Ting, et al.
Published: (2026)
by: Chen, Yan-Ting, et al.
Published: (2026)
$Z^2$-Sampling: Zero-Cost Zigzag Trajectories for Semantic Alignment in Diffusion Models
by: Li, Haosen, et al.
Published: (2026)
by: Li, Haosen, et al.
Published: (2026)
Scaling Text-to-Image Diffusion Transformers with Representation Autoencoders
by: Tong, Shengbang, et al.
Published: (2026)
by: Tong, Shengbang, et al.
Published: (2026)
Guided Path Sampling: Steering Diffusion Models Back on Track with Principled Path Guidance
by: Li, Haosen, et al.
Published: (2025)
by: Li, Haosen, et al.
Published: (2025)
Masked Autoencoders Are Effective Tokenizers for Diffusion Models
by: Chen, Hao, et al.
Published: (2025)
by: Chen, Hao, et al.
Published: (2025)
Delta Score Matters! Spatial Adaptive Multi Guidance in Diffusion Models
by: Li, Haosen, et al.
Published: (2026)
by: Li, Haosen, et al.
Published: (2026)
Rethinking Target Label Conditioning in Adversarial Attacks: A 2D Tensor-Guided Generative Approach
by: Liu, Hangyu, et al.
Published: (2025)
by: Liu, Hangyu, et al.
Published: (2025)
DM-SegNet: Dual-Mamba Architecture for 3D Medical Image Segmentation with Global Context Modeling
by: Ji, Hangyu
Published: (2025)
by: Ji, Hangyu
Published: (2025)
SVG-T2I: Scaling Up Text-to-Image Latent Diffusion Model Without Variational Autoencoder
by: Shi, Minglei, et al.
Published: (2025)
by: Shi, Minglei, et al.
Published: (2025)
A General Framework to Boost 3D GS Initialization for Text-to-3D Generation by Lexical Richness
by: Jiang, Lutao, et al.
Published: (2024)
by: Jiang, Lutao, et al.
Published: (2024)
A Geometric Perspective on Diffusion Models
by: Chen, Defang, et al.
Published: (2023)
by: Chen, Defang, et al.
Published: (2023)
Normative Diffusion Autoencoders: Application to Amyotrophic Lateral Sclerosis
by: Ijishakin, Ayodeji, et al.
Published: (2024)
by: Ijishakin, Ayodeji, et al.
Published: (2024)
Geometric Trajectory Diffusion Models
by: Han, Jiaqi, et al.
Published: (2024)
by: Han, Jiaqi, et al.
Published: (2024)
Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation
by: Shi, Dingyuan, et al.
Published: (2024)
by: Shi, Dingyuan, et al.
Published: (2024)
Latent-Compressed Variational Autoencoder for Video Diffusion Models
by: Guan, Jiarui, et al.
Published: (2026)
by: Guan, Jiarui, et al.
Published: (2026)
DiffPMAE: Diffusion Masked Autoencoders for Point Cloud Reconstruction
by: Li, Yanlong, et al.
Published: (2023)
by: Li, Yanlong, et al.
Published: (2023)
Disentangled Diffusion Autoencoder for Harmonization of Multi-site Neuroimaging Data
by: Ijishakin, Ayodeji, et al.
Published: (2024)
by: Ijishakin, Ayodeji, et al.
Published: (2024)
Diffusion Transformers with Representation Autoencoders
by: Zheng, Boyang, et al.
Published: (2025)
by: Zheng, Boyang, et al.
Published: (2025)
TransParking: A Dual-Decoder Transformer Framework with Soft Localization for End-to-End Automatic Parking
by: Du, Hangyu, et al.
Published: (2025)
by: Du, Hangyu, et al.
Published: (2025)
POLARIS: Projection-Orthogonal Least Squares for Robust and Adaptive Inversion in Diffusion Models
by: Chen, Wenshuo, et al.
Published: (2025)
by: Chen, Wenshuo, et al.
Published: (2025)
GeoSceneGraph: Geometric Scene Graph Diffusion Model for Text-guided 3D Indoor Scene Synthesis
by: Ruiz, Antonio, et al.
Published: (2025)
by: Ruiz, Antonio, et al.
Published: (2025)
Flexible Geometric Guidance for Probabilistic Human Pose Estimation with Diffusion Models
by: Snelgar, Francis, et al.
Published: (2026)
by: Snelgar, Francis, et al.
Published: (2026)
MARMOT: Masked Autoencoder for Modeling Transient Imaging
by: Shen, Siyuan, et al.
Published: (2025)
by: Shen, Siyuan, et al.
Published: (2025)
A Tilted Seesaw: Revisiting Autoencoder Trade-off for Controllable Diffusion
by: Cao, Pu, et al.
Published: (2026)
by: Cao, Pu, et al.
Published: (2026)
Feature Denoising Diffusion Model for Blind Image Quality Assessment
by: Li, Xudong, et al.
Published: (2024)
by: Li, Xudong, et al.
Published: (2024)
Enhancing Adversarial Transferability via Component-Wise Transformation
by: Liu, Hangyu, et al.
Published: (2025)
by: Liu, Hangyu, et al.
Published: (2025)
Deep Compression Autoencoder for Efficient High-Resolution Diffusion Models
by: Chen, Junyu, et al.
Published: (2024)
by: Chen, Junyu, et al.
Published: (2024)
SAEmnesia: Erasing Concepts in Diffusion Models with Supervised Sparse Autoencoders
by: Cassano, Enrico, et al.
Published: (2025)
by: Cassano, Enrico, et al.
Published: (2025)
Similar Items
-
Multimodal Latent Language Modeling with Next-Token Diffusion
by: Sun, Yutao, et al.
Published: (2024) -
Adaptive 1D Video Diffusion Autoencoder
by: Teng, Yao, et al.
Published: (2026) -
GeoGPT4V: Towards Geometric Multi-modal Large Language Models with Geometric Image Generation
by: Cai, Shihao, et al.
Published: (2024) -
What Matters for Diffusion-Friendly Latent Manifold? Prior-Aligned Autoencoders for Latent Diffusion
by: Yue, Zhengrong, et al.
Published: (2026) -
Latent Diffusion Model without Variational Autoencoder
by: Shi, Minglei, et al.
Published: (2025)