VolumeDiffusion: Flexible Text-to-3D Generation with Efficient Volumetric Encoder
Fuente:
arXiv
Saved in:
| Main Authors: | Tang, Zhicong, Gu, Shuyang, Wang, Chunyu, Zhang, Ting, Bao, Jianmin, Chen, Dong, Guo, Baining |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Diffusion Models without Classifier-free Guidance
by: Tang, Zhicong, et al.
Published: (2025)
by: Tang, Zhicong, et al.
Published: (2025)
Simplified Diffusion Schrödinger Bridge
by: Tang, Zhicong, et al.
Published: (2024)
by: Tang, Zhicong, et al.
Published: (2024)
Incorporating Pre-trained Diffusion Models in Solving the Schrödinger Bridge Problem
by: Tang, Zhicong, et al.
Published: (2025)
by: Tang, Zhicong, et al.
Published: (2025)
Efficient Diffusion Training via Min-SNR Weighting Strategy
by: Hang, Tiankai, et al.
Published: (2023)
by: Hang, Tiankai, et al.
Published: (2023)
DesignDiffusion: High-Quality Text-to-Design Image Generation with Diffusion Models
by: Wang, Zhendong, et al.
Published: (2025)
by: Wang, Zhendong, et al.
Published: (2025)
RodinHD: High-Fidelity 3D Avatar Generation with Diffusion Models
by: Zhang, Bowen, et al.
Published: (2024)
by: Zhang, Bowen, et al.
Published: (2024)
Improved Noise Schedule for Diffusion Training
by: Hang, Tiankai, et al.
Published: (2024)
by: Hang, Tiankai, et al.
Published: (2024)
CCA: Collaborative Competitive Agents for Image Editing
by: Hang, Tiankai, et al.
Published: (2024)
by: Hang, Tiankai, et al.
Published: (2024)
GaussianCube: A Structured and Explicit Radiance Representation for 3D Generative Modeling
by: Zhang, Bowen, et al.
Published: (2024)
by: Zhang, Bowen, et al.
Published: (2024)
FontStudio: Shape-Adaptive Diffusion Model for Coherent and Consistent Font Effect Generation
by: Mu, Xinzhi, et al.
Published: (2024)
by: Mu, Xinzhi, et al.
Published: (2024)
Distribution Matching Variational AutoEncoder
by: Ye, Sen, et al.
Published: (2025)
by: Ye, Sen, et al.
Published: (2025)
MicroCinema: A Divide-and-Conquer Approach for Text-to-Video Generation
by: Wang, Yanhui, et al.
Published: (2023)
by: Wang, Yanhui, et al.
Published: (2023)
CCEdit: Creative and Controllable Video Editing via Diffusion Models
by: Feng, Ruoyu, et al.
Published: (2023)
by: Feng, Ruoyu, et al.
Published: (2023)
GVGEN: Text-to-3D Generation with Volumetric Representation
by: He, Xianglong, et al.
Published: (2024)
by: He, Xianglong, et al.
Published: (2024)
PI3D: Efficient Text-to-3D Generation with Pseudo-Image Diffusion
by: Liu, Ying-Tian, et al.
Published: (2023)
by: Liu, Ying-Tian, et al.
Published: (2023)
ART: Anonymous Region Transformer for Variable Multi-Layer Transparent Image Generation
by: Pu, Yifan, et al.
Published: (2025)
by: Pu, Yifan, et al.
Published: (2025)
Gaussian Variation Field Diffusion for High-fidelity Video-to-4D Synthesis
by: Zhang, Bowen, et al.
Published: (2025)
by: Zhang, Bowen, et al.
Published: (2025)
Optimal Stepsize for Diffusion Sampling
by: Pei, Jianning, et al.
Published: (2025)
by: Pei, Jianning, et al.
Published: (2025)
Think-Then-Generate: Reasoning-Aware Text-to-Image Diffusion with LLM Encoders
by: Kou, Siqi, et al.
Published: (2026)
by: Kou, Siqi, et al.
Published: (2026)
Harnessing Caption Detailness for Data-Efficient Text-to-Image Generation
by: Wang, Xinran, et al.
Published: (2025)
by: Wang, Xinran, et al.
Published: (2025)
Dual3D: Efficient and Consistent Text-to-3D Generation with Dual-mode Multi-view Latent Diffusion
by: Li, Xinyang, et al.
Published: (2024)
by: Li, Xinyang, et al.
Published: (2024)
Novel 3D Binary Indexed Tree for Volume Computation of 3D Reconstructed Models from Volumetric Data
by: Nguyen-Le, Quoc-Bao, et al.
Published: (2024)
by: Nguyen-Le, Quoc-Bao, et al.
Published: (2024)
Scaling Down Text Encoders of Text-to-Image Diffusion Models
by: Wang, Lifu, et al.
Published: (2025)
by: Wang, Lifu, et al.
Published: (2025)
CogView3: Finer and Faster Text-to-Image Generation via Relay Diffusion
by: Zheng, Wendi, et al.
Published: (2024)
by: Zheng, Wendi, et al.
Published: (2024)
3D Face Reconstruction Using A Spectral-Based Graph Convolution Encoder
by: Xu, Haoxin, et al.
Published: (2024)
by: Xu, Haoxin, et al.
Published: (2024)
Enhancing Diffusion Models with Text-Encoder Reinforcement Learning
by: Chen, Chaofeng, et al.
Published: (2023)
by: Chen, Chaofeng, et al.
Published: (2023)
RefAny3D: 3D Asset-Referenced Diffusion Models for Image Generation
by: Huang, Hanzhuo, et al.
Published: (2026)
by: Huang, Hanzhuo, et al.
Published: (2026)
GenerateCT: Text-Conditional Generation of 3D Chest CT Volumes
by: Hamamci, Ibrahim Ethem, et al.
Published: (2023)
by: Hamamci, Ibrahim Ethem, et al.
Published: (2023)
TextBoost: Boosting Text Encoder for Personalized Text-to-Image Generation
by: Park, NaHyeon, et al.
Published: (2024)
by: Park, NaHyeon, et al.
Published: (2024)
Neural Point-based Volumetric Avatar: Surface-guided Neural Points for Efficient and Photorealistic Volumetric Head Avatar
by: Wang, Cong, et al.
Published: (2023)
by: Wang, Cong, et al.
Published: (2023)
3DTopia: Large Text-to-3D Generation Model with Hybrid Diffusion Priors
by: Hong, Fangzhou, et al.
Published: (2024)
by: Hong, Fangzhou, et al.
Published: (2024)
Efficient Part-level 3D Object Generation via Dual Volume Packing
by: Tang, Jiaxiang, et al.
Published: (2025)
by: Tang, Jiaxiang, et al.
Published: (2025)
Volumetric Conditioning Module to Control Pretrained Diffusion Models for 3D Medical Images
by: Ahn, Suhyun, et al.
Published: (2024)
by: Ahn, Suhyun, et al.
Published: (2024)
Text2CT: Towards 3D CT Volume Generation from Free-text Descriptions Using Diffusion Model
by: Guo, Pengfei, et al.
Published: (2025)
by: Guo, Pengfei, et al.
Published: (2025)
Several questions of visual generation in 2024
by: Gu, Shuyang
Published: (2024)
by: Gu, Shuyang
Published: (2024)
Lens: Rethinking Training Efficiency for Foundational Text-to-Image Models
by: Chen, Dong, et al.
Published: (2026)
by: Chen, Dong, et al.
Published: (2026)
Fast Autoregressive Models for Continuous Latent Generation
by: Hang, Tiankai, et al.
Published: (2025)
by: Hang, Tiankai, et al.
Published: (2025)
Learning Coherent Matrixized Representation in Latent Space for Volumetric 4D Generation
by: Yang, Qitong, et al.
Published: (2024)
by: Yang, Qitong, et al.
Published: (2024)
How to Continually Adapt Text-to-Image Diffusion Models for Flexible Customization?
by: Dong, Jiahua, et al.
Published: (2024)
by: Dong, Jiahua, et al.
Published: (2024)
Semantic Image Synthesis via Diffusion Models
by: Zhou, Wengang, et al.
Published: (2022)
by: Zhou, Wengang, et al.
Published: (2022)
Similar Items
-
Diffusion Models without Classifier-free Guidance
by: Tang, Zhicong, et al.
Published: (2025) -
Simplified Diffusion Schrödinger Bridge
by: Tang, Zhicong, et al.
Published: (2024) -
Incorporating Pre-trained Diffusion Models in Solving the Schrödinger Bridge Problem
by: Tang, Zhicong, et al.
Published: (2025) -
Efficient Diffusion Training via Min-SNR Weighting Strategy
by: Hang, Tiankai, et al.
Published: (2023) -
DesignDiffusion: High-Quality Text-to-Design Image Generation with Diffusion Models
by: Wang, Zhendong, et al.
Published: (2025)