GVGEN: Text-to-3D Generation with Volumetric Representation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | He, Xianglong, Chen, Junyi, Peng, Sida, Huang, Di, Li, Yangguang, Huang, Xiaoshui, Yuan, Chun, Ouyang, Wanli, He, Tong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MeshCraft: Exploring Efficient and Controllable Mesh Generation with Flow-based DiTs
von: He, Xianglong, et al.
Veröffentlicht: (2025)
von: He, Xianglong, et al.
Veröffentlicht: (2025)
UniDream: Unifying Diffusion Priors for Relightable Text-to-3D Generation
von: Liu, Zexiang, et al.
Veröffentlicht: (2023)
von: Liu, Zexiang, et al.
Veröffentlicht: (2023)
Where Am I and What Will I See: An Auto-Regressive Model for Spatial Localization and View Prediction
von: Chen, Junyi, et al.
Veröffentlicht: (2024)
von: Chen, Junyi, et al.
Veröffentlicht: (2024)
SparseFlex: High-Resolution and Arbitrary-Topology 3D Shape Modeling
von: He, Xianglong, et al.
Veröffentlicht: (2025)
von: He, Xianglong, et al.
Veröffentlicht: (2025)
ShapeGen: Towards High-Quality 3D Shape Synthesis
von: Li, Yangguang, et al.
Veröffentlicht: (2025)
von: Li, Yangguang, et al.
Veröffentlicht: (2025)
Faithful Contouring: Near-Lossless 3D Voxel Representation Free from Iso-surface
von: Luo, Yihao, et al.
Veröffentlicht: (2025)
von: Luo, Yihao, et al.
Veröffentlicht: (2025)
NeRF-Det++: Incorporating Semantic Cues and Perspective-aware Depth Supervision for Indoor Multi-View 3D Detection
von: Huang, Chenxi, et al.
Veröffentlicht: (2024)
von: Huang, Chenxi, et al.
Veröffentlicht: (2024)
GigaGS: Scaling up Planar-Based 3D Gaussians for Large Scene Surface Reconstruction
von: Chen, Junyi, et al.
Veröffentlicht: (2024)
von: Chen, Junyi, et al.
Veröffentlicht: (2024)
Taming Stable Diffusion for Text to 360° Panorama Image Generation
von: Zhang, Cheng, et al.
Veröffentlicht: (2024)
von: Zhang, Cheng, et al.
Veröffentlicht: (2024)
Agent3D-Zero: An Agent for Zero-shot 3D Understanding
von: Zhang, Sha, et al.
Veröffentlicht: (2024)
von: Zhang, Sha, et al.
Veröffentlicht: (2024)
DATAP-SfM: Dynamic-Aware Tracking Any Point for Robust Structure from Motion in the Wild
von: Ye, Weicai, et al.
Veröffentlicht: (2024)
von: Ye, Weicai, et al.
Veröffentlicht: (2024)
PonderV2: Pave the Way for 3D Foundation Model with A Universal Pre-training Paradigm
von: Zhu, Haoyi, et al.
Veröffentlicht: (2023)
von: Zhu, Haoyi, et al.
Veröffentlicht: (2023)
COMOGen: A Controllable Text-to-3D Multi-object Generation Framework
von: Sun, Shaorong, et al.
Veröffentlicht: (2024)
von: Sun, Shaorong, et al.
Veröffentlicht: (2024)
DiffPano: Scalable and Consistent Text to Panorama Generation with Spherical Epipolar-Aware Diffusion
von: Ye, Weicai, et al.
Veröffentlicht: (2024)
von: Ye, Weicai, et al.
Veröffentlicht: (2024)
Semi-supervised 3D Object Detection with PatchTeacher and PillarMix
von: Wu, Xiaopei, et al.
Veröffentlicht: (2024)
von: Wu, Xiaopei, et al.
Veröffentlicht: (2024)
Uni3D-LLM: Unifying Point Cloud Perception, Generation and Editing with Large Language Models
von: Liu, Dingning, et al.
Veröffentlicht: (2024)
von: Liu, Dingning, et al.
Veröffentlicht: (2024)
A Comprehensive Survey on 3D Content Generation
von: Liu, Jian, et al.
Veröffentlicht: (2024)
von: Liu, Jian, et al.
Veröffentlicht: (2024)
TELA: Text to Layer-wise 3D Clothed Human Generation
von: Dong, Junting, et al.
Veröffentlicht: (2024)
von: Dong, Junting, et al.
Veröffentlicht: (2024)
Holistic-Motion2D: Scalable Whole-body Human Motion Generation in 2D Space
von: Wang, Yuan, et al.
Veröffentlicht: (2024)
von: Wang, Yuan, et al.
Veröffentlicht: (2024)
Adapter-X: A Novel General Parameter-Efficient Fine-Tuning Framework for Vision
von: Li, Minglei, et al.
Veröffentlicht: (2024)
von: Li, Minglei, et al.
Veröffentlicht: (2024)
EMR-Merging: Tuning-Free High-Performance Model Merging
von: Huang, Chenyu, et al.
Veröffentlicht: (2024)
von: Huang, Chenyu, et al.
Veröffentlicht: (2024)
A Self-Conditioned Representation Guided Diffusion Model for Realistic Text-to-LiDAR Scene Generation
von: Qu, Wentao, et al.
Veröffentlicht: (2025)
von: Qu, Wentao, et al.
Veröffentlicht: (2025)
NeuRodin: A Two-stage Framework for High-Fidelity Neural Surface Reconstruction
von: Wang, Yifan, et al.
Veröffentlicht: (2024)
von: Wang, Yifan, et al.
Veröffentlicht: (2024)
TASeg: Temporal Aggregation Network for LiDAR Semantic Segmentation
von: Wu, Xiaopei, et al.
Veröffentlicht: (2024)
von: Wu, Xiaopei, et al.
Veröffentlicht: (2024)
DeepVerse: 4D Autoregressive Video Generation as a World Model
von: Chen, Junyi, et al.
Veröffentlicht: (2025)
von: Chen, Junyi, et al.
Veröffentlicht: (2025)
Human-Centric Foundation Models: Perception, Generation and Agentic Modeling
von: Tang, Shixiang, et al.
Veröffentlicht: (2025)
von: Tang, Shixiang, et al.
Veröffentlicht: (2025)
VolumeDiffusion: Flexible Text-to-3D Generation with Efficient Volumetric Encoder
von: Tang, Zhicong, et al.
Veröffentlicht: (2023)
von: Tang, Zhicong, et al.
Veröffentlicht: (2023)
TAR3D: Creating High-Quality 3D Assets via Next-Part Prediction
von: Zhang, Xuying, et al.
Veröffentlicht: (2024)
von: Zhang, Xuying, et al.
Veröffentlicht: (2024)
Point Cloud Matters: Rethinking the Impact of Different Observation Spaces on Robot Learning
von: Zhu, Haoyi, et al.
Veröffentlicht: (2024)
von: Zhu, Haoyi, et al.
Veröffentlicht: (2024)
EgoAgent: A Joint Predictive Agent Model in Egocentric Worlds
von: Chen, Lu, et al.
Veröffentlicht: (2025)
von: Chen, Lu, et al.
Veröffentlicht: (2025)
3DBench: A Scalable 3D Benchmark and Instruction-Tuning Dataset
von: Zhang, Junjie, et al.
Veröffentlicht: (2024)
von: Zhang, Junjie, et al.
Veröffentlicht: (2024)
ViG3D-UNet: Volumetric Vascular Connectivity-Aware Segmentation via 3D Vision Graph Representation
von: Liu, Bowen, et al.
Veröffentlicht: (2025)
von: Liu, Bowen, et al.
Veröffentlicht: (2025)
Generating Human Motion in 3D Scenes from Text Descriptions
von: Cen, Zhi, et al.
Veröffentlicht: (2024)
von: Cen, Zhi, et al.
Veröffentlicht: (2024)
PredBench: Benchmarking Spatio-Temporal Prediction across Diverse Disciplines
von: Wang, ZiDong, et al.
Veröffentlicht: (2024)
von: Wang, ZiDong, et al.
Veröffentlicht: (2024)
SAM-Med3D: Towards General-purpose Segmentation Models for Volumetric Medical Images
von: Wang, Haoyu, et al.
Veröffentlicht: (2023)
von: Wang, Haoyu, et al.
Veröffentlicht: (2023)
Depth Any Video with Scalable Synthetic Data
von: Yang, Honghui, et al.
Veröffentlicht: (2024)
von: Yang, Honghui, et al.
Veröffentlicht: (2024)
HoloPart: Generative 3D Part Amodal Segmentation
von: Yang, Yunhan, et al.
Veröffentlicht: (2025)
von: Yang, Yunhan, et al.
Veröffentlicht: (2025)
Transition Models: Rethinking the Generative Learning Objective
von: Wang, Zidong, et al.
Veröffentlicht: (2025)
von: Wang, Zidong, et al.
Veröffentlicht: (2025)
LEO-VL: Efficient Scene Representation for Scalable 3D Vision-Language Learning
von: Huang, Jiangyong, et al.
Veröffentlicht: (2025)
von: Huang, Jiangyong, et al.
Veröffentlicht: (2025)
SIC3D: Style Image Conditioned Text-to-3D Gaussian Splatting Generation
von: He, Ming, et al.
Veröffentlicht: (2026)
von: He, Ming, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
MeshCraft: Exploring Efficient and Controllable Mesh Generation with Flow-based DiTs
von: He, Xianglong, et al.
Veröffentlicht: (2025) -
UniDream: Unifying Diffusion Priors for Relightable Text-to-3D Generation
von: Liu, Zexiang, et al.
Veröffentlicht: (2023) -
Where Am I and What Will I See: An Auto-Regressive Model for Spatial Localization and View Prediction
von: Chen, Junyi, et al.
Veröffentlicht: (2024) -
SparseFlex: High-Resolution and Arbitrary-Topology 3D Shape Modeling
von: He, Xianglong, et al.
Veröffentlicht: (2025) -
ShapeGen: Towards High-Quality 3D Shape Synthesis
von: Li, Yangguang, et al.
Veröffentlicht: (2025)