COMOGen: A Controllable Text-to-3D Multi-object Generation Framework
Fuente:
arXiv
Saved in:
| Main Authors: | Sun, Shaorong, Pang, Shuchao, Yao, Yazhou, Huang, Xiaoshui |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Rethinking Point Clouds as Sequences: A Causal Next-Token Predictive Learning Framework
by: Yao, Yumeng, et al.
Published: (2026)
by: Yao, Yumeng, et al.
Published: (2026)
Foster Adaptivity and Balance in Learning with Noisy Labels
by: Sheng, Mengmeng, et al.
Published: (2024)
by: Sheng, Mengmeng, et al.
Published: (2024)
GVGEN: Text-to-3D Generation with Volumetric Representation
by: He, Xianglong, et al.
Published: (2024)
by: He, Xianglong, et al.
Published: (2024)
UniDream: Unifying Diffusion Priors for Relightable Text-to-3D Generation
by: Liu, Zexiang, et al.
Published: (2023)
by: Liu, Zexiang, et al.
Published: (2023)
3DBench: A Scalable 3D Benchmark and Instruction-Tuning Dataset
by: Zhang, Junjie, et al.
Published: (2024)
by: Zhang, Junjie, et al.
Published: (2024)
Synergizing Motion and Appearance: Multi-Scale Compensatory Codebooks for Talking Head Video Generation
by: Zhao, Shuling, et al.
Published: (2024)
by: Zhao, Shuling, et al.
Published: (2024)
Learning 3D Representations for Spatial Intelligence from Unposed Multi-View Images
by: Zhou, Bo, et al.
Published: (2026)
by: Zhou, Bo, et al.
Published: (2026)
A Self-Conditioned Representation Guided Diffusion Model for Realistic Text-to-LiDAR Scene Generation
by: Qu, Wentao, et al.
Published: (2025)
by: Qu, Wentao, et al.
Published: (2025)
FTMoMamba: Motion Generation with Frequency and Text State Space Models
by: Li, Chengjian, et al.
Published: (2024)
by: Li, Chengjian, et al.
Published: (2024)
Taming Stable Diffusion for Text to 360° Panorama Image Generation
by: Zhang, Cheng, et al.
Published: (2024)
by: Zhang, Cheng, et al.
Published: (2024)
ESCT3D: Efficient and Selectively Controllable Text-Driven 3D Content Generation with Gaussian Splatting
by: Wu, Huiqi, et al.
Published: (2025)
by: Wu, Huiqi, et al.
Published: (2025)
NeRF-Det++: Incorporating Semantic Cues and Perspective-aware Depth Supervision for Indoor Multi-View 3D Detection
by: Huang, Chenxi, et al.
Published: (2024)
by: Huang, Chenxi, et al.
Published: (2024)
DreamControl: Control-Based Text-to-3D Generation with 3D Self-Prior
by: Huang, Tianyu, et al.
Published: (2023)
by: Huang, Tianyu, et al.
Published: (2023)
THOR: Text to Human-Object Interaction Diffusion via Relation Intervention
by: Wu, Qianyang, et al.
Published: (2024)
by: Wu, Qianyang, et al.
Published: (2024)
LAHNet: Local Attentive Hashing Network for Point Cloud Registration
by: Qu, Wentao, et al.
Published: (2025)
by: Qu, Wentao, et al.
Published: (2025)
CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation
by: Wang, Qinghe, et al.
Published: (2025)
by: Wang, Qinghe, et al.
Published: (2025)
Multimodal Robust Prompt Distillation for 3D Point Cloud Models
by: Gu, Xiang, et al.
Published: (2025)
by: Gu, Xiang, et al.
Published: (2025)
MeshCraft: Exploring Efficient and Controllable Mesh Generation with Flow-based DiTs
by: He, Xianglong, et al.
Published: (2025)
by: He, Xianglong, et al.
Published: (2025)
Robust Single-Stage Fully Sparse 3D Object Detection via Detachable Latent Diffusion
by: Qu, Wentao, et al.
Published: (2025)
by: Qu, Wentao, et al.
Published: (2025)
ForgeDreamer: Industrial Text-to-3D Generation with Multi-Expert LoRA and Cross-View Hypergraph
by: Cai, Junhao, et al.
Published: (2026)
by: Cai, Junhao, et al.
Published: (2026)
OrientDream: Streamlining Text-to-3D Generation with Explicit Orientation Control
by: Huang, Yuzhong, et al.
Published: (2024)
by: Huang, Yuzhong, et al.
Published: (2024)
Direct2.5: Diverse Text-to-3D Generation via Multi-view 2.5D Diffusion
by: Lu, Yuanxun, et al.
Published: (2023)
by: Lu, Yuanxun, et al.
Published: (2023)
A Comprehensive Survey on 3D Content Generation
by: Liu, Jian, et al.
Published: (2024)
by: Liu, Jian, et al.
Published: (2024)
A General Framework to Boost 3D GS Initialization for Text-to-3D Generation by Lexical Richness
by: Jiang, Lutao, et al.
Published: (2024)
by: Jiang, Lutao, et al.
Published: (2024)
UNICA: A Unified Neural Framework for Controllable 3D Avatars
by: Zhu, Jiahe, et al.
Published: (2026)
by: Zhu, Jiahe, et al.
Published: (2026)
BrightDreamer: Generic 3D Gaussian Generative Framework for Fast Text-to-3D Synthesis
by: Jiang, Lutao, et al.
Published: (2024)
by: Jiang, Lutao, et al.
Published: (2024)
InterFusion: Text-Driven Generation of 3D Human-Object Interaction
by: Dai, Sisi, et al.
Published: (2024)
by: Dai, Sisi, et al.
Published: (2024)
X-Oscar: A Progressive Framework for High-quality Text-guided 3D Animatable Avatar Generation
by: Ma, Yiwei, et al.
Published: (2024)
by: Ma, Yiwei, et al.
Published: (2024)
Towards a 3D Transfer-based Black-box Attack via Critical Feature Guidance
by: Pang, Shuchao, et al.
Published: (2025)
by: Pang, Shuchao, et al.
Published: (2025)
RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation
by: Pang, Lexi, et al.
Published: (2025)
by: Pang, Lexi, et al.
Published: (2025)
Sculpt3D: Multi-View Consistent Text-to-3D Generation with Sparse 3D Prior
by: Chen, Cheng, et al.
Published: (2024)
by: Chen, Cheng, et al.
Published: (2024)
TPA3D: Triplane Attention for Fast Text-to-3D Generation
by: Wu, Bin-Shih, et al.
Published: (2023)
by: Wu, Bin-Shih, et al.
Published: (2023)
A Text-to-3D Framework for Joint Generation of CG-Ready Humans and Compatible Garments
by: Sun, Zhiyao, et al.
Published: (2025)
by: Sun, Zhiyao, et al.
Published: (2025)
TextField3D: Towards Enhancing Open-Vocabulary 3D Generation with Noisy Text Fields
by: Huang, Tianyu, et al.
Published: (2023)
by: Huang, Tianyu, et al.
Published: (2023)
Semi-supervised 3D Object Detection with PatchTeacher and PillarMix
by: Wu, Xiaopei, et al.
Published: (2024)
by: Wu, Xiaopei, et al.
Published: (2024)
MAUP: Training-free Multi-center Adaptive Uncertainty-aware Prompting for Cross-domain Few-shot Medical Image Segmentation
by: Zhu, Yazhou, et al.
Published: (2025)
by: Zhu, Yazhou, et al.
Published: (2025)
YouDream: Generating Anatomically Controllable Consistent Text-to-3D Animals
by: Mishra, Sandeep, et al.
Published: (2024)
by: Mishra, Sandeep, et al.
Published: (2024)
Controllable Text-to-3D Generation via Surface-Aligned Gaussian Splatting
by: Li, Zhiqi, et al.
Published: (2024)
by: Li, Zhiqi, et al.
Published: (2024)
If you're waiting for a sign... that might not be it! Mitigating Trust Boundary Confusion from Visual Injections on Vision-Language Agentic Systems
by: Chang, Jiamin, et al.
Published: (2026)
by: Chang, Jiamin, et al.
Published: (2026)
SeMv-3D: Towards Concurrency of Semantic and Multi-view Consistency in General Text-to-3D Generation
by: Cai, Xiao, et al.
Published: (2024)
by: Cai, Xiao, et al.
Published: (2024)
Similar Items
-
Rethinking Point Clouds as Sequences: A Causal Next-Token Predictive Learning Framework
by: Yao, Yumeng, et al.
Published: (2026) -
Foster Adaptivity and Balance in Learning with Noisy Labels
by: Sheng, Mengmeng, et al.
Published: (2024) -
GVGEN: Text-to-3D Generation with Volumetric Representation
by: He, Xianglong, et al.
Published: (2024) -
UniDream: Unifying Diffusion Priors for Relightable Text-to-3D Generation
by: Liu, Zexiang, et al.
Published: (2023) -
3DBench: A Scalable 3D Benchmark and Instruction-Tuning Dataset
by: Zhang, Junjie, et al.
Published: (2024)