DreamControl: Control-Based Text-to-3D Generation with 3D Self-Prior
Fuente:
arXiv
Saved in:
| Main Authors: | Huang, Tianyu, Zeng, Yihan, Zhang, Zhilu, Xu, Wan, Xu, Hang, Xu, Songcen, Lau, Rynson W. H., Zuo, Wangmeng |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TextField3D: Towards Enhancing Open-Vocabulary 3D Generation with Noisy Text Fields
by: Huang, Tianyu, et al.
Published: (2023)
by: Huang, Tianyu, et al.
Published: (2023)
DreamPhysics: Learning Physics-Based 3D Dynamics with Video Diffusion Priors
by: Huang, Tianyu, et al.
Published: (2024)
by: Huang, Tianyu, et al.
Published: (2024)
FILP-3D: Enhancing 3D Few-shot Class-incremental Learning with Pre-trained Vision-Language Models
by: Xu, Wan, et al.
Published: (2023)
by: Xu, Wan, et al.
Published: (2023)
Color3D: Controllable and Consistent 3D Colorization with Personalized Colorizer
by: Wan, Yecong, et al.
Published: (2025)
by: Wan, Yecong, et al.
Published: (2025)
StyleSculptor: Zero-Shot Style-Controllable 3D Asset Generation with Texture-Geometry Dual Guidance
by: Qu, Zefan, et al.
Published: (2025)
by: Qu, Zefan, et al.
Published: (2025)
DreamVideo: High-Fidelity Image-to-Video Generation with Image Retention and Text Guidance
by: Wang, Cong, et al.
Published: (2023)
by: Wang, Cong, et al.
Published: (2023)
DreamControl-v2: Simpler and Scalable Autonomous Humanoid Skills via Trainable Guided Diffusion Priors
by: Harithas, Sudarshan, et al.
Published: (2026)
by: Harithas, Sudarshan, et al.
Published: (2026)
FramePainter: Endowing Interactive Image Editing with Video Diffusion Priors
by: Zhang, Yabo, et al.
Published: (2025)
by: Zhang, Yabo, et al.
Published: (2025)
Rethinking Transformer-Based Blind-Spot Network for Self-Supervised Image Denoising
by: Li, Junyi, et al.
Published: (2024)
by: Li, Junyi, et al.
Published: (2024)
S2AM3D: Scale-controllable Part Segmentation of 3D Point Clouds
by: Su, Han, et al.
Published: (2025)
by: Su, Han, et al.
Published: (2025)
OrientDream: Streamlining Text-to-3D Generation with Explicit Orientation Control
by: Huang, Yuzhong, et al.
Published: (2024)
by: Huang, Yuzhong, et al.
Published: (2024)
Voyager: Long-Range and World-Consistent Video Diffusion for Explorable 3D Scene Generation
by: Huang, Tianyu, et al.
Published: (2025)
by: Huang, Tianyu, et al.
Published: (2025)
SelfHVD: Self-Supervised Handheld Video Deblurring
by: Xu, Honglei, et al.
Published: (2025)
by: Xu, Honglei, et al.
Published: (2025)
DreamControl: Human-Inspired Whole-Body Humanoid Control for Scene Interaction via Guided Diffusion
by: Kalaria, Dvij, et al.
Published: (2025)
by: Kalaria, Dvij, et al.
Published: (2025)
Bridging Geometry-Coherent Text-to-3D Generation with Multi-View Diffusion Priors and Gaussian Splatting
by: Yang, Feng, et al.
Published: (2025)
by: Yang, Feng, et al.
Published: (2025)
GLaVE-Cap: Global-Local Aligned Video Captioning with Vision Expert Integration
by: Xu, Wan, et al.
Published: (2025)
by: Xu, Wan, et al.
Published: (2025)
Phidias: A Generative Model for Creating 3D Content from Text, Image, and 3D Conditions with Reference-Augmented Diffusion
by: Wang, Zhenwei, et al.
Published: (2024)
by: Wang, Zhenwei, et al.
Published: (2024)
NIR-Assisted Image Denoising: A Selective Fusion Approach and A Real-World Benchmark Dataset
by: Xu, Rongjian, et al.
Published: (2024)
by: Xu, Rongjian, et al.
Published: (2024)
Hierarchical Cross-Modal Alignment for Open-Vocabulary 3D Object Detection
by: Zhao, Youjun, et al.
Published: (2025)
by: Zhao, Youjun, et al.
Published: (2025)
JointDreamer: Ensuring Geometry Consistency and Text Congruence in Text-to-3D Generation via Joint Score Distillation
by: Jiang, Chenhan, et al.
Published: (2024)
by: Jiang, Chenhan, et al.
Published: (2024)
CoherenDream: Boosting Holistic Text Coherence in 3D Generation via Multimodal Large Language Models Feedback
by: Jiang, Chenhan, et al.
Published: (2025)
by: Jiang, Chenhan, et al.
Published: (2025)
SeHDR: Single-Exposure HDR Novel View Synthesis via 3D Gaussian Bracketing
by: Li, Yiyu, et al.
Published: (2025)
by: Li, Yiyu, et al.
Published: (2025)
PhysWorld: From Real Videos to World Models of Deformable Objects via Physics-Aware Demonstration Synthesis
by: Yang, Yu, et al.
Published: (2025)
by: Yang, Yu, et al.
Published: (2025)
UniDream: Unifying Diffusion Priors for Relightable Text-to-3D Generation
by: Liu, Zexiang, et al.
Published: (2023)
by: Liu, Zexiang, et al.
Published: (2023)
Collaborative Novel Object Discovery and Box-Guided Cross-Modal Alignment for Open-Vocabulary 3D Object Detection
by: Cao, Yang, et al.
Published: (2024)
by: Cao, Yang, et al.
Published: (2024)
Self-Supervised Learning for Real-World Super-Resolution from Dual and Multiple Zoomed Observations
by: Zhang, Zhilu, et al.
Published: (2024)
by: Zhang, Zhilu, et al.
Published: (2024)
OpenScan: A Benchmark for Generalized Open-Vocabulary 3D Scene Understanding
by: Zhao, Youjun, et al.
Published: (2024)
by: Zhao, Youjun, et al.
Published: (2024)
ThemeStation: Generating Theme-Aware 3D Assets from Few Exemplars
by: Wang, Zhenwei, et al.
Published: (2024)
by: Wang, Zhenwei, et al.
Published: (2024)
Delving into Dark Regions for Robust Shadow Detection
by: Guan, Huankang, et al.
Published: (2024)
by: Guan, Huankang, et al.
Published: (2024)
TexGen: Text-Guided 3D Texture Generation with Multi-view Sampling and Resampling
by: Huo, Dong, et al.
Published: (2024)
by: Huo, Dong, et al.
Published: (2024)
YouDream: Generating Anatomically Controllable Consistent Text-to-3D Animals
by: Mishra, Sandeep, et al.
Published: (2024)
by: Mishra, Sandeep, et al.
Published: (2024)
Enhanced Generative Structure Prior for Chinese Text Image Super-resolution
by: Li, Xiaoming, et al.
Published: (2025)
by: Li, Xiaoming, et al.
Published: (2025)
Deblur4DGS: 4D Gaussian Splatting from Blurry Monocular Video
by: Wu, Renlong, et al.
Published: (2024)
by: Wu, Renlong, et al.
Published: (2024)
PhysChoreo: Physics-Controllable Video Generation with Part-Aware Semantic Grounding
by: Zhang, Haoze, et al.
Published: (2025)
by: Zhang, Haoze, et al.
Published: (2025)
ArtHOI: Taming Foundation Models for Monocular 4D Reconstruction of Hand-Articulated-Object Interactions
by: Wang, Zikai, et al.
Published: (2026)
by: Wang, Zikai, et al.
Published: (2026)
RoamScene3D: Immersive Text-to-3D Scene Generation via Adaptive Object-aware Roaming
by: Chu, Jisheng, et al.
Published: (2026)
by: Chu, Jisheng, et al.
Published: (2026)
Shape-for-Motion: Precise and Consistent Video Editing with 3D Proxy
by: Liu, Yuhao, et al.
Published: (2025)
by: Liu, Yuhao, et al.
Published: (2025)
UniRestorer: Universal Image Restoration via Adaptively Estimating Image Degradation at Proper Granularity
by: Lin, Jingbo, et al.
Published: (2024)
by: Lin, Jingbo, et al.
Published: (2024)
DreamDissector: Learning Disentangled Text-to-3D Generation from 2D Diffusion Priors
by: Yan, Zizheng, et al.
Published: (2024)
by: Yan, Zizheng, et al.
Published: (2024)
S2Gaussian: Sparse-View Super-Resolution 3D Gaussian Splatting
by: Wan, Yecong, et al.
Published: (2025)
by: Wan, Yecong, et al.
Published: (2025)
Similar Items
-
TextField3D: Towards Enhancing Open-Vocabulary 3D Generation with Noisy Text Fields
by: Huang, Tianyu, et al.
Published: (2023) -
DreamPhysics: Learning Physics-Based 3D Dynamics with Video Diffusion Priors
by: Huang, Tianyu, et al.
Published: (2024) -
FILP-3D: Enhancing 3D Few-shot Class-incremental Learning with Pre-trained Vision-Language Models
by: Xu, Wan, et al.
Published: (2023) -
Color3D: Controllable and Consistent 3D Colorization with Personalized Colorizer
by: Wan, Yecong, et al.
Published: (2025) -
StyleSculptor: Zero-Shot Style-Controllable 3D Asset Generation with Texture-Geometry Dual Guidance
by: Qu, Zefan, et al.
Published: (2025)