HOLODECK 2.0: Vision-Language-Guided 3D World Generation with Editing
Fuente:
arXiv
Saved in:
| Main Authors: | Bian, Zixuan, Ren, Ruohan, Yang, Yue, Callison-Burch, Chris |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
BlenderAlchemy: Editing 3D Graphics with Vision-Language Models
by: Huang, Ian, et al.
Published: (2024)
by: Huang, Ian, et al.
Published: (2024)
3D-Generalist: Self-Improving Vision-Language-Action Models for Crafting 3D Worlds
by: Sun, Fan-Yun, et al.
Published: (2025)
by: Sun, Fan-Yun, et al.
Published: (2025)
WorldGrow: Generating Infinite 3D World
by: Li, Sikuang, et al.
Published: (2025)
by: Li, Sikuang, et al.
Published: (2025)
Improved 3D Scene Stylization via Text-Guided Generative Image Editing with Region-Based Control
by: Fujiwara, Haruo, et al.
Published: (2025)
by: Fujiwara, Haruo, et al.
Published: (2025)
InfiniCube: Unbounded and Controllable Dynamic 3D Driving Scene Generation with World-Guided Video Models
by: Lu, Yifan, et al.
Published: (2024)
by: Lu, Yifan, et al.
Published: (2024)
Garment Particles: A 2D--3D Symmetric Garment Representation for Generation and Editing
by: Nakayama, Kiyohiro, et al.
Published: (2026)
by: Nakayama, Kiyohiro, et al.
Published: (2026)
Matrix-3D: Omnidirectional Explorable 3D World Generation
by: Yang, Zhongqi, et al.
Published: (2025)
by: Yang, Zhongqi, et al.
Published: (2025)
ReplaceAnything3D:Text-Guided 3D Scene Editing with Compositional Neural Radiance Fields
by: Bartrum, Edward, et al.
Published: (2024)
by: Bartrum, Edward, et al.
Published: (2024)
Velocity-Space 3D Asset Editing
by: Liu, Hao, et al.
Published: (2026)
by: Liu, Hao, et al.
Published: (2026)
HPR3D: Hierarchical Proxy Representation for High-Fidelity 3D Reconstruction and Controllable Editing
by: Wang, Tielong, et al.
Published: (2025)
by: Wang, Tielong, et al.
Published: (2025)
Native 3D Editing with Full Attention
by: Cai, Weiwei, et al.
Published: (2025)
by: Cai, Weiwei, et al.
Published: (2025)
BlenderFusion: 3D-Grounded Visual Editing and Generative Compositing
by: Chen, Jiacheng, et al.
Published: (2025)
by: Chen, Jiacheng, et al.
Published: (2025)
3DGS$^2$: Near Second-order Converging 3D Gaussian Splatting
by: Lan, Lei, et al.
Published: (2025)
by: Lan, Lei, et al.
Published: (2025)
CoMo: Controllable Motion Generation through Language Guided Pose Code Editing
by: Huang, Yiming, et al.
Published: (2024)
by: Huang, Yiming, et al.
Published: (2024)
3DEditSafe: Defending 3D Editing Pipelines from Unsafe Generation
by: Meng, Nicole, et al.
Published: (2026)
by: Meng, Nicole, et al.
Published: (2026)
Generic 3D Diffusion Adapter Using Controlled Multi-View Editing
by: Chen, Hansheng, et al.
Published: (2024)
by: Chen, Hansheng, et al.
Published: (2024)
Coin3D: Controllable and Interactive 3D Assets Generation with Proxy-Guided Conditioning
by: Dong, Wenqi, et al.
Published: (2024)
by: Dong, Wenqi, et al.
Published: (2024)
C3Editor: Achieving Controllable Consistency in 2D Model for 3D Editing
by: Tao, Zeng, et al.
Published: (2025)
by: Tao, Zeng, et al.
Published: (2025)
CTRL-D: Controllable Dynamic 3D Scene Editing with Personalized 2D Diffusion
by: He, Kai, et al.
Published: (2024)
by: He, Kai, et al.
Published: (2024)
GEN3C: 3D-Informed World-Consistent Video Generation with Precise Camera Control
by: Ren, Xuanchi, et al.
Published: (2025)
by: Ren, Xuanchi, et al.
Published: (2025)
AvatarMMC: 3D Head Avatar Generation and Editing with Multi-Modal Conditioning
by: Para, Wamiq Reyaz, et al.
Published: (2024)
by: Para, Wamiq Reyaz, et al.
Published: (2024)
CRAFT: Designing Creative and Functional 3D Objects
by: Guo, Michelle, et al.
Published: (2024)
by: Guo, Michelle, et al.
Published: (2024)
ProGDF: Progressive Gaussian Differential Field for Controllable and Flexible 3D Editing
by: Zhao, Yian, et al.
Published: (2024)
by: Zhao, Yian, et al.
Published: (2024)
SF3D: Stable Fast 3D Mesh Reconstruction with UV-unwrapping and Illumination Disentanglement
by: Boss, Mark, et al.
Published: (2024)
by: Boss, Mark, et al.
Published: (2024)
Sketch3DVE: Sketch-based 3D-Aware Scene Video Editing
by: Liu, Feng-Lin, et al.
Published: (2025)
by: Liu, Feng-Lin, et al.
Published: (2025)
PrEditor3D: Fast and Precise 3D Shape Editing
by: Erkoç, Ziya, et al.
Published: (2024)
by: Erkoç, Ziya, et al.
Published: (2024)
Proc3D: Procedural 3D Generation and Parametric Editing of 3D Shapes with Large Language Models
by: Raji, Fadlullah, et al.
Published: (2026)
by: Raji, Fadlullah, et al.
Published: (2026)
Taking Language Embedded 3D Gaussian Splatting into the Wild
by: Wang, Yuze, et al.
Published: (2025)
by: Wang, Yuze, et al.
Published: (2025)
View-Consistent 3D Editing with Gaussian Splatting
by: Wang, Yuxuan, et al.
Published: (2024)
by: Wang, Yuxuan, et al.
Published: (2024)
ImmerseGen: Agent-Guided Immersive World Generation with Alpha-Textured Proxies
by: Yuan, Jinyan, et al.
Published: (2025)
by: Yuan, Jinyan, et al.
Published: (2025)
3D-LATTE: Latent Space 3D Editing from Textual Instructions
by: Parelli, Maria, et al.
Published: (2025)
by: Parelli, Maria, et al.
Published: (2025)
Handle-based Mesh Deformation Guided By Vision Language Model
by: Sun, Xingpeng, et al.
Published: (2025)
by: Sun, Xingpeng, et al.
Published: (2025)
ShapeUP: Scalable Image-Conditioned 3D Editing
by: Gat, Inbar, et al.
Published: (2026)
by: Gat, Inbar, et al.
Published: (2026)
WordRobe: Text-Guided Generation of Textured 3D Garments
by: Srivastava, Astitva, et al.
Published: (2024)
by: Srivastava, Astitva, et al.
Published: (2024)
SPAR3D: Stable Point-Aware Reconstruction of 3D Objects from Single Images
by: Huang, Zixuan, et al.
Published: (2025)
by: Huang, Zixuan, et al.
Published: (2025)
Instant3dit: Multiview Inpainting for Fast Editing of 3D Objects
by: Barda, Amir, et al.
Published: (2024)
by: Barda, Amir, et al.
Published: (2024)
Image Sculpting: Precise Object Editing with 3D Geometry Control
by: Yenphraphai, Jiraphon, et al.
Published: (2024)
by: Yenphraphai, Jiraphon, et al.
Published: (2024)
MotionFix: Text-Driven 3D Human Motion Editing
by: Athanasiou, Nikos, et al.
Published: (2024)
by: Athanasiou, Nikos, et al.
Published: (2024)
GaussianEditor: Editing 3D Gaussians Delicately with Text Instructions
by: Wang, Junjie, et al.
Published: (2023)
by: Wang, Junjie, et al.
Published: (2023)
StyleTex: Style Image-Guided Texture Generation for 3D Models
by: Xie, Zhiyu, et al.
Published: (2024)
by: Xie, Zhiyu, et al.
Published: (2024)
Similar Items
-
BlenderAlchemy: Editing 3D Graphics with Vision-Language Models
by: Huang, Ian, et al.
Published: (2024) -
3D-Generalist: Self-Improving Vision-Language-Action Models for Crafting 3D Worlds
by: Sun, Fan-Yun, et al.
Published: (2025) -
WorldGrow: Generating Infinite 3D World
by: Li, Sikuang, et al.
Published: (2025) -
Improved 3D Scene Stylization via Text-Guided Generative Image Editing with Region-Based Control
by: Fujiwara, Haruo, et al.
Published: (2025) -
InfiniCube: Unbounded and Controllable Dynamic 3D Driving Scene Generation with World-Guided Video Models
by: Lu, Yifan, et al.
Published: (2024)