Agentic 3D Scene Generation with Spatially Contextualized VLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Xinhang, Tai, Yu-Wing, Tang, Chi-Keung |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
WorldCraft: Photo-Realistic 3D World Creation and Customization via LLM Agents
by: Liu, Xinhang, et al.
Published: (2025)
by: Liu, Xinhang, et al.
Published: (2025)
DragVideo: Interactive Drag-style Video Editing
by: Deng, Yufan, et al.
Published: (2023)
by: Deng, Yufan, et al.
Published: (2023)
Multimodal Generation of Animatable 3D Human Models with AvatarForge
by: Liu, Xinhang, et al.
Published: (2025)
by: Liu, Xinhang, et al.
Published: (2025)
Gear-NeRF: Free-Viewpoint Rendering and Tracking with Motion-aware Spatio-Temporal Sampling
by: Liu, Xinhang, et al.
Published: (2024)
by: Liu, Xinhang, et al.
Published: (2024)
InceptionHuman: Controllable Prompt-to-NeRF for Photorealistic 3D Human Generation
by: Kao, Shiu-hong, et al.
Published: (2023)
by: Kao, Shiu-hong, et al.
Published: (2023)
ChatCam: Empowering Camera Control through Conversational AI
by: Liu, Xinhang, et al.
Published: (2024)
by: Liu, Xinhang, et al.
Published: (2024)
Human-Aware 3D Scene Generation with Spatially-constrained Diffusion Models
by: Hong, Xiaolin, et al.
Published: (2024)
by: Hong, Xiaolin, et al.
Published: (2024)
Recent Advances in 3D Object and Scene Generation: A Survey
by: Tang, Xiang, et al.
Published: (2025)
by: Tang, Xiang, et al.
Published: (2025)
ZeroScene: A Zero-Shot Framework for 3D Scene Generation from a Single Image and Controllable Texture Editing
by: Tang, Xiang, et al.
Published: (2025)
by: Tang, Xiang, et al.
Published: (2025)
CNS-Edit: 3D Shape Editing via Coupled Neural Shape Optimization
by: Hu, Jingyu, et al.
Published: (2024)
by: Hu, Jingyu, et al.
Published: (2024)
Towards Geometric and Textural Consistency 3D Scene Generation via Single Image-guided Model Generation and Layout Optimization
by: Tang, Xiang, et al.
Published: (2025)
by: Tang, Xiang, et al.
Published: (2025)
ReelWave: Multi-Agentic Movie Sound Generation through Multimodal LLM Conversation
by: Wang, Zixuan, et al.
Published: (2025)
by: Wang, Zixuan, et al.
Published: (2025)
PEGAsus: 3D Personalization of Geometry and Appearance
by: Hu, Jingyu, et al.
Published: (2026)
by: Hu, Jingyu, et al.
Published: (2026)
Deceptive-NeRF/3DGS: Diffusion-Generated Pseudo-Observations for High-Quality Sparse-View Reconstruction
by: Liu, Xinhang, et al.
Published: (2023)
by: Liu, Xinhang, et al.
Published: (2023)
WonderWorld: Interactive 3D Scene Generation from a Single Image
by: Yu, Hong-Xing, et al.
Published: (2024)
by: Yu, Hong-Xing, et al.
Published: (2024)
VividDream: Generating 3D Scene with Ambient Dynamics
by: Lee, Yao-Chih, et al.
Published: (2024)
by: Lee, Yao-Chih, et al.
Published: (2024)
DreamAnywhere: Object-Centric Panoramic 3D Scene Generation
by: Dominici, Edoardo Alberto, et al.
Published: (2025)
by: Dominici, Edoardo Alberto, et al.
Published: (2025)
Recent Trends in 3D Reconstruction of General Non-Rigid Scenes
by: Yunus, Raza, et al.
Published: (2024)
by: Yunus, Raza, et al.
Published: (2024)
Beyond Inpainting: Unleash 3D Understanding for Precise Camera-Controlled Video Generation
by: Chen, Dong-Yu, et al.
Published: (2026)
by: Chen, Dong-Yu, et al.
Published: (2026)
Make-A-Shape: a Ten-Million-scale 3D Shape Model
by: Hui, Ka-Hei, et al.
Published: (2024)
by: Hui, Ka-Hei, et al.
Published: (2024)
Sketch2Scene: Automatic Generation of Interactive 3D Game Scenes from User's Casual Sketches
by: Xu, Yongzhi, et al.
Published: (2024)
by: Xu, Yongzhi, et al.
Published: (2024)
DreamGaussian4D: Generative 4D Gaussian Splatting
by: Ren, Jiawei, et al.
Published: (2023)
by: Ren, Jiawei, et al.
Published: (2023)
Generating 360° Video is What You Need For a 3D Scene
by: Zhang, Zhaoyang, et al.
Published: (2025)
by: Zhang, Zhaoyang, et al.
Published: (2025)
Diverse 3D Human Pose Generation in Scenes based on Decoupled Structure
by: Dang, Bowen, et al.
Published: (2024)
by: Dang, Bowen, et al.
Published: (2024)
HiScene: Creating Hierarchical 3D Scenes with Isometric View Generation
by: Dong, Wenqi, et al.
Published: (2025)
by: Dong, Wenqi, et al.
Published: (2025)
ZeroHSI: Zero-Shot 4D Human-Scene Interaction by Video Generation
by: Li, Hongjie, et al.
Published: (2024)
by: Li, Hongjie, et al.
Published: (2024)
DragScene: Interactive 3D Scene Editing with Single-view Drag Instructions
by: Gu, Chenghao, et al.
Published: (2024)
by: Gu, Chenghao, et al.
Published: (2024)
Lyra: Generative 3D Scene Reconstruction via Video Diffusion Model Self-Distillation
by: Bahmani, Sherwin, et al.
Published: (2025)
by: Bahmani, Sherwin, et al.
Published: (2025)
An evaluation of SVBRDF Prediction from Generative Image Models for Appearance Modeling of 3D Scenes
by: Gauthier, Alban, et al.
Published: (2025)
by: Gauthier, Alban, et al.
Published: (2025)
SceneSmith: Agentic Generation of Simulation-Ready Indoor Scenes
by: Pfaff, Nicholas, et al.
Published: (2026)
by: Pfaff, Nicholas, et al.
Published: (2026)
Sketch3DVE: Sketch-based 3D-Aware Scene Video Editing
by: Liu, Feng-Lin, et al.
Published: (2025)
by: Liu, Feng-Lin, et al.
Published: (2025)
Neural 3D Strokes: Creating Stylized 3D Scenes with Vectorized 3D Strokes
by: Duan, Hao-Bin, et al.
Published: (2023)
by: Duan, Hao-Bin, et al.
Published: (2023)
SceneEval: Evaluating Semantic Coherence in Text-Conditioned 3D Indoor Scene Synthesis
by: Tam, Hou In Ivan, et al.
Published: (2025)
by: Tam, Hou In Ivan, et al.
Published: (2025)
Articraft: An Agentic System for Scalable Articulated 3D Asset Generation
by: Zhou, Matt, et al.
Published: (2026)
by: Zhou, Matt, et al.
Published: (2026)
Motion-Agent: A Conversational Framework for Human Motion Generation with LLMs
by: Wu, Qi, et al.
Published: (2024)
by: Wu, Qi, et al.
Published: (2024)
Text2NeRF: Text-Driven 3D Scene Generation with Neural Radiance Fields
by: Zhang, Jingbo, et al.
Published: (2023)
by: Zhang, Jingbo, et al.
Published: (2023)
Graph Canvas for Controllable 3D Scene Generation
by: Liu, Libin, et al.
Published: (2024)
by: Liu, Libin, et al.
Published: (2024)
Human Geometry Distribution for 3D Animation Generation
by: Tang, Xiangjun, et al.
Published: (2025)
by: Tang, Xiangjun, et al.
Published: (2025)
Lang3D-XL: Language Embedded 3D Gaussians for Large-scale Scenes
by: Krakovsky, Shai, et al.
Published: (2025)
by: Krakovsky, Shai, et al.
Published: (2025)
CompGS: Efficient 3D Scene Representation via Compressed Gaussian Splatting
by: Liu, Xiangrui, et al.
Published: (2024)
by: Liu, Xiangrui, et al.
Published: (2024)
Similar Items
-
WorldCraft: Photo-Realistic 3D World Creation and Customization via LLM Agents
by: Liu, Xinhang, et al.
Published: (2025) -
DragVideo: Interactive Drag-style Video Editing
by: Deng, Yufan, et al.
Published: (2023) -
Multimodal Generation of Animatable 3D Human Models with AvatarForge
by: Liu, Xinhang, et al.
Published: (2025) -
Gear-NeRF: Free-Viewpoint Rendering and Tracking with Motion-aware Spatio-Temporal Sampling
by: Liu, Xinhang, et al.
Published: (2024) -
InceptionHuman: Controllable Prompt-to-NeRF for Photorealistic 3D Human Generation
by: Kao, Shiu-hong, et al.
Published: (2023)