WorldCraft: Photo-Realistic 3D World Creation and Customization via LLM Agents
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Xinhang, Tang, Chi-Keung, Tai, Yu-Wing |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Agentic 3D Scene Generation with Spatially Contextualized VLMs
by: Liu, Xinhang, et al.
Published: (2025)
by: Liu, Xinhang, et al.
Published: (2025)
Gear-NeRF: Free-Viewpoint Rendering and Tracking with Motion-aware Spatio-Temporal Sampling
by: Liu, Xinhang, et al.
Published: (2024)
by: Liu, Xinhang, et al.
Published: (2024)
DragVideo: Interactive Drag-style Video Editing
by: Deng, Yufan, et al.
Published: (2023)
by: Deng, Yufan, et al.
Published: (2023)
Multimodal Generation of Animatable 3D Human Models with AvatarForge
by: Liu, Xinhang, et al.
Published: (2025)
by: Liu, Xinhang, et al.
Published: (2025)
PFAvatar: Pose-Fusion 3D Personalized Avatar Reconstruction from Real-World Outfit-of-the-Day Photos
by: Xi, Dianbing, et al.
Published: (2025)
by: Xi, Dianbing, et al.
Published: (2025)
3D-Generalist: Self-Improving Vision-Language-Action Models for Crafting 3D Worlds
by: Sun, Fan-Yun, et al.
Published: (2025)
by: Sun, Fan-Yun, et al.
Published: (2025)
WorldFlow3D: Flowing Through 3D Distributions for Unbounded World Generation
by: Joshi, Amogh, et al.
Published: (2026)
by: Joshi, Amogh, et al.
Published: (2026)
WonderZoom: Multi-Scale 3D World Generation
by: Cao, Jin, et al.
Published: (2025)
by: Cao, Jin, et al.
Published: (2025)
ChatCam: Empowering Camera Control through Conversational AI
by: Liu, Xinhang, et al.
Published: (2024)
by: Liu, Xinhang, et al.
Published: (2024)
WorldScore: A Unified Evaluation Benchmark for World Generation
by: Duan, Haoyi, et al.
Published: (2025)
by: Duan, Haoyi, et al.
Published: (2025)
InceptionHuman: Controllable Prompt-to-NeRF for Photorealistic 3D Human Generation
by: Kao, Shiu-hong, et al.
Published: (2023)
by: Kao, Shiu-hong, et al.
Published: (2023)
WorldGrow: Generating Infinite 3D World
by: Li, Sikuang, et al.
Published: (2025)
by: Li, Sikuang, et al.
Published: (2025)
CNS-Edit: 3D Shape Editing via Coupled Neural Shape Optimization
by: Hu, Jingyu, et al.
Published: (2024)
by: Hu, Jingyu, et al.
Published: (2024)
WorldCraft: From Camera Navigation to Object Manipulation in Interactive Video World Models
by: Gu, Bohai, et al.
Published: (2026)
by: Gu, Bohai, et al.
Published: (2026)
Matrix-3D: Omnidirectional Explorable 3D World Generation
by: Yang, Zhongqi, et al.
Published: (2025)
by: Yang, Zhongqi, et al.
Published: (2025)
InfiniCube: Unbounded and Controllable Dynamic 3D Driving Scene Generation with World-Guided Video Models
by: Lu, Yifan, et al.
Published: (2024)
by: Lu, Yifan, et al.
Published: (2024)
Taming Video Models for 3D and 4D Generation via Zero-Shot Camera Control
by: Song, Chenxi, et al.
Published: (2025)
by: Song, Chenxi, et al.
Published: (2025)
LAGA: Layered 3D Avatar Generation and Customization via Gaussian Splatting
by: Gong, Jia, et al.
Published: (2024)
by: Gong, Jia, et al.
Published: (2024)
DI-PCG: Diffusion-based Efficient Inverse Procedural Content Generation for High-quality 3D Asset Creation
by: Zhao, Wang, et al.
Published: (2024)
by: Zhao, Wang, et al.
Published: (2024)
PEGAsus: 3D Personalization of Geometry and Appearance
by: Hu, Jingyu, et al.
Published: (2026)
by: Hu, Jingyu, et al.
Published: (2026)
VP-LLM: Text-Driven 3D Volume Completion with Large Language Models through Patchification
by: Liu, Jianmeng, et al.
Published: (2024)
by: Liu, Jianmeng, et al.
Published: (2024)
HoloDreamer: Holistic 3D Panoramic World Generation from Text Descriptions
by: Zhou, Haiyang, et al.
Published: (2024)
by: Zhou, Haiyang, et al.
Published: (2024)
WonderWorld: Interactive 3D Scene Generation from a Single Image
by: Yu, Hong-Xing, et al.
Published: (2024)
by: Yu, Hong-Xing, et al.
Published: (2024)
Navigating Motion Agents in Dynamic and Cluttered Environments through LLM Reasoning
by: Zhao, Yubo, et al.
Published: (2025)
by: Zhao, Yubo, et al.
Published: (2025)
In-Context Brush: Zero-shot Customized Subject Insertion with Context-Aware Latent Space Manipulation
by: Xu, Yu, et al.
Published: (2025)
by: Xu, Yu, et al.
Published: (2025)
ArtiLatent: Realistic Articulated 3D Object Generation via Structured Latents
by: Chen, Honghua, et al.
Published: (2025)
by: Chen, Honghua, et al.
Published: (2025)
Stanford-ORB: A Real-World 3D Object Inverse Rendering Benchmark
by: Kuang, Zhengfei, et al.
Published: (2023)
by: Kuang, Zhengfei, et al.
Published: (2023)
CraftsMan3D: High-fidelity Mesh Generation with 3D Native Generation and Interactive Geometry Refiner
by: Li, Weiyu, et al.
Published: (2024)
by: Li, Weiyu, et al.
Published: (2024)
Visionary: The World Model Carrier Built on WebGPU-Powered Gaussian Splatting Platform
by: Gong, Yuning, et al.
Published: (2025)
by: Gong, Yuning, et al.
Published: (2025)
HairGPT: Strand-as-Language Autoregressive Modeling for Realistic 3D Hairstyle Synthesis
by: Luo, Haimin, et al.
Published: (2026)
by: Luo, Haimin, et al.
Published: (2026)
PrimitiveAnything: Human-Crafted 3D Primitive Assembly Generation with Auto-Regressive Transformer
by: Ye, Jingwen, et al.
Published: (2025)
by: Ye, Jingwen, et al.
Published: (2025)
ImagenWorld: Stress-Testing Image Generation Models with Explainable Human Evaluation on Open-ended Real-World Tasks
by: Sani, Samin Mahdizadeh, et al.
Published: (2026)
by: Sani, Samin Mahdizadeh, et al.
Published: (2026)
EgoGrasp: World-Space Hand-Object Interaction Estimation from Egocentric Videos
by: Fu, Hongming, et al.
Published: (2026)
by: Fu, Hongming, et al.
Published: (2026)
SceneForge: Structured World Supervision from 3D Interventions
by: Li, Jizhizi, et al.
Published: (2026)
by: Li, Jizhizi, et al.
Published: (2026)
Physics3D: Learning Physical Properties of 3D Gaussians via Video Diffusion
by: Liu, Fangfu, et al.
Published: (2024)
by: Liu, Fangfu, et al.
Published: (2024)
MatSpray: Fusing 2D Material World Knowledge on 3D Geometry
by: Langsteiner, Philipp, et al.
Published: (2025)
by: Langsteiner, Philipp, et al.
Published: (2025)
ImmerseGen: Agent-Guided Immersive World Generation with Alpha-Textured Proxies
by: Yuan, Jinyan, et al.
Published: (2025)
by: Yuan, Jinyan, et al.
Published: (2025)
Make-A-Shape: a Ten-Million-scale 3D Shape Model
by: Hui, Ka-Hei, et al.
Published: (2024)
by: Hui, Ka-Hei, et al.
Published: (2024)
POSTA: A Go-to Framework for Customized Artistic Poster Generation
by: Chen, Haoyu, et al.
Published: (2025)
by: Chen, Haoyu, et al.
Published: (2025)
RefGaussian: Disentangling Reflections from 3D Gaussian Splatting for Realistic Rendering
by: Zhang, Rui, et al.
Published: (2024)
by: Zhang, Rui, et al.
Published: (2024)
Similar Items
-
Agentic 3D Scene Generation with Spatially Contextualized VLMs
by: Liu, Xinhang, et al.
Published: (2025) -
Gear-NeRF: Free-Viewpoint Rendering and Tracking with Motion-aware Spatio-Temporal Sampling
by: Liu, Xinhang, et al.
Published: (2024) -
DragVideo: Interactive Drag-style Video Editing
by: Deng, Yufan, et al.
Published: (2023) -
Multimodal Generation of Animatable 3D Human Models with AvatarForge
by: Liu, Xinhang, et al.
Published: (2025) -
PFAvatar: Pose-Fusion 3D Personalized Avatar Reconstruction from Real-World Outfit-of-the-Day Photos
by: Xi, Dianbing, et al.
Published: (2025)