Dreamland: Controllable World Creation with Simulator and Generative Models
Fuente:
arXiv
Saved in:
| Main Authors: | Mo, Sicheng, Leng, Ziyang, Liu, Leon, Wang, Weizhen, He, Honglin, Zhou, Bolei |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SimGen: Simulator-conditioned Driving Scene Generation
by: Zhou, Yunsong, et al.
Published: (2024)
by: Zhou, Yunsong, et al.
Published: (2024)
BEVCon: Advancing Bird's Eye View Perception with Contrastive Learning
by: Leng, Ziyang, et al.
Published: (2025)
by: Leng, Ziyang, et al.
Published: (2025)
Occupancy Learning with Spatiotemporal Memory
by: Leng, Ziyang, et al.
Published: (2025)
by: Leng, Ziyang, et al.
Published: (2025)
Ctrl-X: Controlling Structure and Appearance for Text-To-Image Generation Without Guidance
by: Lin, Kuan Heng, et al.
Published: (2024)
by: Lin, Kuan Heng, et al.
Published: (2024)
From Seeing to Experiencing: Scaling Navigation Foundation Models with Reinforcement Learning
by: He, Honglin, et al.
Published: (2025)
by: He, Honglin, et al.
Published: (2025)
Embodied Scene Understanding for Vision Language Models via MetaVQA
by: Wang, Weizhen, et al.
Published: (2025)
by: Wang, Weizhen, et al.
Published: (2025)
UrbanVerse: Scaling Urban Simulation by Watching City-Tour Videos
by: Liu, Mingxuan, et al.
Published: (2025)
by: Liu, Mingxuan, et al.
Published: (2025)
A Simple Framework Towards Vision-based Traffic Signal Control with Microscopic Simulation
by: He, Pan, et al.
Published: (2024)
by: He, Pan, et al.
Published: (2024)
Vid2Sim: Realistic and Interactive Simulation from Video for Urban Navigation
by: Xie, Ziyang, et al.
Published: (2025)
by: Xie, Ziyang, et al.
Published: (2025)
Learning Sidewalk Autopilot from Multi-Scale Imitation with Corrective Behavior Expansion
by: He, Honglin, et al.
Published: (2026)
by: He, Honglin, et al.
Published: (2026)
MetaUrban: An Embodied AI Simulation Platform for Urban Micromobility
by: Wu, Wayne, et al.
Published: (2024)
by: Wu, Wayne, et al.
Published: (2024)
Towards Autonomous Micromobility through Scalable Urban Simulation
by: Wu, Wayne, et al.
Published: (2025)
by: Wu, Wayne, et al.
Published: (2025)
SceneStreamer: Continuous Scenario Generation as Next Token Group Prediction
by: Peng, Zhenghao, et al.
Published: (2025)
by: Peng, Zhenghao, et al.
Published: (2025)
LiveWorld: Simulating Out-of-Sight Dynamics in Generative Video World Models
by: Duan, Zicheng, et al.
Published: (2026)
by: Duan, Zicheng, et al.
Published: (2026)
Group Diffusion: Enhancing Image Generation by Unlocking Cross-Sample Collaboration
by: Mo, Sicheng, et al.
Published: (2025)
by: Mo, Sicheng, et al.
Published: (2025)
Understanding Real-World Traffic Safety through RoadSafe365 Benchmark
by: Liu, Xinyu, et al.
Published: (2026)
by: Liu, Xinyu, et al.
Published: (2026)
Street-View Image Generation from a Bird's-Eye View Layout
by: Swerdlow, Alexander, et al.
Published: (2023)
by: Swerdlow, Alexander, et al.
Published: (2023)
Learning to Generate Diverse Pedestrian Movements from Web Videos with Noisy Labels
by: Liu, Zhizheng, et al.
Published: (2024)
by: Liu, Zhizheng, et al.
Published: (2024)
WorldSimBench: Towards Video Generation Models as World Simulators
by: Qin, Yiran, et al.
Published: (2024)
by: Qin, Yiran, et al.
Published: (2024)
WISA: World Simulator Assistant for Physics-Aware Text-to-Video Generation
by: Wang, Jing, et al.
Published: (2025)
by: Wang, Jing, et al.
Published: (2025)
X-Fusion: Introducing New Modality to Frozen Large Language Models
by: Mo, Sicheng, et al.
Published: (2025)
by: Mo, Sicheng, et al.
Published: (2025)
Medical World Model: Generative Simulation of Tumor Evolution for Treatment Planning
by: Yang, Yijun, et al.
Published: (2025)
by: Yang, Yijun, et al.
Published: (2025)
SnAG: Scalable and Accurate Video Grounding
by: Mu, Fangzhou, et al.
Published: (2024)
by: Mu, Fangzhou, et al.
Published: (2024)
HY-World 2.0: A Multi-Modal World Model for Reconstructing, Generating, and Simulating 3D Worlds
by: HY-World, Team, et al.
Published: (2026)
by: HY-World, Team, et al.
Published: (2026)
OptiWorld: Optimal Control for Video World Generation under Physical Constraints
by: Yuan, Yu, et al.
Published: (2026)
by: Yuan, Yu, et al.
Published: (2026)
Joint Optimization for 4D Human-Scene Reconstruction in the Wild
by: Liu, Zhizheng, et al.
Published: (2025)
by: Liu, Zhizheng, et al.
Published: (2025)
Is Sora a World Simulator? A Comprehensive Survey on General World Models and Beyond
by: Zhu, Zheng, et al.
Published: (2024)
by: Zhu, Zheng, et al.
Published: (2024)
Physical Object Understanding with a Physically Controllable World Model
by: Venkatesh, Rahul, et al.
Published: (2026)
by: Venkatesh, Rahul, et al.
Published: (2026)
Pre-Trained Video Generative Models as World Simulators
by: He, Haoran, et al.
Published: (2025)
by: He, Haoran, et al.
Published: (2025)
Extended Short- and Long-Range Mesh Learning for Fast and Generalized Garment Simulation
by: Liu, Aoran, et al.
Published: (2025)
by: Liu, Aoran, et al.
Published: (2025)
SciFlow-Bench: Evaluating Structure-Aware Scientific Diagram Generation via Inverse Parsing
by: Zhang, Tong, et al.
Published: (2026)
by: Zhang, Tong, et al.
Published: (2026)
DeepVerse: 4D Autoregressive Video Generation as a World Model
by: Chen, Junyi, et al.
Published: (2025)
by: Chen, Junyi, et al.
Published: (2025)
DreamGaussian: Generative Gaussian Splatting for Efficient 3D Content Creation
by: Tang, Jiaxiang, et al.
Published: (2023)
by: Tang, Jiaxiang, et al.
Published: (2023)
LidarDM: Generative LiDAR Simulation in a Generated World
by: Zyrianov, Vlas, et al.
Published: (2024)
by: Zyrianov, Vlas, et al.
Published: (2024)
Dynamic Worlds, Dynamic Humans: Generating Virtual Human-Scene Interaction Motion in Dynamic Scenes
by: Wang, Yin, et al.
Published: (2026)
by: Wang, Yin, et al.
Published: (2026)
EyeWorld: A Generative World Model of Ocular State and Dynamics
by: Gao, Ziyu, et al.
Published: (2026)
by: Gao, Ziyu, et al.
Published: (2026)
HiCo: Hierarchical Controllable Diffusion Model for Layout-to-image Generation
by: Cheng, Bo, et al.
Published: (2024)
by: Cheng, Bo, et al.
Published: (2024)
LIVE: Long-horizon Interactive Video World Modeling
by: Huang, Junchao, et al.
Published: (2026)
by: Huang, Junchao, et al.
Published: (2026)
Preference Score Distillation: Leveraging 2D Rewards to Align Text-to-3D Generation with Human Preference
by: Leng, Jiaqi, et al.
Published: (2026)
by: Leng, Jiaqi, et al.
Published: (2026)
Qwen-Image-Bench: From Generation to Creation in Text-to-Image Evaluation
by: Li, Niantong, et al.
Published: (2026)
by: Li, Niantong, et al.
Published: (2026)
Similar Items
-
SimGen: Simulator-conditioned Driving Scene Generation
by: Zhou, Yunsong, et al.
Published: (2024) -
BEVCon: Advancing Bird's Eye View Perception with Contrastive Learning
by: Leng, Ziyang, et al.
Published: (2025) -
Occupancy Learning with Spatiotemporal Memory
by: Leng, Ziyang, et al.
Published: (2025) -
Ctrl-X: Controlling Structure and Appearance for Text-To-Image Generation Without Guidance
by: Lin, Kuan Heng, et al.
Published: (2024) -
From Seeing to Experiencing: Scaling Navigation Foundation Models with Reinforcement Learning
by: He, Honglin, et al.
Published: (2025)