Dreamland: Controllable World Creation with Simulator and Generative Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Mo, Sicheng, Leng, Ziyang, Liu, Leon, Wang, Weizhen, He, Honglin, Zhou, Bolei |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SimGen: Simulator-conditioned Driving Scene Generation
di: Zhou, Yunsong, et al.
Pubblicazione: (2024)
di: Zhou, Yunsong, et al.
Pubblicazione: (2024)
BEVCon: Advancing Bird's Eye View Perception with Contrastive Learning
di: Leng, Ziyang, et al.
Pubblicazione: (2025)
di: Leng, Ziyang, et al.
Pubblicazione: (2025)
Occupancy Learning with Spatiotemporal Memory
di: Leng, Ziyang, et al.
Pubblicazione: (2025)
di: Leng, Ziyang, et al.
Pubblicazione: (2025)
Ctrl-X: Controlling Structure and Appearance for Text-To-Image Generation Without Guidance
di: Lin, Kuan Heng, et al.
Pubblicazione: (2024)
di: Lin, Kuan Heng, et al.
Pubblicazione: (2024)
From Seeing to Experiencing: Scaling Navigation Foundation Models with Reinforcement Learning
di: He, Honglin, et al.
Pubblicazione: (2025)
di: He, Honglin, et al.
Pubblicazione: (2025)
Embodied Scene Understanding for Vision Language Models via MetaVQA
di: Wang, Weizhen, et al.
Pubblicazione: (2025)
di: Wang, Weizhen, et al.
Pubblicazione: (2025)
UrbanVerse: Scaling Urban Simulation by Watching City-Tour Videos
di: Liu, Mingxuan, et al.
Pubblicazione: (2025)
di: Liu, Mingxuan, et al.
Pubblicazione: (2025)
A Simple Framework Towards Vision-based Traffic Signal Control with Microscopic Simulation
di: He, Pan, et al.
Pubblicazione: (2024)
di: He, Pan, et al.
Pubblicazione: (2024)
Vid2Sim: Realistic and Interactive Simulation from Video for Urban Navigation
di: Xie, Ziyang, et al.
Pubblicazione: (2025)
di: Xie, Ziyang, et al.
Pubblicazione: (2025)
Learning Sidewalk Autopilot from Multi-Scale Imitation with Corrective Behavior Expansion
di: He, Honglin, et al.
Pubblicazione: (2026)
di: He, Honglin, et al.
Pubblicazione: (2026)
MetaUrban: An Embodied AI Simulation Platform for Urban Micromobility
di: Wu, Wayne, et al.
Pubblicazione: (2024)
di: Wu, Wayne, et al.
Pubblicazione: (2024)
Towards Autonomous Micromobility through Scalable Urban Simulation
di: Wu, Wayne, et al.
Pubblicazione: (2025)
di: Wu, Wayne, et al.
Pubblicazione: (2025)
SceneStreamer: Continuous Scenario Generation as Next Token Group Prediction
di: Peng, Zhenghao, et al.
Pubblicazione: (2025)
di: Peng, Zhenghao, et al.
Pubblicazione: (2025)
LiveWorld: Simulating Out-of-Sight Dynamics in Generative Video World Models
di: Duan, Zicheng, et al.
Pubblicazione: (2026)
di: Duan, Zicheng, et al.
Pubblicazione: (2026)
Group Diffusion: Enhancing Image Generation by Unlocking Cross-Sample Collaboration
di: Mo, Sicheng, et al.
Pubblicazione: (2025)
di: Mo, Sicheng, et al.
Pubblicazione: (2025)
Understanding Real-World Traffic Safety through RoadSafe365 Benchmark
di: Liu, Xinyu, et al.
Pubblicazione: (2026)
di: Liu, Xinyu, et al.
Pubblicazione: (2026)
Street-View Image Generation from a Bird's-Eye View Layout
di: Swerdlow, Alexander, et al.
Pubblicazione: (2023)
di: Swerdlow, Alexander, et al.
Pubblicazione: (2023)
Learning to Generate Diverse Pedestrian Movements from Web Videos with Noisy Labels
di: Liu, Zhizheng, et al.
Pubblicazione: (2024)
di: Liu, Zhizheng, et al.
Pubblicazione: (2024)
WorldSimBench: Towards Video Generation Models as World Simulators
di: Qin, Yiran, et al.
Pubblicazione: (2024)
di: Qin, Yiran, et al.
Pubblicazione: (2024)
WISA: World Simulator Assistant for Physics-Aware Text-to-Video Generation
di: Wang, Jing, et al.
Pubblicazione: (2025)
di: Wang, Jing, et al.
Pubblicazione: (2025)
X-Fusion: Introducing New Modality to Frozen Large Language Models
di: Mo, Sicheng, et al.
Pubblicazione: (2025)
di: Mo, Sicheng, et al.
Pubblicazione: (2025)
Medical World Model: Generative Simulation of Tumor Evolution for Treatment Planning
di: Yang, Yijun, et al.
Pubblicazione: (2025)
di: Yang, Yijun, et al.
Pubblicazione: (2025)
SnAG: Scalable and Accurate Video Grounding
di: Mu, Fangzhou, et al.
Pubblicazione: (2024)
di: Mu, Fangzhou, et al.
Pubblicazione: (2024)
HY-World 2.0: A Multi-Modal World Model for Reconstructing, Generating, and Simulating 3D Worlds
di: HY-World, Team, et al.
Pubblicazione: (2026)
di: HY-World, Team, et al.
Pubblicazione: (2026)
OptiWorld: Optimal Control for Video World Generation under Physical Constraints
di: Yuan, Yu, et al.
Pubblicazione: (2026)
di: Yuan, Yu, et al.
Pubblicazione: (2026)
Joint Optimization for 4D Human-Scene Reconstruction in the Wild
di: Liu, Zhizheng, et al.
Pubblicazione: (2025)
di: Liu, Zhizheng, et al.
Pubblicazione: (2025)
Is Sora a World Simulator? A Comprehensive Survey on General World Models and Beyond
di: Zhu, Zheng, et al.
Pubblicazione: (2024)
di: Zhu, Zheng, et al.
Pubblicazione: (2024)
Physical Object Understanding with a Physically Controllable World Model
di: Venkatesh, Rahul, et al.
Pubblicazione: (2026)
di: Venkatesh, Rahul, et al.
Pubblicazione: (2026)
Pre-Trained Video Generative Models as World Simulators
di: He, Haoran, et al.
Pubblicazione: (2025)
di: He, Haoran, et al.
Pubblicazione: (2025)
Extended Short- and Long-Range Mesh Learning for Fast and Generalized Garment Simulation
di: Liu, Aoran, et al.
Pubblicazione: (2025)
di: Liu, Aoran, et al.
Pubblicazione: (2025)
SciFlow-Bench: Evaluating Structure-Aware Scientific Diagram Generation via Inverse Parsing
di: Zhang, Tong, et al.
Pubblicazione: (2026)
di: Zhang, Tong, et al.
Pubblicazione: (2026)
DeepVerse: 4D Autoregressive Video Generation as a World Model
di: Chen, Junyi, et al.
Pubblicazione: (2025)
di: Chen, Junyi, et al.
Pubblicazione: (2025)
DreamGaussian: Generative Gaussian Splatting for Efficient 3D Content Creation
di: Tang, Jiaxiang, et al.
Pubblicazione: (2023)
di: Tang, Jiaxiang, et al.
Pubblicazione: (2023)
LidarDM: Generative LiDAR Simulation in a Generated World
di: Zyrianov, Vlas, et al.
Pubblicazione: (2024)
di: Zyrianov, Vlas, et al.
Pubblicazione: (2024)
Dynamic Worlds, Dynamic Humans: Generating Virtual Human-Scene Interaction Motion in Dynamic Scenes
di: Wang, Yin, et al.
Pubblicazione: (2026)
di: Wang, Yin, et al.
Pubblicazione: (2026)
EyeWorld: A Generative World Model of Ocular State and Dynamics
di: Gao, Ziyu, et al.
Pubblicazione: (2026)
di: Gao, Ziyu, et al.
Pubblicazione: (2026)
HiCo: Hierarchical Controllable Diffusion Model for Layout-to-image Generation
di: Cheng, Bo, et al.
Pubblicazione: (2024)
di: Cheng, Bo, et al.
Pubblicazione: (2024)
LIVE: Long-horizon Interactive Video World Modeling
di: Huang, Junchao, et al.
Pubblicazione: (2026)
di: Huang, Junchao, et al.
Pubblicazione: (2026)
Preference Score Distillation: Leveraging 2D Rewards to Align Text-to-3D Generation with Human Preference
di: Leng, Jiaqi, et al.
Pubblicazione: (2026)
di: Leng, Jiaqi, et al.
Pubblicazione: (2026)
Qwen-Image-Bench: From Generation to Creation in Text-to-Image Evaluation
di: Li, Niantong, et al.
Pubblicazione: (2026)
di: Li, Niantong, et al.
Pubblicazione: (2026)
Documenti analoghi
-
SimGen: Simulator-conditioned Driving Scene Generation
di: Zhou, Yunsong, et al.
Pubblicazione: (2024) -
BEVCon: Advancing Bird's Eye View Perception with Contrastive Learning
di: Leng, Ziyang, et al.
Pubblicazione: (2025) -
Occupancy Learning with Spatiotemporal Memory
di: Leng, Ziyang, et al.
Pubblicazione: (2025) -
Ctrl-X: Controlling Structure and Appearance for Text-To-Image Generation Without Guidance
di: Lin, Kuan Heng, et al.
Pubblicazione: (2024) -
From Seeing to Experiencing: Scaling Navigation Foundation Models with Reinforcement Learning
di: He, Honglin, et al.
Pubblicazione: (2025)