Generalized Dynamics Generation towards Scannable Physical World Model
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Yichen, Li, Zhiyi, Feng, Brandon, Zhang, Dinghuai, Torralba, Antonio |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MultiModal Action Conditioned Video Generation
by: Li, Yichen, et al.
Published: (2025)
by: Li, Yichen, et al.
Published: (2025)
Face2QR: A Unified Framework for Aesthetic, Face-Preserving, and Scannable QR Code Generation
by: Cui, Xuehao, et al.
Published: (2024)
by: Cui, Xuehao, et al.
Published: (2024)
WildWorld: A Large-Scale Dataset for Dynamic World Modeling with Actions and Explicit State toward Generative ARPG
by: Li, Zhen, et al.
Published: (2026)
by: Li, Zhen, et al.
Published: (2026)
Learning to Sample Effective and Diverse Prompts for Text-to-Image Generation
by: Yun, Taeyoung, et al.
Published: (2025)
by: Yun, Taeyoung, et al.
Published: (2025)
LiveWorld: Simulating Out-of-Sight Dynamics in Generative Video World Models
by: Duan, Zicheng, et al.
Published: (2026)
by: Duan, Zicheng, et al.
Published: (2026)
Vision-Language Binding in In-Context Image Generation
by: Ge, Chris, et al.
Published: (2026)
by: Ge, Chris, et al.
Published: (2026)
Dual Diffusion Models for Multi-modal Guided 3D Avatar Generation
by: Li, Hong, et al.
Published: (2026)
by: Li, Hong, et al.
Published: (2026)
UrbanWorld: An Urban World Model for 3D City Generation
by: Shang, Yu, et al.
Published: (2024)
by: Shang, Yu, et al.
Published: (2024)
VideoSketcher: Video Models Prior Enable Versatile Sequential Sketch Generation
by: Ren, Hui, et al.
Published: (2026)
by: Ren, Hui, et al.
Published: (2026)
SimDiff: Simulator-constrained Diffusion Model for Physically Plausible Motion Generation
by: Watanabe, Akihisa, et al.
Published: (2025)
by: Watanabe, Akihisa, et al.
Published: (2025)
STAGE: A Stream-Centric Generative World Model for Long-Horizon Driving-Scene Simulation
by: Wang, Jiamin, et al.
Published: (2025)
by: Wang, Jiamin, et al.
Published: (2025)
IC-World: In-Context Generation for Shared World Modeling
by: Wu, Fan, et al.
Published: (2025)
by: Wu, Fan, et al.
Published: (2025)
Multi-Garment Customized Model Generation
by: Liu, Yichen, et al.
Published: (2024)
by: Liu, Yichen, et al.
Published: (2024)
Fake It till You Make It: Curricular Dynamic Forgery Augmentations towards General Deepfake Detection
by: Lin, Yuzhen, et al.
Published: (2024)
by: Lin, Yuzhen, et al.
Published: (2024)
EyeWorld: A Generative World Model of Ocular State and Dynamics
by: Gao, Ziyu, et al.
Published: (2026)
by: Gao, Ziyu, et al.
Published: (2026)
Dynamic Worlds, Dynamic Humans: Generating Virtual Human-Scene Interaction Motion in Dynamic Scenes
by: Wang, Yin, et al.
Published: (2026)
by: Wang, Yin, et al.
Published: (2026)
Endora: Video Generation Models as Endoscopy Simulators
by: Li, Chenxin, et al.
Published: (2024)
by: Li, Chenxin, et al.
Published: (2024)
DeCoT: Decomposing Complex Instructions for Enhanced Text-to-Image Generation with Large Language Models
by: Lin, Xiaochuan, et al.
Published: (2025)
by: Lin, Xiaochuan, et al.
Published: (2025)
DreamWorld: Unified World Modeling in Video Generation
by: Tan, Boming, et al.
Published: (2026)
by: Tan, Boming, et al.
Published: (2026)
SketchAgent: Language-Driven Sequential Sketch Generation
by: Vinker, Yael, et al.
Published: (2024)
by: Vinker, Yael, et al.
Published: (2024)
PhysFire-WM: A Physics-Informed World Model for Emulating Fire Spread Dynamics
by: Zhou, Nan, et al.
Published: (2025)
by: Zhou, Nan, et al.
Published: (2025)
DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control
by: Wen, Junjie, et al.
Published: (2025)
by: Wen, Junjie, et al.
Published: (2025)
OptiWorld: Optimal Control for Video World Generation under Physical Constraints
by: Yuan, Yu, et al.
Published: (2026)
by: Yuan, Yu, et al.
Published: (2026)
Siamese-DETR for Generic Multi-Object Tracking
by: Liu, Qiankun, et al.
Published: (2023)
by: Liu, Qiankun, et al.
Published: (2023)
UniFuture: A 4D Driving World Model for Future Generation and Perception
by: Liang, Dingkang, et al.
Published: (2025)
by: Liang, Dingkang, et al.
Published: (2025)
Inference-time Physics Alignment of Video Generative Models with Latent World Models
by: Yuan, Jianhao, et al.
Published: (2026)
by: Yuan, Jianhao, et al.
Published: (2026)
Preconditioned Score-based Generative Models
by: Ma, Hengyuan, et al.
Published: (2023)
by: Ma, Hengyuan, et al.
Published: (2023)
Transferable Physical-World Adversarial Patches Against Object Detection in Autonomous Driving
by: Zhu, Zihui, et al.
Published: (2026)
by: Zhu, Zihui, et al.
Published: (2026)
WorldSimBench: Towards Video Generation Models as World Simulators
by: Qin, Yiran, et al.
Published: (2024)
by: Qin, Yiran, et al.
Published: (2024)
Breaking Barriers in Physical-World Adversarial Examples: Improving Robustness and Transferability via Robust Feature
by: Wang, Yichen, et al.
Published: (2024)
by: Wang, Yichen, et al.
Published: (2024)
WISA: World Simulator Assistant for Physics-Aware Text-to-Video Generation
by: Wang, Jing, et al.
Published: (2025)
by: Wang, Jing, et al.
Published: (2025)
DART: Denoising Autoregressive Transformer for Scalable Text-to-Image Generation
by: Gu, Jiatao, et al.
Published: (2024)
by: Gu, Jiatao, et al.
Published: (2024)
Yume-1.5: A Text-Controlled Interactive World Generation Model
by: Mao, Xiaofeng, et al.
Published: (2025)
by: Mao, Xiaofeng, et al.
Published: (2025)
Towards World Simulator: Crafting Physical Commonsense-Based Benchmark for Video Generation
by: Meng, Fanqing, et al.
Published: (2024)
by: Meng, Fanqing, et al.
Published: (2024)
PanoWorld: A Generative Spatial World Model for Consistent Whole-House Panorama Synthesis
by: Jia, Jinrang, et al.
Published: (2026)
by: Jia, Jinrang, et al.
Published: (2026)
Efficient 3D Instance Mapping and Localization with Neural Fields
by: Tang, George, et al.
Published: (2024)
by: Tang, George, et al.
Published: (2024)
LivingWorld: Interactive 4D World Generation with Environmental Dynamics
by: Mun, Hyeongju, et al.
Published: (2026)
by: Mun, Hyeongju, et al.
Published: (2026)
Memory Regulation and Alignment toward Generalizer RGB-Infrared Person
by: Chen, Feng, et al.
Published: (2021)
by: Chen, Feng, et al.
Published: (2021)
Is Sora a World Simulator? A Comprehensive Survey on General World Models and Beyond
by: Zhu, Zheng, et al.
Published: (2024)
by: Zhu, Zheng, et al.
Published: (2024)
Seeing the World through Your Eyes
by: Alzayer, Hadi, et al.
Published: (2023)
by: Alzayer, Hadi, et al.
Published: (2023)
Similar Items
-
MultiModal Action Conditioned Video Generation
by: Li, Yichen, et al.
Published: (2025) -
Face2QR: A Unified Framework for Aesthetic, Face-Preserving, and Scannable QR Code Generation
by: Cui, Xuehao, et al.
Published: (2024) -
WildWorld: A Large-Scale Dataset for Dynamic World Modeling with Actions and Explicit State toward Generative ARPG
by: Li, Zhen, et al.
Published: (2026) -
Learning to Sample Effective and Diverse Prompts for Text-to-Image Generation
by: Yun, Taeyoung, et al.
Published: (2025) -
LiveWorld: Simulating Out-of-Sight Dynamics in Generative Video World Models
by: Duan, Zicheng, et al.
Published: (2026)