LiveWorld: Simulating Out-of-Sight Dynamics in Generative Video World Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Duan, Zicheng, Xia, Jiatong, Zhang, Zeyu, Zhang, Wenbo, Zhou, Gengze, Gou, Chenhui, He, Yefei, Chen, Feng, Zhang, Xinyu, Liu, Lingqiao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Training-Free Motion-Guided Video Generation with Enhanced Temporal Consistency Using Motion Consistency Loss
von: Zhang, Xinyu, et al.
Veröffentlicht: (2025)
von: Zhang, Xinyu, et al.
Veröffentlicht: (2025)
Points-to-3D: Structure-Aware 3D Generation with Point Cloud Priors
von: Xia, Jiatong, et al.
Veröffentlicht: (2026)
von: Xia, Jiatong, et al.
Veröffentlicht: (2026)
EZIGen: Enhancing zero-shot personalized image generation with precise subject encoding and decoupled guidance
von: Duan, Zicheng, et al.
Veröffentlicht: (2024)
von: Duan, Zicheng, et al.
Veröffentlicht: (2024)
Out of Sight but Not Out of Mind: Hybrid Memory for Dynamic Video World Models
von: Chen, Kaijin, et al.
Veröffentlicht: (2026)
von: Chen, Kaijin, et al.
Veröffentlicht: (2026)
Let Your Video Listen to Your Music!
von: Zhang, Xinyu, et al.
Veröffentlicht: (2025)
von: Zhang, Xinyu, et al.
Veröffentlicht: (2025)
Close-up-GS: Enhancing Close-Up View Synthesis in 3D Gaussian Splatting with Progressive Self-Training
von: Xia, Jiatong, et al.
Veröffentlicht: (2025)
von: Xia, Jiatong, et al.
Veröffentlicht: (2025)
Training-Free Instance-Aware 3D Scene Reconstruction and Diffusion-Based View Synthesis from Sparse Images
von: Xia, Jiatong, et al.
Veröffentlicht: (2026)
von: Xia, Jiatong, et al.
Veröffentlicht: (2026)
Out of Sight, Out of Mind? Evaluating State Evolution in Video World Models
von: Ma, Ziqi, et al.
Veröffentlicht: (2026)
von: Ma, Ziqi, et al.
Veröffentlicht: (2026)
World-R1: Reinforcing 3D Constraints for Text-to-Video Generation
von: Wang, Weijie, et al.
Veröffentlicht: (2026)
von: Wang, Weijie, et al.
Veröffentlicht: (2026)
Enhancing Close-up Novel View Synthesis via Pseudo-labeling
von: Xia, Jiatong, et al.
Veröffentlicht: (2025)
von: Xia, Jiatong, et al.
Veröffentlicht: (2025)
You Can Generate It Again: Data-to-Text Generation with Verification and Correction Prompting
von: Ren, Xuan, et al.
Veröffentlicht: (2023)
von: Ren, Xuan, et al.
Veröffentlicht: (2023)
An Empirical Study on How Video-LLMs Answer Video Questions
von: Gou, Chenhui, et al.
Veröffentlicht: (2025)
von: Gou, Chenhui, et al.
Veröffentlicht: (2025)
VQ-VA World: Towards High-Quality Visual Question-Visual Answering
von: Gou, Chenhui, et al.
Veröffentlicht: (2025)
von: Gou, Chenhui, et al.
Veröffentlicht: (2025)
Pre-Trained Video Generative Models as World Simulators
von: He, Haoran, et al.
Veröffentlicht: (2025)
von: He, Haoran, et al.
Veröffentlicht: (2025)
RISE-Video: Can Video Generators Decode Implicit World Rules?
von: Liu, Mingxin, et al.
Veröffentlicht: (2026)
von: Liu, Mingxin, et al.
Veröffentlicht: (2026)
WorldSimBench: Towards Video Generation Models as World Simulators
von: Qin, Yiran, et al.
Veröffentlicht: (2024)
von: Qin, Yiran, et al.
Veröffentlicht: (2024)
LinkedOut: Linking World Knowledge Representation Out of Video LLM for Next-Generation Video Recommendation
von: Zhang, Haichao, et al.
Veröffentlicht: (2025)
von: Zhang, Haichao, et al.
Veröffentlicht: (2025)
Sparsity Forcing: Reinforcing Token Sparsity of MLLMs
von: Chen, Feng, et al.
Veröffentlicht: (2025)
von: Chen, Feng, et al.
Veröffentlicht: (2025)
Code2Worlds: Empowering Coding LLMs for 4D World Generation
von: Zhang, Yi, et al.
Veröffentlicht: (2026)
von: Zhang, Yi, et al.
Veröffentlicht: (2026)
DreamWorld: Unified World Modeling in Video Generation
von: Tan, Boming, et al.
Veröffentlicht: (2026)
von: Tan, Boming, et al.
Veröffentlicht: (2026)
The Best of Both Worlds: On the Dilemma of Out-of-distribution Detection
von: Zhang, Qingyang, et al.
Veröffentlicht: (2024)
von: Zhang, Qingyang, et al.
Veröffentlicht: (2024)
What Do World Models Learn in RL? Probing Latent Representations in Learned Environment Simulators
von: Zhang, Xinyu
Veröffentlicht: (2026)
von: Zhang, Xinyu
Veröffentlicht: (2026)
Teaching Video Generators to Remember: Eliciting Dynamic Memory for Out-of-Sight State Evolution
von: Xu, Tianshuo, et al.
Veröffentlicht: (2026)
von: Xu, Tianshuo, et al.
Veröffentlicht: (2026)
Towards World Simulator: Crafting Physical Commonsense-Based Benchmark for Video Generation
von: Meng, Fanqing, et al.
Veröffentlicht: (2024)
von: Meng, Fanqing, et al.
Veröffentlicht: (2024)
Dynamic Execution Commitment of Vision-Language-Action Models
von: Chen, Feng, et al.
Veröffentlicht: (2026)
von: Chen, Feng, et al.
Veröffentlicht: (2026)
dWorldEval: Scalable Robotic Policy Evaluation via Discrete Diffusion World Model
von: Li, Yaxuan, et al.
Veröffentlicht: (2026)
von: Li, Yaxuan, et al.
Veröffentlicht: (2026)
Rethinking Training Dynamics in Scale-wise Autoregressive Generation
von: Zhou, Gengze, et al.
Veröffentlicht: (2025)
von: Zhou, Gengze, et al.
Veröffentlicht: (2025)
Pediatric Chronic Monteggia Fractures: Insights From a Comprehensive Review
von: Gengze Li, et al.
Veröffentlicht: (2025)
von: Gengze Li, et al.
Veröffentlicht: (2025)
Strong and Controllable Blind Image Decomposition
von: Zhang, Zeyu, et al.
Veröffentlicht: (2024)
von: Zhang, Zeyu, et al.
Veröffentlicht: (2024)
LivingWorld: Interactive 4D World Generation with Environmental Dynamics
von: Mun, Hyeongju, et al.
Veröffentlicht: (2026)
von: Mun, Hyeongju, et al.
Veröffentlicht: (2026)
WISA: World Simulator Assistant for Physics-Aware Text-to-Video Generation
von: Wang, Jing, et al.
Veröffentlicht: (2025)
von: Wang, Jing, et al.
Veröffentlicht: (2025)
GeoWorld: Geometric World Models
von: Zhang, Zeyu, et al.
Veröffentlicht: (2026)
von: Zhang, Zeyu, et al.
Veröffentlicht: (2026)
Spatial Cognition from Egocentric Video: Out of Sight, Not Out of Mind
von: Plizzari, Chiara, et al.
Veröffentlicht: (2024)
von: Plizzari, Chiara, et al.
Veröffentlicht: (2024)
FinSight: Towards Real-World Financial Deep Research
von: Jin, Jiajie, et al.
Veröffentlicht: (2025)
von: Jin, Jiajie, et al.
Veröffentlicht: (2025)
LiveStar: Live Streaming Assistant for Real-World Online Video Understanding
von: Yang, Zhenyu, et al.
Veröffentlicht: (2025)
von: Yang, Zhenyu, et al.
Veröffentlicht: (2025)
VideoVerse: Does Your T2V Generator Have World Model Capability to Synthesize Videos?
von: Wang, Zeqing, et al.
Veröffentlicht: (2025)
von: Wang, Zeqing, et al.
Veröffentlicht: (2025)
Less Detail, Better Answers: Degradation-Driven Prompting for VQA
von: Han, Haoxuan, et al.
Veröffentlicht: (2026)
von: Han, Haoxuan, et al.
Veröffentlicht: (2026)
Neighboring Autoregressive Modeling for Efficient Visual Generation
von: He, Yefei, et al.
Veröffentlicht: (2025)
von: He, Yefei, et al.
Veröffentlicht: (2025)
Driving in Corner Case: A Real-World Adversarial Closed-Loop Evaluation Platform for End-to-End Autonomous Driving
von: Geng, Jiaheng, et al.
Veröffentlicht: (2025)
von: Geng, Jiaheng, et al.
Veröffentlicht: (2025)
UnityVideo: Unified Multi-Modal Multi-Task Learning for Enhancing World-Aware Video Generation
von: Huang, Jiehui, et al.
Veröffentlicht: (2025)
von: Huang, Jiehui, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Training-Free Motion-Guided Video Generation with Enhanced Temporal Consistency Using Motion Consistency Loss
von: Zhang, Xinyu, et al.
Veröffentlicht: (2025) -
Points-to-3D: Structure-Aware 3D Generation with Point Cloud Priors
von: Xia, Jiatong, et al.
Veröffentlicht: (2026) -
EZIGen: Enhancing zero-shot personalized image generation with precise subject encoding and decoupled guidance
von: Duan, Zicheng, et al.
Veröffentlicht: (2024) -
Out of Sight but Not Out of Mind: Hybrid Memory for Dynamic Video World Models
von: Chen, Kaijin, et al.
Veröffentlicht: (2026) -
Let Your Video Listen to Your Music!
von: Zhang, Xinyu, et al.
Veröffentlicht: (2025)