Position: Interactive Generative Video as Next-Generation Game Engine
Fuente:
arXiv
Guardado en:
| Autores principales: | Yu, Jiwen, Qin, Yiran, Che, Haoxuan, Liu, Quande, Wang, Xintao, Wan, Pengfei, Zhang, Di, Liu, Xihui |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
GameFactory: Creating New Games with Generative Interactive Videos
por: Yu, Jiwen, et al.
Publicado: (2025)
por: Yu, Jiwen, et al.
Publicado: (2025)
A Survey of Interactive Generative Video
por: Yu, Jiwen, et al.
Publicado: (2025)
por: Yu, Jiwen, et al.
Publicado: (2025)
Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval
por: Yu, Jiwen, et al.
Publicado: (2025)
por: Yu, Jiwen, et al.
Publicado: (2025)
GameGen-X: Interactive Open-world Game Video Generation
por: Che, Haoxuan, et al.
Publicado: (2024)
por: Che, Haoxuan, et al.
Publicado: (2024)
ShotStream: Streaming Multi-Shot Video Generation for Interactive Storytelling
por: Luo, Yawen, et al.
Publicado: (2026)
por: Luo, Yawen, et al.
Publicado: (2026)
UniVideo: Unified Understanding, Generation, and Editing for Videos
por: Wei, Cong, et al.
Publicado: (2025)
por: Wei, Cong, et al.
Publicado: (2025)
FullDiT: Multi-Task Video Generative Foundation Model with Full Attention
por: Ju, Xuan, et al.
Publicado: (2025)
por: Ju, Xuan, et al.
Publicado: (2025)
OmniX: From Unified Panoramic Generation and Perception to Graphics-Ready 3D Scenes
por: Huang, Yukun, et al.
Publicado: (2025)
por: Huang, Yukun, et al.
Publicado: (2025)
CineScene: Implicit 3D as Effective Scene Representation for Cinematic Video Generation
por: Huang, Kaiyi, et al.
Publicado: (2026)
por: Huang, Kaiyi, et al.
Publicado: (2026)
FilMaster: Bridging Cinematic Principles and Generative AI for Automated Film Generation
por: Huang, Kaiyi, et al.
Publicado: (2025)
por: Huang, Kaiyi, et al.
Publicado: (2025)
UniMMVSR: A Unified Multi-Modal Framework for Cascaded Video Super-Resolution
por: Du, Shian, et al.
Publicado: (2025)
por: Du, Shian, et al.
Publicado: (2025)
A Reason-then-Describe Instruction Interpreter for Controllable Video Generation
por: Wu, Shengqiong, et al.
Publicado: (2025)
por: Wu, Shengqiong, et al.
Publicado: (2025)
MultiShotMaster: A Controllable Multi-Shot Video Generation Framework
por: Wang, Qinghe, et al.
Publicado: (2025)
por: Wang, Qinghe, et al.
Publicado: (2025)
ConceptMaster: Multi-Concept Video Customization on Diffusion Transformer Models Without Test-Time Tuning
por: Huang, Yuzhou, et al.
Publicado: (2025)
por: Huang, Yuzhou, et al.
Publicado: (2025)
SketchVideo: Sketch-based Video Generation and Editing
por: Liu, Feng-Lin, et al.
Publicado: (2025)
por: Liu, Feng-Lin, et al.
Publicado: (2025)
UNIC: Unified In-Context Video Editing
por: Ye, Zixuan, et al.
Publicado: (2025)
por: Ye, Zixuan, et al.
Publicado: (2025)
Any2Caption:Interpreting Any Condition to Caption for Controllable Video Generation
por: Wu, Shengqiong, et al.
Publicado: (2025)
por: Wu, Shengqiong, et al.
Publicado: (2025)
StyleMaster: Stylize Your Video with Artistic Generation and Translation
por: Ye, Zixuan, et al.
Publicado: (2024)
por: Ye, Zixuan, et al.
Publicado: (2024)
FullDiT2: Efficient In-Context Conditioning for Video Diffusion Transformers
por: He, Xuanhua, et al.
Publicado: (2025)
por: He, Xuanhua, et al.
Publicado: (2025)
Learning Video Generation for Robotic Manipulation with Collaborative Trajectory Control
por: Fu, Xiao, et al.
Publicado: (2025)
por: Fu, Xiao, et al.
Publicado: (2025)
VideoCanvas: Unified Video Completion from Arbitrary Spatiotemporal Patches via In-Context Conditioning
por: Cai, Minghong, et al.
Publicado: (2025)
por: Cai, Minghong, et al.
Publicado: (2025)
MultiWorld: Scalable Multi-Agent Multi-View Video World Models
por: Wu, Haoyu, et al.
Publicado: (2026)
por: Wu, Haoyu, et al.
Publicado: (2026)
WorldSimBench: Towards Video Generation Models as World Simulators
por: Qin, Yiran, et al.
Publicado: (2024)
por: Qin, Yiran, et al.
Publicado: (2024)
In-Context Audio Control of Video Diffusion Transformers
por: Liu, Wenze, et al.
Publicado: (2025)
por: Liu, Wenze, et al.
Publicado: (2025)
SynCamMaster: Synchronizing Multi-Camera Video Generation from Diverse Viewpoints
por: Bai, Jianhong, et al.
Publicado: (2024)
por: Bai, Jianhong, et al.
Publicado: (2024)
Visual-Aware CoT: Achieving High-Fidelity Visual Consistency in Unified Models
por: Ye, Zixuan, et al.
Publicado: (2025)
por: Ye, Zixuan, et al.
Publicado: (2025)
VideoTetris: Towards Compositional Text-to-Video Generation
por: Tian, Ye, et al.
Publicado: (2024)
por: Tian, Ye, et al.
Publicado: (2024)
3DTrajMaster: Mastering 3D Trajectory for Multi-Entity Motion in Video Generation
por: Fu, Xiao, et al.
Publicado: (2024)
por: Fu, Xiao, et al.
Publicado: (2024)
PatchVSR: Breaking Video Diffusion Resolution Limits with Patch-wise Video Super-Resolution
por: Du, Shian, et al.
Publicado: (2025)
por: Du, Shian, et al.
Publicado: (2025)
CamCloneMaster: Enabling Reference-based Camera Control for Video Generation
por: Luo, Yawen, et al.
Publicado: (2025)
por: Luo, Yawen, et al.
Publicado: (2025)
ReCamMaster: Camera-Controlled Generative Rendering from A Single Video
por: Bai, Jianhong, et al.
Publicado: (2025)
por: Bai, Jianhong, et al.
Publicado: (2025)
Scaling Image and Video Generation via Test-Time Evolutionary Search
por: He, Haoran, et al.
Publicado: (2025)
por: He, Haoran, et al.
Publicado: (2025)
MIDAS: Multimodal Interactive Digital-humAn Synthesis via Real-time Autoregressive Video Generation
por: Chen, Ming, et al.
Publicado: (2025)
por: Chen, Ming, et al.
Publicado: (2025)
SemanticGen: Video Generation in Semantic Space
por: Bai, Jianhong, et al.
Publicado: (2025)
por: Bai, Jianhong, et al.
Publicado: (2025)
Boosting Resolution Generalization of Diffusion Transformers with Randomized Positional Encodings
por: Hou, Liang, et al.
Publicado: (2025)
por: Hou, Liang, et al.
Publicado: (2025)
4Diffusion: Multi-view Video Diffusion Model for 4D Generation
por: Zhang, Haiyu, et al.
Publicado: (2024)
por: Zhang, Haiyu, et al.
Publicado: (2024)
4Dynamic: Text-to-4D Generation with Hybrid Priors
por: Yuan, Yu-Jie, et al.
Publicado: (2024)
por: Yuan, Yu-Jie, et al.
Publicado: (2024)
CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation
por: Wang, Qinghe, et al.
Publicado: (2025)
por: Wang, Qinghe, et al.
Publicado: (2025)
GenMAC: Compositional Text-to-Video Generation with Multi-Agent Collaboration
por: Huang, Kaiyi, et al.
Publicado: (2024)
por: Huang, Kaiyi, et al.
Publicado: (2024)
Text-Animator: Controllable Visual Text Video Generation
por: Liu, Lin, et al.
Publicado: (2024)
por: Liu, Lin, et al.
Publicado: (2024)
Ejemplares similares
-
GameFactory: Creating New Games with Generative Interactive Videos
por: Yu, Jiwen, et al.
Publicado: (2025) -
A Survey of Interactive Generative Video
por: Yu, Jiwen, et al.
Publicado: (2025) -
Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval
por: Yu, Jiwen, et al.
Publicado: (2025) -
GameGen-X: Interactive Open-world Game Video Generation
por: Che, Haoxuan, et al.
Publicado: (2024) -
ShotStream: Streaming Multi-Shot Video Generation for Interactive Storytelling
por: Luo, Yawen, et al.
Publicado: (2026)