UniMLVG: Unified Framework for Multi-view Long Video Generation with Comprehensive Control Capabilities for Autonomous Driving
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Rui, Wu, Zehuan, Liu, Yichen, Guo, Yuxin, Ni, Jingcheng, Xia, Haifeng, Xia, Siyu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MaskGWM: A Generalizable Driving World Model with Video Mask Reconstruction
von: Ni, Jingcheng, et al.
Veröffentlicht: (2025)
von: Ni, Jingcheng, et al.
Veröffentlicht: (2025)
HoloDrive: Holistic 2D-3D Multi-Modal Street Scene Generation for Autonomous Driving
von: Wu, Zehuan, et al.
Veröffentlicht: (2024)
von: Wu, Zehuan, et al.
Veröffentlicht: (2024)
CVD-STORM: Cross-View Video Diffusion with Spatial-Temporal Reconstruction Model for Autonomous Driving
von: Zhang, Tianrui, et al.
Veröffentlicht: (2025)
von: Zhang, Tianrui, et al.
Veröffentlicht: (2025)
UniMMVSR: A Unified Multi-Modal Framework for Cascaded Video Super-Resolution
von: Du, Shian, et al.
Veröffentlicht: (2025)
von: Du, Shian, et al.
Veröffentlicht: (2025)
UniSTPA: A Safety Analysis Framework for End-to-End Autonomous Driving
von: Kou, Hongrui, et al.
Veröffentlicht: (2025)
von: Kou, Hongrui, et al.
Veröffentlicht: (2025)
MTDrive: Multi-turn Interactive Reinforcement Learning for Autonomous Driving
von: Li, Xidong, et al.
Veröffentlicht: (2026)
von: Li, Xidong, et al.
Veröffentlicht: (2026)
UniDrive-WM: Unified Understanding, Planning and Generation World Model For Autonomous Driving
von: Xiong, Zhexiao, et al.
Veröffentlicht: (2026)
von: Xiong, Zhexiao, et al.
Veröffentlicht: (2026)
UniGeo: Taming Video Diffusion for Unified Consistent Geometry Estimation
von: Sun, Yang-Tian, et al.
Veröffentlicht: (2025)
von: Sun, Yang-Tian, et al.
Veröffentlicht: (2025)
MiLA: Multi-view Intensive-fidelity Long-term Video Generation World Model for Autonomous Driving
von: Wang, Haiguang, et al.
Veröffentlicht: (2025)
von: Wang, Haiguang, et al.
Veröffentlicht: (2025)
UniUGP: Unifying Understanding, Generation, and Planing For End-to-end Autonomous Driving
von: Lu, Hao, et al.
Veröffentlicht: (2025)
von: Lu, Hao, et al.
Veröffentlicht: (2025)
MCTrack: A Unified 3D Multi-Object Tracking Framework for Autonomous Driving
von: Wang, Xiyang, et al.
Veröffentlicht: (2024)
von: Wang, Xiyang, et al.
Veröffentlicht: (2024)
UniDriveVLA: Unifying Understanding, Perception, and Action Planning for Autonomous Driving
von: Li, Yongkang, et al.
Veröffentlicht: (2026)
von: Li, Yongkang, et al.
Veröffentlicht: (2026)
Sketch3D: Style-Consistent Guidance for Sketch-to-3D Generation
von: Zheng, Wangguandong, et al.
Veröffentlicht: (2024)
von: Zheng, Wangguandong, et al.
Veröffentlicht: (2024)
UniCtrl: Improving the Spatiotemporal Consistency of Text-to-Video Diffusion Models via Training-Free Unified Attention Control
von: Xia, Tian, et al.
Veröffentlicht: (2024)
von: Xia, Tian, et al.
Veröffentlicht: (2024)
UniTok: A Unified Tokenizer for Visual Generation and Understanding
von: Ma, Chuofan, et al.
Veröffentlicht: (2025)
von: Ma, Chuofan, et al.
Veröffentlicht: (2025)
Unleashing Generalization of End-to-End Autonomous Driving with Controllable Long Video Generation
von: Ma, Enhui, et al.
Veröffentlicht: (2024)
von: Ma, Enhui, et al.
Veröffentlicht: (2024)
From Parts to Whole: A Unified Reference Framework for Controllable Human Image Generation
von: Huang, Zehuan, et al.
Veröffentlicht: (2024)
von: Huang, Zehuan, et al.
Veröffentlicht: (2024)
UniDebugger: Hierarchical Multi-Agent Framework for Unified Software Debugging
von: Lee, Cheryl, et al.
Veröffentlicht: (2024)
von: Lee, Cheryl, et al.
Veröffentlicht: (2024)
Embedded Representation Learning Network for Animating Styled Video Portrait
von: Wang, Tianyong, et al.
Veröffentlicht: (2024)
von: Wang, Tianyong, et al.
Veröffentlicht: (2024)
UniCorn: A Unified Contrastive Learning Approach for Multi-view Molecular Representation Learning
von: Feng, Shikun, et al.
Veröffentlicht: (2024)
von: Feng, Shikun, et al.
Veröffentlicht: (2024)
UniGen: Unified Modeling of Initial Agent States and Trajectories for Generating Autonomous Driving Scenarios
von: Mahjourian, Reza, et al.
Veröffentlicht: (2024)
von: Mahjourian, Reza, et al.
Veröffentlicht: (2024)
UniScene: Unified Occupancy-centric Driving Scene Generation
von: Li, Bohan, et al.
Veröffentlicht: (2024)
von: Li, Bohan, et al.
Veröffentlicht: (2024)
UniMM-V2X: MoE-Enhanced Multi-Level Fusion for End-to-End Cooperative Autonomous Driving
von: Song, Ziyi, et al.
Veröffentlicht: (2025)
von: Song, Ziyi, et al.
Veröffentlicht: (2025)
MagicDrive-V2: High-Resolution Long Video Generation for Autonomous Driving with Adaptive Control
von: Gao, Ruiyuan, et al.
Veröffentlicht: (2024)
von: Gao, Ruiyuan, et al.
Veröffentlicht: (2024)
UniOcc: A Unified Benchmark for Occupancy Forecasting and Prediction in Autonomous Driving
von: Wang, Yuping, et al.
Veröffentlicht: (2025)
von: Wang, Yuping, et al.
Veröffentlicht: (2025)
UniLION: Towards Unified Autonomous Driving Model with Linear Group RNNs
von: Liu, Zhe, et al.
Veröffentlicht: (2025)
von: Liu, Zhe, et al.
Veröffentlicht: (2025)
MMTL-UniAD: A Unified Framework for Multimodal and Multi-Task Learning in Assistive Driving Perception
von: Liu, Wenzhuo, et al.
Veröffentlicht: (2025)
von: Liu, Wenzhuo, et al.
Veröffentlicht: (2025)
UniVideo: Unified Understanding, Generation, and Editing for Videos
von: Wei, Cong, et al.
Veröffentlicht: (2025)
von: Wei, Cong, et al.
Veröffentlicht: (2025)
UniPhys: Unified Planner and Controller with Diffusion for Flexible Physics-Based Character Control
von: Wu, Yan, et al.
Veröffentlicht: (2025)
von: Wu, Yan, et al.
Veröffentlicht: (2025)
RLGF: Reinforcement Learning with Geometric Feedback for Autonomous Driving Video Generation
von: Yan, Tianyi, et al.
Veröffentlicht: (2025)
von: Yan, Tianyi, et al.
Veröffentlicht: (2025)
UniVBench: Towards Unified Evaluation for Video Foundation Models
von: Wei, Jianhui, et al.
Veröffentlicht: (2026)
von: Wei, Jianhui, et al.
Veröffentlicht: (2026)
DrivingGen: A Comprehensive Benchmark for Generative Video World Models in Autonomous Driving
von: Zhou, Yang, et al.
Veröffentlicht: (2026)
von: Zhou, Yang, et al.
Veröffentlicht: (2026)
UniScene: Multi-Camera Unified Pre-training via 3D Scene Reconstruction for Autonomous Driving
von: Min, Chen, et al.
Veröffentlicht: (2023)
von: Min, Chen, et al.
Veröffentlicht: (2023)
UniM$^2$AE: Multi-modal Masked Autoencoders with Unified 3D Representation for 3D Perception in Autonomous Driving
von: Zou, Jian, et al.
Veröffentlicht: (2023)
von: Zou, Jian, et al.
Veröffentlicht: (2023)
UniPose: A Unified Multimodal Framework for Human Pose Comprehension, Generation and Editing
von: Li, Yiheng, et al.
Veröffentlicht: (2024)
von: Li, Yiheng, et al.
Veröffentlicht: (2024)
UniCP: A Unified Caching and Pruning Framework for Efficient Video Generation
von: Sun, Wenzhang, et al.
Veröffentlicht: (2025)
von: Sun, Wenzhang, et al.
Veröffentlicht: (2025)
UniTalking: A Unified Audio-Video Framework for Talking Portrait Generation
von: Li, Hebeizi, et al.
Veröffentlicht: (2026)
von: Li, Hebeizi, et al.
Veröffentlicht: (2026)
Panacea+: Panoramic and Controllable Video Generation for Autonomous Driving
von: Wen, Yuqing, et al.
Veröffentlicht: (2024)
von: Wen, Yuqing, et al.
Veröffentlicht: (2024)
MV-Adapter: Multi-view Consistent Image Generation Made Easy
von: Huang, Zehuan, et al.
Veröffentlicht: (2024)
von: Huang, Zehuan, et al.
Veröffentlicht: (2024)
Cross-Block Fine-Grained Semantic Cascade for Skeleton-Based Sports Action Recognition
von: Liu, Zhendong, et al.
Veröffentlicht: (2024)
von: Liu, Zhendong, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
MaskGWM: A Generalizable Driving World Model with Video Mask Reconstruction
von: Ni, Jingcheng, et al.
Veröffentlicht: (2025) -
HoloDrive: Holistic 2D-3D Multi-Modal Street Scene Generation for Autonomous Driving
von: Wu, Zehuan, et al.
Veröffentlicht: (2024) -
CVD-STORM: Cross-View Video Diffusion with Spatial-Temporal Reconstruction Model for Autonomous Driving
von: Zhang, Tianrui, et al.
Veröffentlicht: (2025) -
UniMMVSR: A Unified Multi-Modal Framework for Cascaded Video Super-Resolution
von: Du, Shian, et al.
Veröffentlicht: (2025) -
UniSTPA: A Safety Analysis Framework for End-to-End Autonomous Driving
von: Kou, Hongrui, et al.
Veröffentlicht: (2025)