OmniGen: Unified Multimodal Sensor Generation for Autonomous Driving
Fuente:
arXiv
Saved in:
| Main Authors: | Tang, Tao, Ma, Enhui, zhou, xia, Wang, Letian, Yan, Tianyi, Zhang, Xueyang, Zhan, Kun, Jia, Peng, Lang, XianPeng, Bian, Jia-Wang, Yu, Kaicheng, Liang, Xiaodan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
OmniGen: Unified Image Generation
by: Xiao, Shitao, et al.
Published: (2024)
by: Xiao, Shitao, et al.
Published: (2024)
OmniGen2: Towards Instruction-Aligned Multimodal Generation
by: Wu, Chenyuan, et al.
Published: (2025)
by: Wu, Chenyuan, et al.
Published: (2025)
CorrectAD: A Self-Correcting Agentic System to Improve End-to-end Planning in Autonomous Driving
by: Ma, Enhui, et al.
Published: (2025)
by: Ma, Enhui, et al.
Published: (2025)
LiSTAR: Ray-Centric World Models for 4D LiDAR Sequences in Autonomous Driving
by: Liu, Pei, et al.
Published: (2025)
by: Liu, Pei, et al.
Published: (2025)
Unleashing Generalization of End-to-End Autonomous Driving with Controllable Long Video Generation
by: Ma, Enhui, et al.
Published: (2024)
by: Ma, Enhui, et al.
Published: (2024)
RoboPearls: Editable Video Simulation for Robot Manipulation
by: Tang, Tao, et al.
Published: (2025)
by: Tang, Tao, et al.
Published: (2025)
DriveCombo: Benchmarking Compositional Traffic Rule Reasoning in Autonomous Driving
by: Ma, Enhui, et al.
Published: (2026)
by: Ma, Enhui, et al.
Published: (2026)
HiNeuS: High-fidelity Neural Surface Mitigating Low-texture and Reflective Ambiguity
by: Wang, Yida, et al.
Published: (2025)
by: Wang, Yida, et al.
Published: (2025)
WorldRFT: Latent World Model Planning with Reinforcement Fine-Tuning for Autonomous Driving
by: Yang, Pengxuan, et al.
Published: (2025)
by: Yang, Pengxuan, et al.
Published: (2025)
Vec-QMDP: Vectorized POMDP Planning on CPUs for Real-Time Autonomous Driving
by: Jin, Xuanjin, et al.
Published: (2026)
by: Jin, Xuanjin, et al.
Published: (2026)
StyledStreets: Multi-style Street Simulator with Spatial and Temporal Consistency
by: Chen, Yuyin, et al.
Published: (2025)
by: Chen, Yuyin, et al.
Published: (2025)
S2-Track: A Simple yet Strong Approach for End-to-End 3D Multi-Object Tracking
by: Tang, Tao, et al.
Published: (2024)
by: Tang, Tao, et al.
Published: (2024)
Generalizing Motion Planners with Mixture of Experts for Autonomous Driving
by: Sun, Qiao, et al.
Published: (2024)
by: Sun, Qiao, et al.
Published: (2024)
DriveVLM: The Convergence of Autonomous Driving and Large Vision-Language Models
by: Tian, Xiaoyu, et al.
Published: (2024)
by: Tian, Xiaoyu, et al.
Published: (2024)
GeoDrive: 3D Geometry-Informed Driving World Model with Precise Action Control
by: Chen, Anthony, et al.
Published: (2025)
by: Chen, Anthony, et al.
Published: (2025)
BEV-TSR: Text-Scene Retrieval in BEV Space for Autonomous Driving
by: Tang, Tao, et al.
Published: (2024)
by: Tang, Tao, et al.
Published: (2024)
RLGF: Reinforcement Learning with Geometric Feedback for Autonomous Driving Video Generation
by: Yan, Tianyi, et al.
Published: (2025)
by: Yan, Tianyi, et al.
Published: (2025)
MagicRoad: Semantic-Aware 3D Road Surface Reconstruction via Obstacle Inpainting
by: Peng, Xingyue, et al.
Published: (2025)
by: Peng, Xingyue, et al.
Published: (2025)
DriveLiDAR4D: Sequential and Controllable LiDAR Scene Generation for Autonomous Driving
by: Cai, Kaiwen, et al.
Published: (2025)
by: Cai, Kaiwen, et al.
Published: (2025)
DrivingGen: A Comprehensive Benchmark for Generative Video World Models in Autonomous Driving
by: Zhou, Yang, et al.
Published: (2026)
by: Zhou, Yang, et al.
Published: (2026)
Unifying Language-Action Understanding and Generation for Autonomous Driving
by: Wang, Xinyang, et al.
Published: (2026)
by: Wang, Xinyang, et al.
Published: (2026)
AudioGen-Omni: A Unified Multimodal Diffusion Transformer for Video-Synchronized Audio, Speech, and Song Generation
by: Wang, Le, et al.
Published: (2025)
by: Wang, Le, et al.
Published: (2025)
MiLA: Multi-view Intensive-fidelity Long-term Video Generation World Model for Autonomous Driving
by: Wang, Haiguang, et al.
Published: (2025)
by: Wang, Haiguang, et al.
Published: (2025)
TransDiffuser: Diverse Trajectory Generation with Decorrelated Multi-modal Representation for End-to-end Autonomous Driving
by: Jiang, Xuefeng, et al.
Published: (2025)
by: Jiang, Xuefeng, et al.
Published: (2025)
ReconDreamer: Crafting World Models for Driving Scene Reconstruction via Online Restoration
by: Ni, Chaojun, et al.
Published: (2024)
by: Ni, Chaojun, et al.
Published: (2024)
Discrete Diffusion for Reflective Vision-Language-Action Models in Autonomous Driving
by: Li, Pengxiang, et al.
Published: (2025)
by: Li, Pengxiang, et al.
Published: (2025)
Unified Sensor Simulation for Autonomous Driving
by: Patakin, Nikolay, et al.
Published: (2026)
by: Patakin, Nikolay, et al.
Published: (2026)
DriveTransformer: Unified Transformer for Scalable End-to-End Autonomous Driving
by: Jia, Xiaosong, et al.
Published: (2025)
by: Jia, Xiaosong, et al.
Published: (2025)
Other Vehicle Trajectories Are Also Needed: A Driving World Model Unifies Ego-Other Vehicle Trajectories in Video Latent Space
by: Zhu, Jian, et al.
Published: (2025)
by: Zhu, Jian, et al.
Published: (2025)
$μ$Drive: User-Controlled Autonomous Driving
by: Wang, Kun, et al.
Published: (2024)
by: Wang, Kun, et al.
Published: (2024)
Protons Selectively Regulate the Asymmetric Activation of the C─H Bonds of Methanol Molecules
by: Kun Jia, et al.
Published: (2025)
by: Kun Jia, et al.
Published: (2025)
OmniHD-Scenes: A Next-Generation Multimodal Dataset for Autonomous Driving
by: Zheng, Lianqing, et al.
Published: (2024)
by: Zheng, Lianqing, et al.
Published: (2024)
CALMM-Drive: Confidence-Aware Autonomous Driving with Large Multimodal Model
by: Yao, Ruoyu, et al.
Published: (2024)
by: Yao, Ruoyu, et al.
Published: (2024)
UltraGen: Extremely Fine-grained Controllable Generation via Attribute Reconstruction and Global Preference Optimization
by: Yun, Longfei, et al.
Published: (2025)
by: Yun, Longfei, et al.
Published: (2025)
The Origin of Order K: Information, Structure and Stability: A Unified Theoretical Framework for Emergence
by: zhou, changzheng, et al.
Published: (2025)
by: zhou, changzheng, et al.
Published: (2025)
DreamOmni: Unified Image Generation and Editing
by: Xia, Bin, et al.
Published: (2024)
by: Xia, Bin, et al.
Published: (2024)
Spectral Selection Rule I: A Unified Framework from Anyon Models to Cosmological Observables
by: zhou, changzheng, et al.
Published: (2026)
by: zhou, changzheng, et al.
Published: (2026)
Do LLM Modules Generalize? A Study on Motion Generation for Autonomous Driving
by: Wang, Mingyi, et al.
Published: (2025)
by: Wang, Mingyi, et al.
Published: (2025)
OmniVLA: Physically-Grounded Multimodal VLA with Unified Multi-Sensor Perception for Robotic Manipulation
by: Guo, Heyu, et al.
Published: (2025)
by: Guo, Heyu, et al.
Published: (2025)
GenAD: Generalized Predictive Model for Autonomous Driving
by: Yang, Jiazhi, et al.
Published: (2024)
by: Yang, Jiazhi, et al.
Published: (2024)
Similar Items
-
OmniGen: Unified Image Generation
by: Xiao, Shitao, et al.
Published: (2024) -
OmniGen2: Towards Instruction-Aligned Multimodal Generation
by: Wu, Chenyuan, et al.
Published: (2025) -
CorrectAD: A Self-Correcting Agentic System to Improve End-to-end Planning in Autonomous Driving
by: Ma, Enhui, et al.
Published: (2025) -
LiSTAR: Ray-Centric World Models for 4D LiDAR Sequences in Autonomous Driving
by: Liu, Pei, et al.
Published: (2025) -
Unleashing Generalization of End-to-End Autonomous Driving with Controllable Long Video Generation
by: Ma, Enhui, et al.
Published: (2024)