UniLION: Towards Unified Autonomous Driving Model with Linear Group RNNs
Fuente:
arXiv
Salvato in:
| Autori principali: | Liu, Zhe, Hou, Jinghua, Ye, Xiaoqing, Wang, Jingdong, Zhao, Hengshuang, Bai, Xiang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
LION: Linear Group RNN for 3D Object Detection in Point Clouds
di: Liu, Zhe, et al.
Pubblicazione: (2024)
di: Liu, Zhe, et al.
Pubblicazione: (2024)
SEED: A Simple and Effective 3D DETR in Point Clouds
di: Liu, Zhe, et al.
Pubblicazione: (2024)
di: Liu, Zhe, et al.
Pubblicazione: (2024)
DrivePI: Spatial-aware 4D MLLM for Unified Autonomous Driving Understanding, Perception, Prediction and Planning
di: Liu, Zhe, et al.
Pubblicazione: (2025)
di: Liu, Zhe, et al.
Pubblicazione: (2025)
GenieDrive: Towards Physics-Aware Driving World Model with 4D Occupancy Guided Video Generation
di: Yang, Zhenya, et al.
Pubblicazione: (2025)
di: Yang, Zhenya, et al.
Pubblicazione: (2025)
OPEN: Object-wise Position Embedding for Multi-view 3D Object Detection
di: Hou, Jinghua, et al.
Pubblicazione: (2024)
di: Hou, Jinghua, et al.
Pubblicazione: (2024)
HERMES++: Toward a Unified Driving World Model for 3D Scene Understanding and Generation
di: Zhou, Xin, et al.
Pubblicazione: (2026)
di: Zhou, Xin, et al.
Pubblicazione: (2026)
UniDrive-WM: Unified Understanding, Planning and Generation World Model For Autonomous Driving
di: Xiong, Zhexiao, et al.
Pubblicazione: (2026)
di: Xiong, Zhexiao, et al.
Pubblicazione: (2026)
OV-Uni3DETR: Towards Unified Open-Vocabulary 3D Object Detection via Cycle-Modality Propagation
di: Wang, Zhenyu, et al.
Pubblicazione: (2024)
di: Wang, Zhenyu, et al.
Pubblicazione: (2024)
UniDriveVLA: Unifying Understanding, Perception, and Action Planning for Autonomous Driving
di: Li, Yongkang, et al.
Pubblicazione: (2026)
di: Li, Yongkang, et al.
Pubblicazione: (2026)
UniPAD: A Universal Pre-training Paradigm for Autonomous Driving
di: Yang, Honghui, et al.
Pubblicazione: (2023)
di: Yang, Honghui, et al.
Pubblicazione: (2023)
FocalClick-XL: Towards Unified and High-quality Interactive Segmentation
di: Chen, Xi, et al.
Pubblicazione: (2025)
di: Chen, Xi, et al.
Pubblicazione: (2025)
Liquid: Language Models are Scalable and Unified Multi-modal Generators
di: Wu, Junfeng, et al.
Pubblicazione: (2024)
di: Wu, Junfeng, et al.
Pubblicazione: (2024)
HERMES: A Unified Self-Driving World Model for Simultaneous 3D Scene Understanding and Generation
di: Zhou, Xin, et al.
Pubblicazione: (2025)
di: Zhou, Xin, et al.
Pubblicazione: (2025)
UniUGP: Unifying Understanding, Generation, and Planing For End-to-end Autonomous Driving
di: Lu, Hao, et al.
Pubblicazione: (2025)
di: Lu, Hao, et al.
Pubblicazione: (2025)
UniMatch V2: Pushing the Limit of Semi-Supervised Semantic Segmentation
di: Yang, Lihe, et al.
Pubblicazione: (2024)
di: Yang, Lihe, et al.
Pubblicazione: (2024)
UniSER: A Foundation Model for Unified Soft Effects Removal
di: Zhang, Jingdong, et al.
Pubblicazione: (2025)
di: Zhang, Jingdong, et al.
Pubblicazione: (2025)
UniDrive: Towards Universal Driving Perception Across Camera Configurations
di: Li, Ye, et al.
Pubblicazione: (2024)
di: Li, Ye, et al.
Pubblicazione: (2024)
UniDriveDreamer: A Single-Stage Multimodal World Model for Autonomous Driving
di: Zhao, Guosheng, et al.
Pubblicazione: (2026)
di: Zhao, Guosheng, et al.
Pubblicazione: (2026)
Uni$^2$Det: Unified and Universal Framework for Prompt-Guided Multi-dataset 3D Detection
di: Wang, Yubin, et al.
Pubblicazione: (2024)
di: Wang, Yubin, et al.
Pubblicazione: (2024)
FASTER: Rethinking Real-Time Flow VLAs
di: Lu, Yuxiang, et al.
Pubblicazione: (2026)
di: Lu, Yuxiang, et al.
Pubblicazione: (2026)
UniDWM: Towards a Unified Driving World Model via Multifaceted Representation Learning
di: Liu, Shuai, et al.
Pubblicazione: (2026)
di: Liu, Shuai, et al.
Pubblicazione: (2026)
DriveGPT4: Interpretable End-to-end Autonomous Driving via Large Language Model
di: Xu, Zhenhua, et al.
Pubblicazione: (2023)
di: Xu, Zhenhua, et al.
Pubblicazione: (2023)
UniMLVG: Unified Framework for Multi-view Long Video Generation with Comprehensive Control Capabilities for Autonomous Driving
di: Chen, Rui, et al.
Pubblicazione: (2024)
di: Chen, Rui, et al.
Pubblicazione: (2024)
UniScene: Unified Occupancy-centric Driving Scene Generation
di: Li, Bohan, et al.
Pubblicazione: (2024)
di: Li, Bohan, et al.
Pubblicazione: (2024)
Towards Unified 3D Object Detection via Algorithm and Data Unification
di: Li, Zhuoling, et al.
Pubblicazione: (2024)
di: Li, Zhuoling, et al.
Pubblicazione: (2024)
Uni-World VLA: Interleaved World Modeling and Planning for Autonomous Driving
di: Liu, Qiqi, et al.
Pubblicazione: (2026)
di: Liu, Qiqi, et al.
Pubblicazione: (2026)
LION: Implicit Vision Prompt Tuning
di: Wang, Haixin, et al.
Pubblicazione: (2023)
di: Wang, Haixin, et al.
Pubblicazione: (2023)
Make Your ViT-based Multi-view 3D Detectors Faster via Token Compression
di: Zhang, Dingyuan, et al.
Pubblicazione: (2024)
di: Zhang, Dingyuan, et al.
Pubblicazione: (2024)
UniVBench: Towards Unified Evaluation for Video Foundation Models
di: Wei, Jianhui, et al.
Pubblicazione: (2026)
di: Wei, Jianhui, et al.
Pubblicazione: (2026)
UniAlignment: Semantic Alignment for Unified Image Generation, Understanding, Manipulation and Perception
di: Song, Xinyang, et al.
Pubblicazione: (2025)
di: Song, Xinyang, et al.
Pubblicazione: (2025)
UniScene: Multi-Camera Unified Pre-training via 3D Scene Reconstruction for Autonomous Driving
di: Min, Chen, et al.
Pubblicazione: (2023)
di: Min, Chen, et al.
Pubblicazione: (2023)
UniReal: Universal Image Generation and Editing via Learning Real-world Dynamics
di: Chen, Xi, et al.
Pubblicazione: (2024)
di: Chen, Xi, et al.
Pubblicazione: (2024)
UniSTD: Towards Unified Spatio-Temporal Learning across Diverse Disciplines
di: Tang, Chen, et al.
Pubblicazione: (2025)
di: Tang, Chen, et al.
Pubblicazione: (2025)
UniT: Unified Geometry Learning with Group Autoregressive Transformer
di: Wang, Haotian, et al.
Pubblicazione: (2026)
di: Wang, Haotian, et al.
Pubblicazione: (2026)
DriveWorld-VLA: Unified Latent-Space World Modeling with Vision-Language-Action for Autonomous Driving
di: jia, Feiyang, et al.
Pubblicazione: (2026)
di: jia, Feiyang, et al.
Pubblicazione: (2026)
UniHuman: A Unified Model for Editing Human Images in the Wild
di: Li, Nannan, et al.
Pubblicazione: (2023)
di: Li, Nannan, et al.
Pubblicazione: (2023)
Uni-Sign: Toward Unified Sign Language Understanding at Scale
di: Li, Zecheng, et al.
Pubblicazione: (2025)
di: Li, Zecheng, et al.
Pubblicazione: (2025)
HybridTM: Combining Transformer and Mamba for 3D Semantic Segmentation
di: Wang, Xinyu, et al.
Pubblicazione: (2025)
di: Wang, Xinyu, et al.
Pubblicazione: (2025)
PADriver: Towards Personalized Autonomous Driving
di: Kou, Genghua, et al.
Pubblicazione: (2025)
di: Kou, Genghua, et al.
Pubblicazione: (2025)
Uni-Animator: Towards Unified Visual Colorization
di: Chen, Xinyuan, et al.
Pubblicazione: (2026)
di: Chen, Xinyuan, et al.
Pubblicazione: (2026)
Documenti analoghi
-
LION: Linear Group RNN for 3D Object Detection in Point Clouds
di: Liu, Zhe, et al.
Pubblicazione: (2024) -
SEED: A Simple and Effective 3D DETR in Point Clouds
di: Liu, Zhe, et al.
Pubblicazione: (2024) -
DrivePI: Spatial-aware 4D MLLM for Unified Autonomous Driving Understanding, Perception, Prediction and Planning
di: Liu, Zhe, et al.
Pubblicazione: (2025) -
GenieDrive: Towards Physics-Aware Driving World Model with 4D Occupancy Guided Video Generation
di: Yang, Zhenya, et al.
Pubblicazione: (2025) -
OPEN: Object-wise Position Embedding for Multi-view 3D Object Detection
di: Hou, Jinghua, et al.
Pubblicazione: (2024)