HiST-VLA: A Hierarchical Spatio-Temporal Vision-Language-Action Model for End-to-End Autonomous Driving
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Yiru, Gu, Zichong, Gao, Yu, Jiang, Anqing, Sun, Zhigang, Wang, Shuo, Heng, Yuwen, Sun, Hao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ETA-VLA: Efficient Token Adaptation via Temporal Fusion and Intra-LLM Sparsification for Vision-Language-Action Models
von: Wang, Yiru, et al.
Veröffentlicht: (2026)
von: Wang, Yiru, et al.
Veröffentlicht: (2026)
DiffVLA++: Bridging Cognitive Reasoning and End-to-End Driving through Metric-Guided Alignment
von: Gao, Yu, et al.
Veröffentlicht: (2025)
von: Gao, Yu, et al.
Veröffentlicht: (2025)
FlowDrive: Energy Flow Field for End-to-End Autonomous Driving
von: Jiang, Hao, et al.
Veröffentlicht: (2025)
von: Jiang, Hao, et al.
Veröffentlicht: (2025)
IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model
von: Jiang, Anqing, et al.
Veröffentlicht: (2025)
von: Jiang, Anqing, et al.
Veröffentlicht: (2025)
UniUncer: Unified Dynamic Static Uncertainty for End to End Driving
von: Gao, Yu, et al.
Veröffentlicht: (2026)
von: Gao, Yu, et al.
Veröffentlicht: (2026)
AnchDrive: Bootstrapping Diffusion Policies with Hybrid Trajectory Anchors for End-to-End Driving
von: Chai, Jinhao, et al.
Veröffentlicht: (2025)
von: Chai, Jinhao, et al.
Veröffentlicht: (2025)
DiffVLA: Vision-Language Guided Diffusion Planning for Autonomous Driving
von: Jiang, Anqing, et al.
Veröffentlicht: (2025)
von: Jiang, Anqing, et al.
Veröffentlicht: (2025)
Do Open-Loop Metrics Predict Closed-Loop Driving? A Cross-Benchmark Correlation Study of NAVSIM and Bench2Drive
von: Wang, Yiru, et al.
Veröffentlicht: (2026)
von: Wang, Yiru, et al.
Veröffentlicht: (2026)
DiffSemanticFusion: Semantic Raster BEV Fusion for Autonomous Driving via Online HD Map Diffusion
von: Sun, Zhigang, et al.
Veröffentlicht: (2025)
von: Sun, Zhigang, et al.
Veröffentlicht: (2025)
TeraSim-World: Worldwide Safety-Critical Data Synthesis for End-to-End Autonomous Driving
von: Wang, Jiawei, et al.
Veröffentlicht: (2025)
von: Wang, Jiawei, et al.
Veröffentlicht: (2025)
PIE: Perception and Interaction Enhanced End-to-End Motion Planning for Autonomous Driving
von: Yuan, Chengran, et al.
Veröffentlicht: (2025)
von: Yuan, Chengran, et al.
Veröffentlicht: (2025)
ST-$π$: Structured SpatioTemporal VLA for Robotic Manipulation
von: Ma, Chuanhao, et al.
Veröffentlicht: (2026)
von: Ma, Chuanhao, et al.
Veröffentlicht: (2026)
LaST$_{0}$: Latent Spatio-Temporal Chain-of-Thought for Robotic Vision-Language-Action Model
von: Liu, Zhuoyang, et al.
Veröffentlicht: (2026)
von: Liu, Zhuoyang, et al.
Veröffentlicht: (2026)
Exploring the Causality of End-to-End Autonomous Driving
von: Li, Jiankun, et al.
Veröffentlicht: (2024)
von: Li, Jiankun, et al.
Veröffentlicht: (2024)
StyleVLA: Driving Style-Aware Vision Language Action Model for Autonomous Driving
von: Gao, Yuan, et al.
Veröffentlicht: (2026)
von: Gao, Yuan, et al.
Veröffentlicht: (2026)
Latent-WAM: Latent World Action Modeling for End-to-End Autonomous Driving
von: Wang, Linbo, et al.
Veröffentlicht: (2026)
von: Wang, Linbo, et al.
Veröffentlicht: (2026)
ComDrive: Comfort-Oriented End-to-End Autonomous Driving
von: Wang, Junming, et al.
Veröffentlicht: (2024)
von: Wang, Junming, et al.
Veröffentlicht: (2024)
VDRive: Leveraging Reinforced VLA and Diffusion Policy for End-to-end Autonomous Driving
von: Guo, Ziang, et al.
Veröffentlicht: (2025)
von: Guo, Ziang, et al.
Veröffentlicht: (2025)
DriveMoE: Mixture-of-Experts for Vision-Language-Action Model in End-to-End Autonomous Driving
von: Yang, Zhenjie, et al.
Veröffentlicht: (2025)
von: Yang, Zhenjie, et al.
Veröffentlicht: (2025)
Mimir: Hierarchical Goal-Driven Diffusion with Uncertainty Propagation for End-to-End Autonomous Driving
von: Xing, Zebin, et al.
Veröffentlicht: (2025)
von: Xing, Zebin, et al.
Veröffentlicht: (2025)
Cognitive-Hierarchy Guided End-to-End Planning for Autonomous Driving
von: Wang, Zhennan, et al.
Veröffentlicht: (2025)
von: Wang, Zhennan, et al.
Veröffentlicht: (2025)
CoReVLA: A Dual-Stage End-to-End Autonomous Driving Framework for Long-Tail Scenarios via Collect-and-Refine
von: Fang, Shiyu, et al.
Veröffentlicht: (2025)
von: Fang, Shiyu, et al.
Veröffentlicht: (2025)
ST4VLA: Spatially Guided Training for Vision-Language-Action Models
von: Ye, Jinhui, et al.
Veröffentlicht: (2026)
von: Ye, Jinhui, et al.
Veröffentlicht: (2026)
AerialVLA: A Vision-Language-Action Model for UAV Navigation via Minimalist End-to-End Control
von: Xu, Peng, et al.
Veröffentlicht: (2026)
von: Xu, Peng, et al.
Veröffentlicht: (2026)
FocalAD: Local Motion Planning for End-to-End Autonomous Driving
von: Sun, Bin, et al.
Veröffentlicht: (2025)
von: Sun, Bin, et al.
Veröffentlicht: (2025)
VADv2: End-to-End Vectorized Autonomous Driving via Probabilistic Planning
von: Jiang, Bo, et al.
Veröffentlicht: (2024)
von: Jiang, Bo, et al.
Veröffentlicht: (2024)
VECTOR-Drive: Tightly Coupled Vision-Language and Trajectory Expert Routing for End-to-End Autonomous Driving
von: Zhao, Rui, et al.
Veröffentlicht: (2026)
von: Zhao, Rui, et al.
Veröffentlicht: (2026)
AutoMoT: A Unified Vision-Language-Action Model with Asynchronous Mixture-of-Transformers for End-to-End Autonomous Driving
von: Huang, Wenhui, et al.
Veröffentlicht: (2026)
von: Huang, Wenhui, et al.
Veröffentlicht: (2026)
DRAMA: An Efficient End-to-end Motion Planner for Autonomous Driving with Mamba
von: Yuan, Chengran, et al.
Veröffentlicht: (2024)
von: Yuan, Chengran, et al.
Veröffentlicht: (2024)
Communication Resources Constrained Hierarchical Federated Learning for End-to-End Autonomous Driving
von: Kou, Wei-Bin, et al.
Veröffentlicht: (2023)
von: Kou, Wei-Bin, et al.
Veröffentlicht: (2023)
Senna: Bridging Large Vision-Language Models and End-to-End Autonomous Driving
von: Jiang, Bo, et al.
Veröffentlicht: (2024)
von: Jiang, Bo, et al.
Veröffentlicht: (2024)
ARTEMIS: Autoregressive End-to-End Trajectory Planning with Mixture of Experts for Autonomous Driving
von: Feng, Renju, et al.
Veröffentlicht: (2025)
von: Feng, Renju, et al.
Veröffentlicht: (2025)
End-To-End Planning of Autonomous Driving in Industry and Academia: 2022-2023
von: Lan, Gongjin, et al.
Veröffentlicht: (2023)
von: Lan, Gongjin, et al.
Veröffentlicht: (2023)
Generalizing End-To-End Autonomous Driving In Real-World Environments Using Zero-Shot LLMs
von: Dong, Zeyu, et al.
Veröffentlicht: (2024)
von: Dong, Zeyu, et al.
Veröffentlicht: (2024)
HAD: Combining Hierarchical Diffusion with Metric-Decoupled RL for End-to-End Driving
von: Yao, Wenhao, et al.
Veröffentlicht: (2026)
von: Yao, Wenhao, et al.
Veröffentlicht: (2026)
ReCogDrive: A Reinforced Cognitive Framework for End-to-End Autonomous Driving
von: Li, Yongkang, et al.
Veröffentlicht: (2025)
von: Li, Yongkang, et al.
Veröffentlicht: (2025)
Decentralized End-to-End Multi-AAV Pursuit Using Predictive Spatio-Temporal Observation via Deep Reinforcement Learning
von: Li, Yude, et al.
Veröffentlicht: (2026)
von: Li, Yude, et al.
Veröffentlicht: (2026)
A Unified Candidate Set with Scene-Adaptive Refinement via Diffusion for End-to-End Autonomous Driving
von: Wu, Zhengfei, et al.
Veröffentlicht: (2026)
von: Wu, Zhengfei, et al.
Veröffentlicht: (2026)
DiffusionDrive: Truncated Diffusion Model for End-to-End Autonomous Driving
von: Liao, Bencheng, et al.
Veröffentlicht: (2024)
von: Liao, Bencheng, et al.
Veröffentlicht: (2024)
Reinforced Refinement with Self-Aware Expansion for End-to-End Autonomous Driving
von: Liu, Haochen, et al.
Veröffentlicht: (2025)
von: Liu, Haochen, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
ETA-VLA: Efficient Token Adaptation via Temporal Fusion and Intra-LLM Sparsification for Vision-Language-Action Models
von: Wang, Yiru, et al.
Veröffentlicht: (2026) -
DiffVLA++: Bridging Cognitive Reasoning and End-to-End Driving through Metric-Guided Alignment
von: Gao, Yu, et al.
Veröffentlicht: (2025) -
FlowDrive: Energy Flow Field for End-to-End Autonomous Driving
von: Jiang, Hao, et al.
Veröffentlicht: (2025) -
IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model
von: Jiang, Anqing, et al.
Veröffentlicht: (2025) -
UniUncer: Unified Dynamic Static Uncertainty for End to End Driving
von: Gao, Yu, et al.
Veröffentlicht: (2026)