SA-VLA: Spatially-Aware Flow-Matching for Vision-Language-Action Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Pan, Xu, Wan, Zhenglin, Yu, Xingrui, Zheng, Xianwei, Ke, Youkai, Sun, Ming, Wang, Rui, Wang, Ziwei, Tsang, Ivor |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards Backdoor-Based Ownership Verification for Vision-Language-Action Models
von: Sun, Ming, et al.
Veröffentlicht: (2026)
von: Sun, Ming, et al.
Veröffentlicht: (2026)
FM-IRL: Flow-Matching for Reward Modeling and Policy Regularization in Reinforcement Learning
von: Wan, Zhenglin, et al.
Veröffentlicht: (2025)
von: Wan, Zhenglin, et al.
Veröffentlicht: (2025)
PAPO-VLA: Planning-Aware Policy Optimization for Vision-Language-Action Models
von: Guo, Peizheng, et al.
Veröffentlicht: (2026)
von: Guo, Peizheng, et al.
Veröffentlicht: (2026)
AsyncVLA: Asynchronous Flow Matching for Vision-Language-Action Models
von: Jiang, Yuhua, et al.
Veröffentlicht: (2025)
von: Jiang, Yuhua, et al.
Veröffentlicht: (2025)
ST4VLA: Spatially Guided Training for Vision-Language-Action Models
von: Ye, Jinhui, et al.
Veröffentlicht: (2026)
von: Ye, Jinhui, et al.
Veröffentlicht: (2026)
RobustVLA: Robustness-Aware Reinforcement Post-Training for Vision-Language-Action Models
von: Zhang, Hongyin, et al.
Veröffentlicht: (2025)
von: Zhang, Hongyin, et al.
Veröffentlicht: (2025)
StreamingVLA: Streaming Vision-Language-Action Model with Action Flow Matching and Adaptive Early Observation
von: Shi, Yiran, et al.
Veröffentlicht: (2026)
von: Shi, Yiran, et al.
Veröffentlicht: (2026)
HBVLA: Pushing 1-Bit Post-Training Quantization for Vision-Language-Action Models
von: Yan, Xin, et al.
Veröffentlicht: (2026)
von: Yan, Xin, et al.
Veröffentlicht: (2026)
Letting Trajectories Spread: Quality-Preserving Control for Diverse Flow Matching
von: Wu, Jingxuan, et al.
Veröffentlicht: (2025)
von: Wu, Jingxuan, et al.
Veröffentlicht: (2025)
STARE-VLA: Progressive Stage-Aware Reinforcement for Fine-Tuning Vision-Language-Action Models
von: Xu, Feng, et al.
Veröffentlicht: (2025)
von: Xu, Feng, et al.
Veröffentlicht: (2025)
StereoVLA: Enhancing Vision-Language-Action Models with Stereo Vision
von: Deng, Shengliang, et al.
Veröffentlicht: (2025)
von: Deng, Shengliang, et al.
Veröffentlicht: (2025)
TA-VLA: Elucidating the Design Space of Torque-aware Vision-Language-Action Models
von: Zhang, Zongzheng, et al.
Veröffentlicht: (2025)
von: Zhang, Zongzheng, et al.
Veröffentlicht: (2025)
FlowVLA: Visual Chain of Thought-based Motion Reasoning for Vision-Language-Action Models
von: Zhong, Zhide, et al.
Veröffentlicht: (2025)
von: Zhong, Zhide, et al.
Veröffentlicht: (2025)
GeoAware-VLA: Implicit Geometry Aware Vision-Language-Action Model
von: Abouzeid, Ali, et al.
Veröffentlicht: (2025)
von: Abouzeid, Ali, et al.
Veröffentlicht: (2025)
VLA-Reasoner: Empowering Vision-Language-Action Models with Reasoning via Online Monte Carlo Tree Search
von: Guo, Wenkai, et al.
Veröffentlicht: (2025)
von: Guo, Wenkai, et al.
Veröffentlicht: (2025)
RoboNurse-VLA: Robotic Scrub Nurse System based on Vision-Language-Action Model
von: Li, Shunlei, et al.
Veröffentlicht: (2024)
von: Li, Shunlei, et al.
Veröffentlicht: (2024)
OmniVLA-RL: A Vision-Language-Action Model with Spatial Understanding and Online RL
von: Jie, Haoxiang, et al.
Veröffentlicht: (2026)
von: Jie, Haoxiang, et al.
Veröffentlicht: (2026)
FocusVLA: Focused Visual Utilization for Vision-Language-Action Models
von: Zhang, Yichi, et al.
Veröffentlicht: (2026)
von: Zhang, Yichi, et al.
Veröffentlicht: (2026)
DynamicVLA: A Vision-Language-Action Model for Dynamic Object Manipulation
von: Xie, Haozhe, et al.
Veröffentlicht: (2026)
von: Xie, Haozhe, et al.
Veröffentlicht: (2026)
RLinf-VLA: A Unified and Efficient Framework for Reinforcement Learning of Vision-Language-Action Models
von: Zang, Hongzhi, et al.
Veröffentlicht: (2025)
von: Zang, Hongzhi, et al.
Veröffentlicht: (2025)
MobileVLA-R1: Reinforcing Vision-Language-Action for Mobile Robots
von: Huang, Ting, et al.
Veröffentlicht: (2025)
von: Huang, Ting, et al.
Veröffentlicht: (2025)
ProbeFlow: Training-Free Adaptive Flow Matching for Vision-Language-Action Models
von: Fang, Zhou, et al.
Veröffentlicht: (2026)
von: Fang, Zhou, et al.
Veröffentlicht: (2026)
SwitchVLA: Execution-Aware Task Switching for Vision-Language-Action Models
von: Li, Meng, et al.
Veröffentlicht: (2025)
von: Li, Meng, et al.
Veröffentlicht: (2025)
Evolving Diffusion and Flow Matching Policies for Online Reinforcement Learning
von: Zhang, Chubin, et al.
Veröffentlicht: (2025)
von: Zhang, Chubin, et al.
Veröffentlicht: (2025)
SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model
von: Qu, Delin, et al.
Veröffentlicht: (2025)
von: Qu, Delin, et al.
Veröffentlicht: (2025)
MAP-VLA: Memory-Augmented Prompting for Vision-Language-Action Model in Robotic Manipulation
von: Li, Runhao, et al.
Veröffentlicht: (2025)
von: Li, Runhao, et al.
Veröffentlicht: (2025)
SG-VLA: Learning Spatially-Grounded Vision-Language-Action Models for Mobile Manipulation
von: Tu, Ruisen, et al.
Veröffentlicht: (2026)
von: Tu, Ruisen, et al.
Veröffentlicht: (2026)
RetoVLA: Reusing Register Tokens for Spatial Reasoning in Vision-Language-Action Models
von: Koo, Jiyeon, et al.
Veröffentlicht: (2025)
von: Koo, Jiyeon, et al.
Veröffentlicht: (2025)
Reinforcement Fine-Tuning of Flow-Matching Policies for Vision-Language-Action Models
von: Lyu, Mingyang, et al.
Veröffentlicht: (2025)
von: Lyu, Mingyang, et al.
Veröffentlicht: (2025)
VLA-R1: Enhancing Reasoning in Vision-Language-Action Models
von: Ye, Angen, et al.
Veröffentlicht: (2025)
von: Ye, Angen, et al.
Veröffentlicht: (2025)
AC^2-VLA: Action-Context-Aware Adaptive Computation in Vision-Language-Action Models for Efficient Robotic Manipulation
von: Yu, Wenda, et al.
Veröffentlicht: (2026)
von: Yu, Wenda, et al.
Veröffentlicht: (2026)
SpecPrune-VLA: Accelerating Vision-Language-Action Models via Action-Aware Self-Speculative Pruning
von: Wang, Hanzhen, et al.
Veröffentlicht: (2025)
von: Wang, Hanzhen, et al.
Veröffentlicht: (2025)
FlowHijack: A Dynamics-Aware Backdoor Attack on Flow-Matching Vision-Language-Action Models
von: An, Xinyuan, et al.
Veröffentlicht: (2026)
von: An, Xinyuan, et al.
Veröffentlicht: (2026)
StyleVLA: Driving Style-Aware Vision Language Action Model for Autonomous Driving
von: Gao, Yuan, et al.
Veröffentlicht: (2026)
von: Gao, Yuan, et al.
Veröffentlicht: (2026)
TacVLA: Contact-Aware Tactile Fusion for Robust Vision-Language-Action Manipulation
von: Zhang, Kaidi, et al.
Veröffentlicht: (2026)
von: Zhang, Kaidi, et al.
Veröffentlicht: (2026)
RationalVLA: A Rational Vision-Language-Action Model with Dual System
von: Song, Wenxuan, et al.
Veröffentlicht: (2025)
von: Song, Wenxuan, et al.
Veröffentlicht: (2025)
DiG-Flow: Discrepancy-Guided Flow Matching for Robust VLA Models
von: Zhang, Wanpeng, et al.
Veröffentlicht: (2025)
von: Zhang, Wanpeng, et al.
Veröffentlicht: (2025)
Latent Reasoning VLA: Latent Thinking and Prediction for Vision-Language-Action Models
von: Bai, Shuanghao, et al.
Veröffentlicht: (2026)
von: Bai, Shuanghao, et al.
Veröffentlicht: (2026)
PriorVLA: Prior-Preserving Adaptation for Vision-Language-Action Models
von: Guo, Xinyu, et al.
Veröffentlicht: (2026)
von: Guo, Xinyu, et al.
Veröffentlicht: (2026)
CRL-VLA: Continual Vision-Language-Action Learning
von: Zeng, Qixin, et al.
Veröffentlicht: (2026)
von: Zeng, Qixin, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Towards Backdoor-Based Ownership Verification for Vision-Language-Action Models
von: Sun, Ming, et al.
Veröffentlicht: (2026) -
FM-IRL: Flow-Matching for Reward Modeling and Policy Regularization in Reinforcement Learning
von: Wan, Zhenglin, et al.
Veröffentlicht: (2025) -
PAPO-VLA: Planning-Aware Policy Optimization for Vision-Language-Action Models
von: Guo, Peizheng, et al.
Veröffentlicht: (2026) -
AsyncVLA: Asynchronous Flow Matching for Vision-Language-Action Models
von: Jiang, Yuhua, et al.
Veröffentlicht: (2025) -
ST4VLA: Spatially Guided Training for Vision-Language-Action Models
von: Ye, Jinhui, et al.
Veröffentlicht: (2026)