Agentic-VLA: Efficient Online Adaptation for Vision-Language-Action Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jin, Ruofan, Zhang, Zaixi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
HyperVLA: Efficient Inference in Vision-Language-Action Models via Hypernetworks
von: Xiong, Zheng, et al.
Veröffentlicht: (2025)
von: Xiong, Zheng, et al.
Veröffentlicht: (2025)
CRL-VLA: Continual Vision-Language-Action Learning
von: Zeng, Qixin, et al.
Veröffentlicht: (2026)
von: Zeng, Qixin, et al.
Veröffentlicht: (2026)
AsyncVLA: Asynchronous Flow Matching for Vision-Language-Action Models
von: Jiang, Yuhua, et al.
Veröffentlicht: (2025)
von: Jiang, Yuhua, et al.
Veröffentlicht: (2025)
DeepThinkVLA: Enhancing Reasoning Capability of Vision-Language-Action Models
von: Yin, Cheng, et al.
Veröffentlicht: (2025)
von: Yin, Cheng, et al.
Veröffentlicht: (2025)
CoT-VLA: Visual Chain-of-Thought Reasoning for Vision-Language-Action Models
von: Zhao, Qingqing, et al.
Veröffentlicht: (2025)
von: Zhao, Qingqing, et al.
Veröffentlicht: (2025)
villa-X: Enhancing Latent Action Modeling in Vision-Language-Action Models
von: Chen, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Chen, Xiaoyu, et al.
Veröffentlicht: (2025)
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
von: Yang, Ruihan, et al.
Veröffentlicht: (2025)
von: Yang, Ruihan, et al.
Veröffentlicht: (2025)
TTF-VLA: Temporal Token Fusion via Pixel-Attention Integration for Vision-Language-Action Models
von: Liu, Chenghao, et al.
Veröffentlicht: (2025)
von: Liu, Chenghao, et al.
Veröffentlicht: (2025)
DeeR-VLA: Dynamic Inference of Multimodal Large Language Models for Efficient Robot Execution
von: Yue, Yang, et al.
Veröffentlicht: (2024)
von: Yue, Yang, et al.
Veröffentlicht: (2024)
Understanding Asynchronous Inference Methods for Vision-Language-Action Models
von: Agouzoul, Ayoub
Veröffentlicht: (2026)
von: Agouzoul, Ayoub
Veröffentlicht: (2026)
Continuous Reasoning for Vision-Language-Action
von: Wu, Yueh-Hua, et al.
Veröffentlicht: (2026)
von: Wu, Yueh-Hua, et al.
Veröffentlicht: (2026)
UniVLA: Learning to Act Anywhere with Task-centric Latent Actions
von: Bu, Qingwen, et al.
Veröffentlicht: (2025)
von: Bu, Qingwen, et al.
Veröffentlicht: (2025)
Atomic Action Slicing: Planner-Aligned Options for Generalist VLA Agents
von: Tabakov, Stefan, et al.
Veröffentlicht: (2025)
von: Tabakov, Stefan, et al.
Veröffentlicht: (2025)
A Survey on Efficient Vision-Language-Action Models
von: Yu, Zhaoshu, et al.
Veröffentlicht: (2025)
von: Yu, Zhaoshu, et al.
Veröffentlicht: (2025)
Verifier-free Test-Time Sampling for Vision Language Action Models
von: Jang, Suhyeok, et al.
Veröffentlicht: (2025)
von: Jang, Suhyeok, et al.
Veröffentlicht: (2025)
UAV-VLA: Vision-Language-Action System for Large Scale Aerial Mission Generation
von: Sautenkov, Oleg, et al.
Veröffentlicht: (2025)
von: Sautenkov, Oleg, et al.
Veröffentlicht: (2025)
VAMOS: A Hierarchical Vision-Language-Action Model for Capability-Modulated and Steerable Navigation
von: Castro, Mateo Guaman, et al.
Veröffentlicht: (2025)
von: Castro, Mateo Guaman, et al.
Veröffentlicht: (2025)
Pretrained Vision-Language-Action Models are Surprisingly Resistant to Forgetting in Continual Learning
von: Liu, Huihan, et al.
Veröffentlicht: (2026)
von: Liu, Huihan, et al.
Veröffentlicht: (2026)
Bridging Embodiment Gaps: Deploying Vision-Language-Action Models on Soft Robots
von: Su, Haochen, et al.
Veröffentlicht: (2025)
von: Su, Haochen, et al.
Veröffentlicht: (2025)
Do What You Say: Steering Vision-Language-Action Models via Runtime Reasoning-Action Alignment Verification
von: Wu, Yilin, et al.
Veröffentlicht: (2025)
von: Wu, Yilin, et al.
Veröffentlicht: (2025)
RoboGene: Boosting VLA Pre-training via Diversity-Driven Agentic Framework for Real-World Task Generation
von: Zhang, Yixue, et al.
Veröffentlicht: (2026)
von: Zhang, Yixue, et al.
Veröffentlicht: (2026)
SCALE: Self-uncertainty Conditioned Adaptive Looking and Execution for Vision-Language-Action Models
von: Choi, Hyeonbeom, et al.
Veröffentlicht: (2026)
von: Choi, Hyeonbeom, et al.
Veröffentlicht: (2026)
Hi Robot: Open-Ended Instruction Following with Hierarchical Vision-Language-Action Models
von: Shi, Lucy Xiaoyang, et al.
Veröffentlicht: (2025)
von: Shi, Lucy Xiaoyang, et al.
Veröffentlicht: (2025)
Jump-Start Reinforcement Learning with Vision-Language-Action Regularization
von: Moroncelli, Angelo, et al.
Veröffentlicht: (2026)
von: Moroncelli, Angelo, et al.
Veröffentlicht: (2026)
ETA-VLA: Efficient Token Adaptation via Temporal Fusion and Intra-LLM Sparsification for Vision-Language-Action Models
von: Wang, Yiru, et al.
Veröffentlicht: (2026)
von: Wang, Yiru, et al.
Veröffentlicht: (2026)
Contrastive Conceptor Activation Steering (COAST): Unlocking Vision-Language-Action Models through Hidden States
von: Miao, Miranda Muqing, et al.
Veröffentlicht: (2026)
von: Miao, Miranda Muqing, et al.
Veröffentlicht: (2026)
VLA Models Are More Generalizable Than You Think: Revisiting Physical and Spatial Modeling
von: Li, Weiqi, et al.
Veröffentlicht: (2025)
von: Li, Weiqi, et al.
Veröffentlicht: (2025)
Shifting Uncertainty to Critical Moments: Towards Reliable Uncertainty Quantification for VLA Model
von: Tang, Yanchuan, et al.
Veröffentlicht: (2026)
von: Tang, Yanchuan, et al.
Veröffentlicht: (2026)
Online Human Action Detection during Escorting
von: Mondal, Siddhartha, et al.
Veröffentlicht: (2025)
von: Mondal, Siddhartha, et al.
Veröffentlicht: (2025)
Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success
von: Kim, Moo Jin, et al.
Veröffentlicht: (2025)
von: Kim, Moo Jin, et al.
Veröffentlicht: (2025)
RobustVLA: Robustness-Aware Reinforcement Post-Training for Vision-Language-Action Models
von: Zhang, Hongyin, et al.
Veröffentlicht: (2025)
von: Zhang, Hongyin, et al.
Veröffentlicht: (2025)
SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics
von: Shukor, Mustafa, et al.
Veröffentlicht: (2025)
von: Shukor, Mustafa, et al.
Veröffentlicht: (2025)
AR-VLA: True Autoregressive Action Expert for Vision-Language-Action Models
von: Hu, Yutong, et al.
Veröffentlicht: (2026)
von: Hu, Yutong, et al.
Veröffentlicht: (2026)
Pure Vision Language Action (VLA) Models: A Comprehensive Survey
von: Zhang, Dapeng, et al.
Veröffentlicht: (2025)
von: Zhang, Dapeng, et al.
Veröffentlicht: (2025)
VLMgineer: Vision Language Models as Robotic Toolsmiths
von: Gao, George Jiayuan, et al.
Veröffentlicht: (2025)
von: Gao, George Jiayuan, et al.
Veröffentlicht: (2025)
OpenVLA: An Open-Source Vision-Language-Action Model
von: Kim, Moo Jin, et al.
Veröffentlicht: (2024)
von: Kim, Moo Jin, et al.
Veröffentlicht: (2024)
Vision-Language Foundation Models as Effective Robot Imitators
von: Li, Xinghang, et al.
Veröffentlicht: (2023)
von: Li, Xinghang, et al.
Veröffentlicht: (2023)
Vision Language Models are In-Context Value Learners
von: Ma, Yecheng Jason, et al.
Veröffentlicht: (2024)
von: Ma, Yecheng Jason, et al.
Veröffentlicht: (2024)
SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning
von: Li, Haozhan, et al.
Veröffentlicht: (2025)
von: Li, Haozhan, et al.
Veröffentlicht: (2025)
EndoVLA: Dual-Phase Vision-Language-Action Model for Autonomous Tracking in Endoscopy
von: Ng, Chi Kit, et al.
Veröffentlicht: (2025)
von: Ng, Chi Kit, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
HyperVLA: Efficient Inference in Vision-Language-Action Models via Hypernetworks
von: Xiong, Zheng, et al.
Veröffentlicht: (2025) -
CRL-VLA: Continual Vision-Language-Action Learning
von: Zeng, Qixin, et al.
Veröffentlicht: (2026) -
AsyncVLA: Asynchronous Flow Matching for Vision-Language-Action Models
von: Jiang, Yuhua, et al.
Veröffentlicht: (2025) -
DeepThinkVLA: Enhancing Reasoning Capability of Vision-Language-Action Models
von: Yin, Cheng, et al.
Veröffentlicht: (2025) -
CoT-VLA: Visual Chain-of-Thought Reasoning for Vision-Language-Action Models
von: Zhao, Qingqing, et al.
Veröffentlicht: (2025)