Gespeichert in:
| Hauptverfasser: | Mai, Hung, Zhu, Bin, Do, Tuan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2606.01095 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
WAM-Diff: A Masked Diffusion VLA Framework with MoE and Online Reinforcement Learning for Autonomous Driving
von: Xu, Mingwang, et al.
Veröffentlicht: (2025)
von: Xu, Mingwang, et al.
Veröffentlicht: (2025)
Sci-VLA: Agentic VLA Inference Plugin for Long-Horizon Tasks in Scientific Experiments
von: Pang, Yiwen, et al.
Veröffentlicht: (2026)
von: Pang, Yiwen, et al.
Veröffentlicht: (2026)
SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model
von: Qu, Delin, et al.
Veröffentlicht: (2025)
von: Qu, Delin, et al.
Veröffentlicht: (2025)
Task Success is not Enough: Investigating the Use of Video-Language Models as Behavior Critics for Catching Undesirable Agent Behaviors
von: Guan, Lin, et al.
Veröffentlicht: (2024)
von: Guan, Lin, et al.
Veröffentlicht: (2024)
WAM-Flow: Parallel Coarse-to-Fine Motion Planning via Discrete Flow Matching for Autonomous Driving
von: Xu, Yifang, et al.
Veröffentlicht: (2025)
von: Xu, Yifang, et al.
Veröffentlicht: (2025)
DroneVLA: VLA based Aerial Manipulation
von: Mehboob, Fawad, et al.
Veröffentlicht: (2026)
von: Mehboob, Fawad, et al.
Veröffentlicht: (2026)
OxyGen: Unified KV Cache Management for VLA Inference under Multi-Task Parallelism
von: Li, Xiangyu, et al.
Veröffentlicht: (2026)
von: Li, Xiangyu, et al.
Veröffentlicht: (2026)
RaceVLA: VLA-based Racing Drone Navigation with Human-like Behaviour
von: Serpiva, Valerii, et al.
Veröffentlicht: (2025)
von: Serpiva, Valerii, et al.
Veröffentlicht: (2025)
LoHoVLA: A Unified Vision-Language-Action Model for Long-Horizon Embodied Tasks
von: Yang, Yi, et al.
Veröffentlicht: (2025)
von: Yang, Yi, et al.
Veröffentlicht: (2025)
UniVLA: Learning to Act Anywhere with Task-centric Latent Actions
von: Bu, Qingwen, et al.
Veröffentlicht: (2025)
von: Bu, Qingwen, et al.
Veröffentlicht: (2025)
Don't Blind Your VLA: Aligning Visual Representations for OOD Generalization
von: Kachaev, Nikita, et al.
Veröffentlicht: (2025)
von: Kachaev, Nikita, et al.
Veröffentlicht: (2025)
Restoring Linguistic Grounding in VLA Models via Train-Free Attention Recalibration
von: Zhang, Ninghao, et al.
Veröffentlicht: (2026)
von: Zhang, Ninghao, et al.
Veröffentlicht: (2026)
MetaVLA: Unified Meta Co-training For Efficient Embodied Adaption
von: Li, Chen, et al.
Veröffentlicht: (2025)
von: Li, Chen, et al.
Veröffentlicht: (2025)
From Noise to Intent: Anchoring Generative VLA Policies with Residual Bridges
von: Zhong, Yiming, et al.
Veröffentlicht: (2026)
von: Zhong, Yiming, et al.
Veröffentlicht: (2026)
VacuumVLA: Boosting VLA Capabilities via a Unified Suction and Gripping Tool for Complex Robotic Manipulation
von: Zhou, Hui, et al.
Veröffentlicht: (2025)
von: Zhou, Hui, et al.
Veröffentlicht: (2025)
Sample-Efficient Robot Skill Learning for Construction Tasks: Benchmarking Hierarchical Reinforcement Learning and Vision-Language-Action VLA Model
von: Hu, Zhaofeng, et al.
Veröffentlicht: (2025)
von: Hu, Zhaofeng, et al.
Veröffentlicht: (2025)
Qwen-VLA: Unifying Vision-Language-Action Modeling across Tasks, Environments, and Robot Embodiments
von: Wang, Qiuyue, et al.
Veröffentlicht: (2026)
von: Wang, Qiuyue, et al.
Veröffentlicht: (2026)
MimicDreamer: Aligning Human and Robot Demonstrations for Scalable VLA Training
von: Li, Haoyun, et al.
Veröffentlicht: (2025)
von: Li, Haoyun, et al.
Veröffentlicht: (2025)
VLA-AN: An Efficient and Onboard Vision-Language-Action Framework for Aerial Navigation in Complex Environments
von: Wu, Yuze, et al.
Veröffentlicht: (2025)
von: Wu, Yuze, et al.
Veröffentlicht: (2025)
Towards Deploying VLA without Fine-Tuning: Plug-and-Play Inference-Time VLA Policy Steering via Embodied Evolutionary Diffusion
von: Li, Zhuo, et al.
Veröffentlicht: (2025)
von: Li, Zhuo, et al.
Veröffentlicht: (2025)
AnywhereVLA: Language-Conditioned Exploration and Mobile Manipulation
von: Gubernatorov, Konstantin, et al.
Veröffentlicht: (2025)
von: Gubernatorov, Konstantin, et al.
Veröffentlicht: (2025)
WorldVLA: Towards Autoregressive Action World Model
von: Cen, Jun, et al.
Veröffentlicht: (2025)
von: Cen, Jun, et al.
Veröffentlicht: (2025)
VLA-RAIL: A Real-Time Asynchronous Inference Linker for VLA Models and Robots
von: Zhao, Yongsheng, et al.
Veröffentlicht: (2025)
von: Zhao, Yongsheng, et al.
Veröffentlicht: (2025)
DEFLECT: Delay-Robust Execution via Flow-matching Likelihood-Estimated Counterfactual Tuning for VLA Policies
von: Zhu, Yixiang, et al.
Veröffentlicht: (2026)
von: Zhu, Yixiang, et al.
Veröffentlicht: (2026)
Parallels Between VLA Model Post-Training and Human Motor Learning: Progress, Challenges, and Trends
von: Xiang, Tian-Yu, et al.
Veröffentlicht: (2025)
von: Xiang, Tian-Yu, et al.
Veröffentlicht: (2025)
Reconciling Spatial and Temporal Abstractions for Goal Representation
von: Zadem, Mehdi, et al.
Veröffentlicht: (2024)
von: Zadem, Mehdi, et al.
Veröffentlicht: (2024)
DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation
von: Su, Taiyi, et al.
Veröffentlicht: (2026)
von: Su, Taiyi, et al.
Veröffentlicht: (2026)
SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning
von: Li, Haozhan, et al.
Veröffentlicht: (2025)
von: Li, Haozhan, et al.
Veröffentlicht: (2025)
VLA-0: Building State-of-the-Art VLAs with Zero Modification
von: Goyal, Ankit, et al.
Veröffentlicht: (2025)
von: Goyal, Ankit, et al.
Veröffentlicht: (2025)
End-to-End Dexterous Arm-Hand VLA Policies via Shared Autonomy: VR Teleoperation Augmented by Autonomous Hand VLA Policy for Efficient Data Collection
von: Cui, Yu, et al.
Veröffentlicht: (2025)
von: Cui, Yu, et al.
Veröffentlicht: (2025)
PRIME: Scaffolding Manipulation Tasks with Behavior Primitives for Data-Efficient Imitation Learning
von: Gao, Tian, et al.
Veröffentlicht: (2024)
von: Gao, Tian, et al.
Veröffentlicht: (2024)
RoboGene: Boosting VLA Pre-training via Diversity-Driven Agentic Framework for Real-World Task Generation
von: Zhang, Yixue, et al.
Veröffentlicht: (2026)
von: Zhang, Yixue, et al.
Veröffentlicht: (2026)
VLA-Pro: Cross-Task Procedural Memory Transfer for Vision-Language-Action Models
von: Si, Shengyu, et al.
Veröffentlicht: (2026)
von: Si, Shengyu, et al.
Veröffentlicht: (2026)
ForeAct: Steering Your VLA with Efficient Visual Foresight Planning
von: Zhang, Zhuoyang, et al.
Veröffentlicht: (2026)
von: Zhang, Zhuoyang, et al.
Veröffentlicht: (2026)
FUTURE-VLA: Forecasting Unified Trajectories Under Real-time Execution
von: Fan, Jingjing, et al.
Veröffentlicht: (2026)
von: Fan, Jingjing, et al.
Veröffentlicht: (2026)
Hide-and-Seek in Trajectories: Discovering Failure Signals for VLA Runtime Monitoring
von: Park, Seongheon, et al.
Veröffentlicht: (2026)
von: Park, Seongheon, et al.
Veröffentlicht: (2026)
LiteVLA-Edge: Quantized On-Device Multimodal Control for Embedded Robotics
von: Williams, Justin, et al.
Veröffentlicht: (2026)
von: Williams, Justin, et al.
Veröffentlicht: (2026)
Pure Vision Language Action (VLA) Models: A Comprehensive Survey
von: Zhang, Dapeng, et al.
Veröffentlicht: (2025)
von: Zhang, Dapeng, et al.
Veröffentlicht: (2025)
BioProVLA-Agent: An Affordable, Protocol-Driven, Vision-Enhanced VLA-Enabled Embodied Multi-Agent System with Closed-Loop-Capable Reasoning for Biological Laboratory Manipulation
von: Du, Zhaohui, et al.
Veröffentlicht: (2026)
von: Du, Zhaohui, et al.
Veröffentlicht: (2026)
Graphormer-Guided Task Planning: Beyond Static Rules with LLM Safety Perception
von: Huang, Wanjing, et al.
Veröffentlicht: (2025)
von: Huang, Wanjing, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
WAM-Diff: A Masked Diffusion VLA Framework with MoE and Online Reinforcement Learning for Autonomous Driving
von: Xu, Mingwang, et al.
Veröffentlicht: (2025) -
Sci-VLA: Agentic VLA Inference Plugin for Long-Horizon Tasks in Scientific Experiments
von: Pang, Yiwen, et al.
Veröffentlicht: (2026) -
SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model
von: Qu, Delin, et al.
Veröffentlicht: (2025) -
Task Success is not Enough: Investigating the Use of Video-Language Models as Behavior Critics for Catching Undesirable Agent Behaviors
von: Guan, Lin, et al.
Veröffentlicht: (2024) -
WAM-Flow: Parallel Coarse-to-Fine Motion Planning via Discrete Flow Matching for Autonomous Driving
von: Xu, Yifang, et al.
Veröffentlicht: (2025)