APPLV: Adaptive Planner Parameter Learning from Vision-Language-Action Model
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lu, Yuanjie, Wang, Beichen, Wu, Zhengqi, Li, Yang, Lin, Xiaomin, Mao, Chengzhi, Xiao, Xuesu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CORAL: COntextual Reasoning And Local Planning in A Hierarchical VLM Framework for Underwater Monitoring
von: Wu, Zhenqi, et al.
Veröffentlicht: (2026)
von: Wu, Zhenqi, et al.
Veröffentlicht: (2026)
Adaptive Dynamics Planning for Robot Navigation
von: Lu, Yuanjie, et al.
Veröffentlicht: (2025)
von: Lu, Yuanjie, et al.
Veröffentlicht: (2025)
Moving Through Clutter: Scaling Data Collection and Benchmarking for 3D Scene-Aware Humanoid Locomotion via Virtual Reality
von: Wang, Beichen, et al.
Veröffentlicht: (2026)
von: Wang, Beichen, et al.
Veröffentlicht: (2026)
Towards Practical World Model-based Reinforcement Learning for Vision-Language-Action Models
von: Zhang, Zhilong, et al.
Veröffentlicht: (2026)
von: Zhang, Zhilong, et al.
Veröffentlicht: (2026)
From Inference Efficiency to Embodied Efficiency: Revisiting Efficiency Metrics for Vision-Language-Action Models
von: Li, Zhuofan, et al.
Veröffentlicht: (2026)
von: Li, Zhuofan, et al.
Veröffentlicht: (2026)
Reward Training Wheels: Adaptive Auxiliary Rewards for Robotics Reinforcement Learning
von: Wang, Linji, et al.
Veröffentlicht: (2025)
von: Wang, Linji, et al.
Veröffentlicht: (2025)
Dyna-LfLH: Learning Agile Navigation in Dynamic Environments from Learned Hallucination
von: Ghani, Saad Abdul, et al.
Veröffentlicht: (2024)
von: Ghani, Saad Abdul, et al.
Veröffentlicht: (2024)
villa-X: Enhancing Latent Action Modeling in Vision-Language-Action Models
von: Chen, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Chen, Xiaoyu, et al.
Veröffentlicht: (2025)
Imitation Learning with Limited Actions via Diffusion Planners and Deep Koopman Controllers
von: Bi, Jianxin, et al.
Veröffentlicht: (2024)
von: Bi, Jianxin, et al.
Veröffentlicht: (2024)
CO-RFT: Efficient Fine-Tuning of Vision-Language-Action Models through Chunked Offline Reinforcement Learning
von: Huang, Dongchi, et al.
Veröffentlicht: (2025)
von: Huang, Dongchi, et al.
Veröffentlicht: (2025)
Tactile-VLA: Unlocking Vision-Language-Action Model's Physical Knowledge for Tactile Generalization
von: Huang, Jialei, et al.
Veröffentlicht: (2025)
von: Huang, Jialei, et al.
Veröffentlicht: (2025)
FAST: Efficient Action Tokenization for Vision-Language-Action Models
von: Pertsch, Karl, et al.
Veröffentlicht: (2025)
von: Pertsch, Karl, et al.
Veröffentlicht: (2025)
Reinforcement Fine-Tuning of Flow-Matching Policies for Vision-Language-Action Models
von: Lyu, Mingyang, et al.
Veröffentlicht: (2025)
von: Lyu, Mingyang, et al.
Veröffentlicht: (2025)
World-Value-Action Model: Implicit Planning for Vision-Language-Action Systems
von: Li, Runze, et al.
Veröffentlicht: (2026)
von: Li, Runze, et al.
Veröffentlicht: (2026)
Dexterous Legged Locomotion in Confined 3D Spaces with Reinforcement Learning
von: Xu, Zifan, et al.
Veröffentlicht: (2024)
von: Xu, Zifan, et al.
Veröffentlicht: (2024)
Confidence Calibration in Vision-Language-Action Models
von: Zollo, Thomas P, et al.
Veröffentlicht: (2025)
von: Zollo, Thomas P, et al.
Veröffentlicht: (2025)
VLM See, Robot Do: Human Demo Video to Robot Action Plan via Vision Language Model
von: Wang, Beichen, et al.
Veröffentlicht: (2024)
von: Wang, Beichen, et al.
Veröffentlicht: (2024)
Few-Shot Vision-Language Action-Incremental Policy Learning
von: Song, Mingchen, et al.
Veröffentlicht: (2025)
von: Song, Mingchen, et al.
Veröffentlicht: (2025)
Latent Adaptive Planner for Dynamic Manipulation
von: Noh, Donghun, et al.
Veröffentlicht: (2025)
von: Noh, Donghun, et al.
Veröffentlicht: (2025)
PVI: Plug-in Visual Injection for Vision-Language-Action Models
von: Zhang, Zezhou, et al.
Veröffentlicht: (2026)
von: Zhang, Zezhou, et al.
Veröffentlicht: (2026)
OpenVLA: An Open-Source Vision-Language-Action Model
von: Kim, Moo Jin, et al.
Veröffentlicht: (2024)
von: Kim, Moo Jin, et al.
Veröffentlicht: (2024)
Xiaomi-Robotics-0: An Open-Sourced Vision-Language-Action Model with Real-Time Execution
von: Cai, Rui, et al.
Veröffentlicht: (2026)
von: Cai, Rui, et al.
Veröffentlicht: (2026)
Efficient Vision-Language-Action Models for Embodied Manipulation: A Systematic Survey
von: Guan, Weifan, et al.
Veröffentlicht: (2025)
von: Guan, Weifan, et al.
Veröffentlicht: (2025)
SCALE: Self-uncertainty Conditioned Adaptive Looking and Execution for Vision-Language-Action Models
von: Choi, Hyeonbeom, et al.
Veröffentlicht: (2026)
von: Choi, Hyeonbeom, et al.
Veröffentlicht: (2026)
HeiSD: Hybrid Speculative Decoding for Embodied Vision-Language-Action Models with Kinematic Awareness
von: Zheng, Zihao, et al.
Veröffentlicht: (2026)
von: Zheng, Zihao, et al.
Veröffentlicht: (2026)
Simple Recipe Works: Vision-Language-Action Models are Natural Continual Learners with Reinforcement Learning
von: Hu, Jiaheng, et al.
Veröffentlicht: (2026)
von: Hu, Jiaheng, et al.
Veröffentlicht: (2026)
Predictive Planner for Autonomous Driving with Consistency Models
von: Li, Anjian, et al.
Veröffentlicht: (2025)
von: Li, Anjian, et al.
Veröffentlicht: (2025)
Continuous Reasoning for Vision-Language-Action
von: Wu, Yueh-Hua, et al.
Veröffentlicht: (2026)
von: Wu, Yueh-Hua, et al.
Veröffentlicht: (2026)
CRL-VLA: Continual Vision-Language-Action Learning
von: Zeng, Qixin, et al.
Veröffentlicht: (2026)
von: Zeng, Qixin, et al.
Veröffentlicht: (2026)
RobustVLA: Robustness-Aware Reinforcement Post-Training for Vision-Language-Action Models
von: Zhang, Hongyin, et al.
Veröffentlicht: (2025)
von: Zhang, Hongyin, et al.
Veröffentlicht: (2025)
Sigma: The Key for Vision-Language-Action Models toward Telepathic Alignment
von: Wang, Libo
Veröffentlicht: (2025)
von: Wang, Libo
Veröffentlicht: (2025)
MEM: Multi-Scale Embodied Memory for Vision Language Action Models
von: Torne, Marcel, et al.
Veröffentlicht: (2026)
von: Torne, Marcel, et al.
Veröffentlicht: (2026)
RL Token: Bootstrapping Online RL with Vision-Language-Action Models
von: Xu, Charles, et al.
Veröffentlicht: (2026)
von: Xu, Charles, et al.
Veröffentlicht: (2026)
Discrete Diffusion VLA: Bringing Discrete Diffusion to Action Decoding in Vision-Language-Action Policies
von: Liang, Zhixuan, et al.
Veröffentlicht: (2025)
von: Liang, Zhixuan, et al.
Veröffentlicht: (2025)
Temporal Difference Calibration in Sequential Tasks: Application to Vision-Language-Action Models
von: Francis-Meretzki, Shelly, et al.
Veröffentlicht: (2026)
von: Francis-Meretzki, Shelly, et al.
Veröffentlicht: (2026)
OmniVLA: An Omni-Modal Vision-Language-Action Model for Robot Navigation
von: Hirose, Noriaki, et al.
Veröffentlicht: (2025)
von: Hirose, Noriaki, et al.
Veröffentlicht: (2025)
SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics
von: Shukor, Mustafa, et al.
Veröffentlicht: (2025)
von: Shukor, Mustafa, et al.
Veröffentlicht: (2025)
$π_0$: A Vision-Language-Action Flow Model for General Robot Control
von: Black, Kevin, et al.
Veröffentlicht: (2024)
von: Black, Kevin, et al.
Veröffentlicht: (2024)
Do What You Say: Steering Vision-Language-Action Models via Runtime Reasoning-Action Alignment Verification
von: Wu, Yilin, et al.
Veröffentlicht: (2025)
von: Wu, Yilin, et al.
Veröffentlicht: (2025)
$π_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
von: Intelligence, Physical, et al.
Veröffentlicht: (2025)
von: Intelligence, Physical, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
CORAL: COntextual Reasoning And Local Planning in A Hierarchical VLM Framework for Underwater Monitoring
von: Wu, Zhenqi, et al.
Veröffentlicht: (2026) -
Adaptive Dynamics Planning for Robot Navigation
von: Lu, Yuanjie, et al.
Veröffentlicht: (2025) -
Moving Through Clutter: Scaling Data Collection and Benchmarking for 3D Scene-Aware Humanoid Locomotion via Virtual Reality
von: Wang, Beichen, et al.
Veröffentlicht: (2026) -
Towards Practical World Model-based Reinforcement Learning for Vision-Language-Action Models
von: Zhang, Zhilong, et al.
Veröffentlicht: (2026) -
From Inference Efficiency to Embodied Efficiency: Revisiting Efficiency Metrics for Vision-Language-Action Models
von: Li, Zhuofan, et al.
Veröffentlicht: (2026)