Few-Shot Vision-Language Action-Incremental Policy Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Song, Mingchen, Deng, Xiang, Zhong, Guoqiang, Lv, Qi, Wan, Jia, Li, Yinchuan, Hao, Jianye, Guan, Weili |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
STAR: Learning Diverse Robot Skill Abstractions through Rotation-Augmented Vector Quantization
von: Li, Hao, et al.
Veröffentlicht: (2025)
von: Li, Hao, et al.
Veröffentlicht: (2025)
Spatial-Temporal Graph Diffusion Policy with Kinematic Modeling for Bimanual Robotic Manipulation
von: Lv, Qi, et al.
Veröffentlicht: (2025)
von: Lv, Qi, et al.
Veröffentlicht: (2025)
EnergyAction: Unimanual to Bimanual Composition with Energy-Based Models
von: Song, Mingchen, et al.
Veröffentlicht: (2026)
von: Song, Mingchen, et al.
Veröffentlicht: (2026)
3D-AffordanceLLM: Harnessing Large Language Models for Open-Vocabulary Affordance Detection in 3D Worlds
von: Chu, Hengshuo, et al.
Veröffentlicht: (2025)
von: Chu, Hengshuo, et al.
Veröffentlicht: (2025)
Conditioning Matters: Training Diffusion Policies is Faster Than You Think
von: Dong, Zibin, et al.
Veröffentlicht: (2025)
von: Dong, Zibin, et al.
Veröffentlicht: (2025)
Policy Adaptation via Language Optimization: Decomposing Tasks for Few-Shot Imitation
von: Myers, Vivek, et al.
Veröffentlicht: (2024)
von: Myers, Vivek, et al.
Veröffentlicht: (2024)
Incremental Few-Shot Adaptation for Non-Prehensile Object Manipulation using Parallelizable Physics Simulators
von: Baumeister, Fabian, et al.
Veröffentlicht: (2024)
von: Baumeister, Fabian, et al.
Veröffentlicht: (2024)
F1: A Vision-Language-Action Model Bridging Understanding and Generation to Actions
von: Lv, Qi, et al.
Veröffentlicht: (2025)
von: Lv, Qi, et al.
Veröffentlicht: (2025)
Efficient Vision-Language-Action Models for Embodied Manipulation: A Systematic Survey
von: Guan, Weifan, et al.
Veröffentlicht: (2025)
von: Guan, Weifan, et al.
Veröffentlicht: (2025)
Reinforcement Fine-Tuning of Flow-Matching Policies for Vision-Language-Action Models
von: Lyu, Mingyang, et al.
Veröffentlicht: (2025)
von: Lyu, Mingyang, et al.
Veröffentlicht: (2025)
Premier-TACO is a Few-Shot Policy Learner: Pretraining Multitask Representation via Temporal Action-Driven Contrastive Loss
von: Zheng, Ruijie, et al.
Veröffentlicht: (2024)
von: Zheng, Ruijie, et al.
Veröffentlicht: (2024)
Proto-CLIP: Vision-Language Prototypical Network for Few-Shot Learning
von: P, Jishnu Jaykumar, et al.
Veröffentlicht: (2023)
von: P, Jishnu Jaykumar, et al.
Veröffentlicht: (2023)
EmbodiedMAE: A Unified 3D Multi-Modal Representation for Robot Manipulation
von: Dong, Zibin, et al.
Veröffentlicht: (2025)
von: Dong, Zibin, et al.
Veröffentlicht: (2025)
AnySlot: Goal-Conditioned Vision-Language-Action Policies for Zero-Shot Slot-Level Placement
von: Hu, Zhaofeng, et al.
Veröffentlicht: (2026)
von: Hu, Zhaofeng, et al.
Veröffentlicht: (2026)
FAST: Efficient Action Tokenization for Vision-Language-Action Models
von: Pertsch, Karl, et al.
Veröffentlicht: (2025)
von: Pertsch, Karl, et al.
Veröffentlicht: (2025)
Flow-Enabled Generalization to Human Demonstrations in Few-Shot Imitation Learning
von: Tang, Runze, et al.
Veröffentlicht: (2026)
von: Tang, Runze, et al.
Veröffentlicht: (2026)
Improving Vision-Language-Action Model with Online Reinforcement Learning
von: Guo, Yanjiang, et al.
Veröffentlicht: (2025)
von: Guo, Yanjiang, et al.
Veröffentlicht: (2025)
Benchmarking Vision, Language, & Action Models on Robotic Learning Tasks
von: Guruprasad, Pranav, et al.
Veröffentlicht: (2024)
von: Guruprasad, Pranav, et al.
Veröffentlicht: (2024)
Universal Pose Pretraining for Generalizable Vision-Language-Action Policies
von: Lin, Haitao, et al.
Veröffentlicht: (2026)
von: Lin, Haitao, et al.
Veröffentlicht: (2026)
FlowRetrieval: Flow-Guided Data Retrieval for Few-Shot Imitation Learning
von: Lin, Li-Heng, et al.
Veröffentlicht: (2024)
von: Lin, Li-Heng, et al.
Veröffentlicht: (2024)
Lifelong Ensemble Learning based on Multiple Representations for Few-Shot Object Recognition
von: Kasaei, Hamidreza, et al.
Veröffentlicht: (2022)
von: Kasaei, Hamidreza, et al.
Veröffentlicht: (2022)
CRL-VLA: Continual Vision-Language-Action Learning
von: Zeng, Qixin, et al.
Veröffentlicht: (2026)
von: Zeng, Qixin, et al.
Veröffentlicht: (2026)
Confidence Calibration in Vision-Language-Action Models
von: Zollo, Thomas P, et al.
Veröffentlicht: (2025)
von: Zollo, Thomas P, et al.
Veröffentlicht: (2025)
APPLV: Adaptive Planner Parameter Learning from Vision-Language-Action Model
von: Lu, Yuanjie, et al.
Veröffentlicht: (2026)
von: Lu, Yuanjie, et al.
Veröffentlicht: (2026)
VLA-Touch: Enhancing Vision-Language-Action Models with Dual-Level Tactile Feedback
von: Bi, Jianxin, et al.
Veröffentlicht: (2025)
von: Bi, Jianxin, et al.
Veröffentlicht: (2025)
World-Value-Action Model: Implicit Planning for Vision-Language-Action Systems
von: Li, Runze, et al.
Veröffentlicht: (2026)
von: Li, Runze, et al.
Veröffentlicht: (2026)
The Lie We Tell: Correcting the Euclidean Fallacy in Vision Language Action Policies via Score Matching on Tangent Space
von: Chuang, Bing-Cheng, et al.
Veröffentlicht: (2026)
von: Chuang, Bing-Cheng, et al.
Veröffentlicht: (2026)
ActionCodec: What Makes for Good Action Tokenizers
von: Dong, Zibin, et al.
Veröffentlicht: (2026)
von: Dong, Zibin, et al.
Veröffentlicht: (2026)
Discrete Diffusion VLA: Bringing Discrete Diffusion to Action Decoding in Vision-Language-Action Policies
von: Liang, Zhixuan, et al.
Veröffentlicht: (2025)
von: Liang, Zhixuan, et al.
Veröffentlicht: (2025)
LensDFF: Language-enhanced Sparse Feature Distillation for Efficient Few-Shot Dexterous Manipulation
von: Feng, Qian, et al.
Veröffentlicht: (2025)
von: Feng, Qian, et al.
Veröffentlicht: (2025)
Towards Practical World Model-based Reinforcement Learning for Vision-Language-Action Models
von: Zhang, Zhilong, et al.
Veröffentlicht: (2026)
von: Zhang, Zhilong, et al.
Veröffentlicht: (2026)
AsyncVLA: Asynchronous Flow Matching for Vision-Language-Action Models
von: Jiang, Yuhua, et al.
Veröffentlicht: (2025)
von: Jiang, Yuhua, et al.
Veröffentlicht: (2025)
Few-Shot Task Learning through Inverse Generative Modeling
von: Netanyahu, Aviv, et al.
Veröffentlicht: (2024)
von: Netanyahu, Aviv, et al.
Veröffentlicht: (2024)
Jump-Start Reinforcement Learning with Vision-Language-Action Regularization
von: Moroncelli, Angelo, et al.
Veröffentlicht: (2026)
von: Moroncelli, Angelo, et al.
Veröffentlicht: (2026)
See Once, Then Act: Vision-Language-Action Model with Task Learning from One-Shot Video Demonstrations
von: Chen, Guangyan, et al.
Veröffentlicht: (2025)
von: Chen, Guangyan, et al.
Veröffentlicht: (2025)
Mechanistic Finetuning of Vision-Language-Action Models via Few-Shot Demonstrations
von: Mitra, Chancharik, et al.
Veröffentlicht: (2025)
von: Mitra, Chancharik, et al.
Veröffentlicht: (2025)
Simple Recipe Works: Vision-Language-Action Models are Natural Continual Learners with Reinforcement Learning
von: Hu, Jiaheng, et al.
Veröffentlicht: (2026)
von: Hu, Jiaheng, et al.
Veröffentlicht: (2026)
DyQ-VLA: Temporal-Dynamic-Aware Quantization for Embodied Vision-Language-Action Models
von: Zheng, Zihao, et al.
Veröffentlicht: (2026)
von: Zheng, Zihao, et al.
Veröffentlicht: (2026)
Continuous Reasoning for Vision-Language-Action
von: Wu, Yueh-Hua, et al.
Veröffentlicht: (2026)
von: Wu, Yueh-Hua, et al.
Veröffentlicht: (2026)
OpenVLA: An Open-Source Vision-Language-Action Model
von: Kim, Moo Jin, et al.
Veröffentlicht: (2024)
von: Kim, Moo Jin, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
STAR: Learning Diverse Robot Skill Abstractions through Rotation-Augmented Vector Quantization
von: Li, Hao, et al.
Veröffentlicht: (2025) -
Spatial-Temporal Graph Diffusion Policy with Kinematic Modeling for Bimanual Robotic Manipulation
von: Lv, Qi, et al.
Veröffentlicht: (2025) -
EnergyAction: Unimanual to Bimanual Composition with Energy-Based Models
von: Song, Mingchen, et al.
Veröffentlicht: (2026) -
3D-AffordanceLLM: Harnessing Large Language Models for Open-Vocabulary Affordance Detection in 3D Worlds
von: Chu, Hengshuo, et al.
Veröffentlicht: (2025) -
Conditioning Matters: Training Diffusion Policies is Faster Than You Think
von: Dong, Zibin, et al.
Veröffentlicht: (2025)