AtomicVLA: Unlocking the Potential of Atomic Skill Learning in Robots
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhang, Likui, Tang, Tao, Zhan, Zhihao, Chen, Xiuwei, Chen, Zisheng, Han, Jianhua, Zhu, Jiangtong, Xu, Pei, Xu, Hang, Wu, Hefeng, Lin, Liang, Liang, Xiaodan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
RoboPearls: Editable Video Simulation for Robot Manipulation
di: Tang, Tao, et al.
Pubblicazione: (2025)
di: Tang, Tao, et al.
Pubblicazione: (2025)
Compose by Focus: Scene Graph-based Atomic Skills
di: Qi, Han, et al.
Pubblicazione: (2025)
di: Qi, Han, et al.
Pubblicazione: (2025)
VidMan: Exploiting Implicit Dynamics from Video Diffusion Model for Effective Robot Manipulation
di: Wen, Youpeng, et al.
Pubblicazione: (2024)
di: Wen, Youpeng, et al.
Pubblicazione: (2024)
Learning Semantic Atomic Skills for Multi-Task Robotic Manipulation
di: Zhu, Yihang, et al.
Pubblicazione: (2025)
di: Zhu, Yihang, et al.
Pubblicazione: (2025)
RoBridge: A Hierarchical Architecture Bridging Cognition and Execution for General Robotic Manipulation
di: Zhang, Kaidong, et al.
Pubblicazione: (2025)
di: Zhang, Kaidong, et al.
Pubblicazione: (2025)
Gentle Manipulation Policy Learning via Demonstrations from VLM Planned Atomic Skills
di: Zhou, Jiayu, et al.
Pubblicazione: (2025)
di: Zhou, Jiayu, et al.
Pubblicazione: (2025)
SimVLA: A Simple VLA Baseline for Robotic Manipulation
di: Luo, Yuankai, et al.
Pubblicazione: (2026)
di: Luo, Yuankai, et al.
Pubblicazione: (2026)
An Atomic Skill Library Construction Method for Data-Efficient Embodied Manipulation
di: Li, Dongjiang, et al.
Pubblicazione: (2025)
di: Li, Dongjiang, et al.
Pubblicazione: (2025)
SemHiTok: A Unified Image Tokenizer via Semantic-Guided Hierarchical Codebook for Multimodal Understanding and Generation
di: Chen, Zisheng, et al.
Pubblicazione: (2025)
di: Chen, Zisheng, et al.
Pubblicazione: (2025)
E0: Enhancing Generalization and Fine-Grained Control in VLA Models via Tweedie Discrete Diffusion
di: Zhan, Zhihao, et al.
Pubblicazione: (2025)
di: Zhan, Zhihao, et al.
Pubblicazione: (2025)
Learning Diffusion Policy from Primitive Skills for Robot Manipulation
di: Gu, Zhihao, et al.
Pubblicazione: (2026)
di: Gu, Zhihao, et al.
Pubblicazione: (2026)
EchoVLA: Synergistic Declarative Memory for VLA-Driven Mobile Manipulation
di: Lin, Min, et al.
Pubblicazione: (2025)
di: Lin, Min, et al.
Pubblicazione: (2025)
MoManipVLA: Transferring Vision-language-action Models for General Mobile Manipulation
di: Wu, Zhenyu, et al.
Pubblicazione: (2025)
di: Wu, Zhenyu, et al.
Pubblicazione: (2025)
Atomic-Probe Governance for Skill Updates in Compositional Robot Policies
di: Qin, Xue, et al.
Pubblicazione: (2026)
di: Qin, Xue, et al.
Pubblicazione: (2026)
Atomic Action Slicing: Planner-Aligned Options for Generalist VLA Agents
di: Tabakov, Stefan, et al.
Pubblicazione: (2025)
di: Tabakov, Stefan, et al.
Pubblicazione: (2025)
Sample-Efficient Robot Skill Learning for Construction Tasks: Benchmarking Hierarchical Reinforcement Learning and Vision-Language-Action VLA Model
di: Hu, Zhaofeng, et al.
Pubblicazione: (2025)
di: Hu, Zhaofeng, et al.
Pubblicazione: (2025)
TransMamba: Fast Universal Architecture Adaption from Transformers to Mamba
di: Chen, Xiuwei, et al.
Pubblicazione: (2025)
di: Chen, Xiuwei, et al.
Pubblicazione: (2025)
ProgVLA: Progress-Aware Robot Manipulation Skill Learning
di: Kim, Seungsu, et al.
Pubblicazione: (2026)
di: Kim, Seungsu, et al.
Pubblicazione: (2026)
MoTo: A Zero-shot Plug-in Interaction-aware Navigation for General Mobile Manipulation
di: Wu, Zhenyu, et al.
Pubblicazione: (2025)
di: Wu, Zhenyu, et al.
Pubblicazione: (2025)
SELF-VLA: A Skill Enhanced Agentic Vision-Language-Action Framework for Contact-Rich Disassembly
di: Liu, Chang, et al.
Pubblicazione: (2026)
di: Liu, Chang, et al.
Pubblicazione: (2026)
Testing Large Language Models on Driving Theory Knowledge and Skills for Connected Autonomous Vehicles
di: Tang, Zuoyin, et al.
Pubblicazione: (2024)
di: Tang, Zuoyin, et al.
Pubblicazione: (2024)
BlockVLA: Accelerating Autoregressive VLA via Block Diffusion Finetuning
di: Wang, Ruiheng, et al.
Pubblicazione: (2026)
di: Wang, Ruiheng, et al.
Pubblicazione: (2026)
SwiftVLA: Unlocking Spatiotemporal Dynamics for Lightweight VLA Models at Minimal Overhead
di: Ni, Chaojun, et al.
Pubblicazione: (2025)
di: Ni, Chaojun, et al.
Pubblicazione: (2025)
NavCoT: Boosting LLM-Based Vision-and-Language Navigation via Learning Disentangled Reasoning
di: Lin, Bingqian, et al.
Pubblicazione: (2024)
di: Lin, Bingqian, et al.
Pubblicazione: (2024)
ShapeGen: Robotic Data Generation for Category-Level Manipulation
di: Wang, Yirui, et al.
Pubblicazione: (2026)
di: Wang, Yirui, et al.
Pubblicazione: (2026)
GazeVLA: Learning Human Intention for Robotic Manipulation
di: Li, Chengyang, et al.
Pubblicazione: (2026)
di: Li, Chengyang, et al.
Pubblicazione: (2026)
Unleashing Humanoid Reaching Potential via Real-world-Ready Skill Space
di: Zhang, Zhikai, et al.
Pubblicazione: (2025)
di: Zhang, Zhikai, et al.
Pubblicazione: (2025)
VLA-ATTC: Adaptive Test-Time Compute for VLA Models with Relative Action Critic Model
di: Li, Wenhao, et al.
Pubblicazione: (2026)
di: Li, Wenhao, et al.
Pubblicazione: (2026)
RePO-VLA: Recovery-Driven Policy Optimization for Vision-Language-Action Models
di: Liufu, Weijia, et al.
Pubblicazione: (2026)
di: Liufu, Weijia, et al.
Pubblicazione: (2026)
Embodied Instruction Following in Unknown Environments
di: Wu, Zhenyu, et al.
Pubblicazione: (2024)
di: Wu, Zhenyu, et al.
Pubblicazione: (2024)
Review of Essential Generic Technologies for Visual Perception in Underground Coal Mine Robots
di: Yuxin Du, et al.
Pubblicazione: (2026)
di: Yuxin Du, et al.
Pubblicazione: (2026)
Automated Hybrid Reward Scheduling via Large Language Models for Robotic Skill Learning
di: Huang, Changxin, et al.
Pubblicazione: (2025)
di: Huang, Changxin, et al.
Pubblicazione: (2025)
AtomVLA: Scalable Post-Training for Robotic Manipulation via Predictive Latent World Models
di: Sun, Xiaoquan, et al.
Pubblicazione: (2026)
di: Sun, Xiaoquan, et al.
Pubblicazione: (2026)
On-the-Fly VLA Adaptation via Test-Time Reinforcement Learning
di: Liu, Changyu, et al.
Pubblicazione: (2026)
di: Liu, Changyu, et al.
Pubblicazione: (2026)
PIVOT-R: Primitive-Driven Waypoint-Aware World Model for Robotic Manipulation
di: Zhang, Kaidong, et al.
Pubblicazione: (2024)
di: Zhang, Kaidong, et al.
Pubblicazione: (2024)
ManualVLA: A Unified VLA Model for Chain-of-Thought Manual Generation and Robotic Manipulation
di: Gu, Chenyang, et al.
Pubblicazione: (2025)
di: Gu, Chenyang, et al.
Pubblicazione: (2025)
SG-Nav: Online 3D Scene Graph Prompting for LLM-based Zero-shot Object Navigation
di: Yin, Hang, et al.
Pubblicazione: (2024)
di: Yin, Hang, et al.
Pubblicazione: (2024)
TIC-VLA: A Think-in-Control Vision-Language-Action Model for Robot Navigation in Dynamic Environments
di: Huang, Zhiyu, et al.
Pubblicazione: (2026)
di: Huang, Zhiyu, et al.
Pubblicazione: (2026)
MoE-DP: An MoE-Enhanced Diffusion Policy for Robust Long-Horizon Robotic Manipulation with Skill Decomposition and Failure Recovery
di: Cheng, Baiye, et al.
Pubblicazione: (2025)
di: Cheng, Baiye, et al.
Pubblicazione: (2025)
SkillVLA: Tackling Combinatorial Diversity in Dual-Arm Manipulation via Skill Reuse
di: Zhai, Xuanran, et al.
Pubblicazione: (2026)
di: Zhai, Xuanran, et al.
Pubblicazione: (2026)
Documenti analoghi
-
RoboPearls: Editable Video Simulation for Robot Manipulation
di: Tang, Tao, et al.
Pubblicazione: (2025) -
Compose by Focus: Scene Graph-based Atomic Skills
di: Qi, Han, et al.
Pubblicazione: (2025) -
VidMan: Exploiting Implicit Dynamics from Video Diffusion Model for Effective Robot Manipulation
di: Wen, Youpeng, et al.
Pubblicazione: (2024) -
Learning Semantic Atomic Skills for Multi-Task Robotic Manipulation
di: Zhu, Yihang, et al.
Pubblicazione: (2025) -
RoBridge: A Hierarchical Architecture Bridging Cognition and Execution for General Robotic Manipulation
di: Zhang, Kaidong, et al.
Pubblicazione: (2025)