SELF-VLA: A Skill Enhanced Agentic Vision-Language-Action Framework for Contact-Rich Disassembly
Fuente:
arXiv
Guardado en:
| Autores principales: | Liu, Chang, Tian, Sibo, Liang, Xiao, Zheng, Minghui |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Vision-Language-Action Models for Selective Robotic Disassembly: A Case Study on Critical Component Extraction from Desktops
por: Liu, Chang, et al.
Publicado: (2025)
por: Liu, Chang, et al.
Publicado: (2025)
PrediFlow: A Flow-Based Prediction-Refinement Framework for Real-Time Human Motion Prediction in Human-Robot Collaboration
por: Tian, Sibo, et al.
Publicado: (2025)
por: Tian, Sibo, et al.
Publicado: (2025)
Warm-Starting Optimization-Based Motion Planning for Robotic Manipulators via Point Cloud-Conditioned Flow Matching
por: Tian, Sibo, et al.
Publicado: (2025)
por: Tian, Sibo, et al.
Publicado: (2025)
PaCo-VLA: Passivity-Shielded Compliance Prior for Contact-Rich Vision-Language-Action Manipulation
por: Cao, Haofan, et al.
Publicado: (2026)
por: Cao, Haofan, et al.
Publicado: (2026)
A Recurrent Neural Network Enhanced Unscented Kalman Filter for Human Motion Prediction
por: Liu, Wansong, et al.
Publicado: (2024)
por: Liu, Wansong, et al.
Publicado: (2024)
FD-VLA: Force-Distilled Vision-Language-Action Model for Contact-Rich Manipulation
por: Zhao, Ruiteng, et al.
Publicado: (2026)
por: Zhao, Ruiteng, et al.
Publicado: (2026)
Bayesian-Optimized One-Step Diffusion Model with Knowledge Distillation for Real-Time 3D Human Motion Prediction
por: Tian, Sibo, et al.
Publicado: (2024)
por: Tian, Sibo, et al.
Publicado: (2024)
HapticVLA: Contact-Rich Manipulation via Vision-Language-Action Model without Inference-Time Tactile Sensing
por: Gubernatorov, Konstantin, et al.
Publicado: (2026)
por: Gubernatorov, Konstantin, et al.
Publicado: (2026)
Learning When to See and When to Feel: Adaptive Vision-Torque Fusion for Contact-Aware Manipulation
por: Lei, Jiuzhou, et al.
Publicado: (2026)
por: Lei, Jiuzhou, et al.
Publicado: (2026)
StereoVLA: Enhancing Vision-Language-Action Models with Stereo Vision
por: Deng, Shengliang, et al.
Publicado: (2025)
por: Deng, Shengliang, et al.
Publicado: (2025)
VLA^2: Empowering Vision-Language-Action Models with an Agentic Framework for Unseen Concept Manipulation
por: Zhao, Han, et al.
Publicado: (2025)
por: Zhao, Han, et al.
Publicado: (2025)
DeGrip: A Compact Cable-driven Robotic Gripper for Desktop Disassembly
por: Zhang, Bihao, et al.
Publicado: (2025)
por: Zhang, Bihao, et al.
Publicado: (2025)
Integrating Uncertainty-Aware Human Motion Prediction into Graph-Based Manipulator Motion Planning
por: Liu, Wansong, et al.
Publicado: (2024)
por: Liu, Wansong, et al.
Publicado: (2024)
CompliantVLA-adaptor: VLM-Guided Variable Impedance Action for Safe Contact-Rich Manipulation
por: Zhang, Heng, et al.
Publicado: (2026)
por: Zhang, Heng, et al.
Publicado: (2026)
MoS-VLA: A Vision-Language-Action Model with One-Shot Skill Adaptation
por: Zhao, Ruihan, et al.
Publicado: (2025)
por: Zhao, Ruihan, et al.
Publicado: (2025)
TacVLA: Contact-Aware Tactile Fusion for Robust Vision-Language-Action Manipulation
por: Zhang, Kaidi, et al.
Publicado: (2026)
por: Zhang, Kaidi, et al.
Publicado: (2026)
AIR-VLA: Vision-Language-Action Systems for Aerial Manipulation
por: Sun, Jianli, et al.
Publicado: (2026)
por: Sun, Jianli, et al.
Publicado: (2026)
RAISE: A Robot-Assisted Selective Disassembly and Sorting System for End-of-Life Phones
por: Liu, Chang, et al.
Publicado: (2025)
por: Liu, Chang, et al.
Publicado: (2025)
KG-Planner: Knowledge-Informed Graph Neural Planning for Collaborative Manipulators
por: Liu, Wansong, et al.
Publicado: (2024)
por: Liu, Wansong, et al.
Publicado: (2024)
Agentic-VLA: Efficient Online Adaptation for Vision-Language-Action Models
por: Jin, Ruofan, et al.
Publicado: (2026)
por: Jin, Ruofan, et al.
Publicado: (2026)
ACoT-VLA: Action Chain-of-Thought for Vision-Language-Action Models
por: Zhong, Linqing, et al.
Publicado: (2026)
por: Zhong, Linqing, et al.
Publicado: (2026)
Audio-VLA: Adding Contact Audio Perception to Vision-Language-Action Model for Robotic Manipulation
por: Wei, Xiangyi, et al.
Publicado: (2025)
por: Wei, Xiangyi, et al.
Publicado: (2025)
VLA-R1: Enhancing Reasoning in Vision-Language-Action Models
por: Ye, Angen, et al.
Publicado: (2025)
por: Ye, Angen, et al.
Publicado: (2025)
RoVLA: Multi-Consistency Constraints for Robust Vision-Language-Action Models
por: Luo, Jingzhou, et al.
Publicado: (2026)
por: Luo, Jingzhou, et al.
Publicado: (2026)
DAM-VLA: A Dynamic Action Model-Based Vision-Language-Action Framework for Robot Manipulation
por: Peng, Xiongfeng, et al.
Publicado: (2026)
por: Peng, Xiongfeng, et al.
Publicado: (2026)
MergeVLA: Cross-Skill Model Merging Toward a Generalist Vision-Language-Action Agent
por: Fu, Yuxia, et al.
Publicado: (2025)
por: Fu, Yuxia, et al.
Publicado: (2025)
VLA-Touch: Enhancing Vision-Language-Action Models with Dual-Level Tactile Feedback
por: Bi, Jianxin, et al.
Publicado: (2025)
por: Bi, Jianxin, et al.
Publicado: (2025)
HiF-VLA: Hindsight, Insight and Foresight through Motion Representation for Vision-Language-Action Models
por: Lin, Minghui, et al.
Publicado: (2025)
por: Lin, Minghui, et al.
Publicado: (2025)
FPC-VLA: A Vision-Language-Action Framework with a Supervisor for Failure Prediction and Correction
por: Yang, Yifan, et al.
Publicado: (2025)
por: Yang, Yifan, et al.
Publicado: (2025)
Adaptive Motion Planning via Contact-Based Intent Inference for Human-Robot Collaboration
por: Song, Jiurun, et al.
Publicado: (2025)
por: Song, Jiurun, et al.
Publicado: (2025)
RLinf-VLA: A Unified and Efficient Framework for Reinforcement Learning of Vision-Language-Action Models
por: Zang, Hongzhi, et al.
Publicado: (2025)
por: Zang, Hongzhi, et al.
Publicado: (2025)
PAPO-VLA: Planning-Aware Policy Optimization for Vision-Language-Action Models
por: Guo, Peizheng, et al.
Publicado: (2026)
por: Guo, Peizheng, et al.
Publicado: (2026)
RationalVLA: A Rational Vision-Language-Action Model with Dual System
por: Song, Wenxuan, et al.
Publicado: (2025)
por: Song, Wenxuan, et al.
Publicado: (2025)
OpenVLA: An Open-Source Vision-Language-Action Model
por: Kim, Moo Jin, et al.
Publicado: (2024)
por: Kim, Moo Jin, et al.
Publicado: (2024)
Grasping Force Control and Adaptation for a Cable-Driven Robotic Hand
por: Mountain, Eric, et al.
Publicado: (2024)
por: Mountain, Eric, et al.
Publicado: (2024)
Sample-Efficient Robot Skill Learning for Construction Tasks: Benchmarking Hierarchical Reinforcement Learning and Vision-Language-Action VLA Model
por: Hu, Zhaofeng, et al.
Publicado: (2025)
por: Hu, Zhaofeng, et al.
Publicado: (2025)
DexGraspVLA: A Vision-Language-Action Framework Towards General Dexterous Grasping
por: Zhong, Yifan, et al.
Publicado: (2025)
por: Zhong, Yifan, et al.
Publicado: (2025)
VP-VLA: Visual Prompting as an Interface for Vision-Language-Action Models
por: Wang, Zixuan, et al.
Publicado: (2026)
por: Wang, Zixuan, et al.
Publicado: (2026)
AT-VLA: Adaptive Tactile Injection for Enhanced Feedback Reaction in Vision-Language-Action Models
por: Li, Xiaoqi, et al.
Publicado: (2026)
por: Li, Xiaoqi, et al.
Publicado: (2026)
DeepThinkVLA: Enhancing Reasoning Capability of Vision-Language-Action Models
por: Yin, Cheng, et al.
Publicado: (2025)
por: Yin, Cheng, et al.
Publicado: (2025)
Ejemplares similares
-
Vision-Language-Action Models for Selective Robotic Disassembly: A Case Study on Critical Component Extraction from Desktops
por: Liu, Chang, et al.
Publicado: (2025) -
PrediFlow: A Flow-Based Prediction-Refinement Framework for Real-Time Human Motion Prediction in Human-Robot Collaboration
por: Tian, Sibo, et al.
Publicado: (2025) -
Warm-Starting Optimization-Based Motion Planning for Robotic Manipulators via Point Cloud-Conditioned Flow Matching
por: Tian, Sibo, et al.
Publicado: (2025) -
PaCo-VLA: Passivity-Shielded Compliance Prior for Contact-Rich Vision-Language-Action Manipulation
por: Cao, Haofan, et al.
Publicado: (2026) -
A Recurrent Neural Network Enhanced Unscented Kalman Filter for Human Motion Prediction
por: Liu, Wansong, et al.
Publicado: (2024)