Saved in:
| Main Authors: | Li, Chengyang, Xiong, Kaiyi, Xu, Yuan, Qian, Lei, Wang, Yizhou, Zhu, Wentao |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2604.22615 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Gaze-Guided Robotic Vascular Ultrasound Leveraging Human Intention Estimation
by: Bi, Yuan, et al.
Published: (2025)
by: Bi, Yuan, et al.
Published: (2025)
Learning Human-Aware Robot Policies for Adaptive Assistance
by: Qin, Jason, et al.
Published: (2024)
by: Qin, Jason, et al.
Published: (2024)
FAVLA: A Force-Adaptive Fast-Slow VLA model for Contact-Rich Robotic Manipulation
by: Li, Yao, et al.
Published: (2026)
by: Li, Yao, et al.
Published: (2026)
Gaze-Based Intention Recognition for Human-Robot Collaboration
by: Belcamino, Valerio, et al.
Published: (2024)
by: Belcamino, Valerio, et al.
Published: (2024)
IntentionVLA: Generalizable and Efficient Embodied Intention Reasoning for Human-Robot Interaction
by: Chen, Yandu, et al.
Published: (2025)
by: Chen, Yandu, et al.
Published: (2025)
Learning Human-Intention Priors from Large-Scale Human Demonstrations for Robotic Manipulation
by: Xie, Yifan, et al.
Published: (2026)
by: Xie, Yifan, et al.
Published: (2026)
SimVLA: A Simple VLA Baseline for Robotic Manipulation
by: Luo, Yuankai, et al.
Published: (2026)
by: Luo, Yuankai, et al.
Published: (2026)
HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model
by: Zhu, Xiang, et al.
Published: (2026)
by: Zhu, Xiang, et al.
Published: (2026)
Gaze2Act: Gaze-Conditioned Vision-Language-Action Policies for Interactive Robot Manipulation
by: Zuo, Kuangji, et al.
Published: (2026)
by: Zuo, Kuangji, et al.
Published: (2026)
RynnVLA-001: Using Human Demonstrations to Improve Robot Manipulation
by: Jiang, Yuming, et al.
Published: (2025)
by: Jiang, Yuming, et al.
Published: (2025)
InternVLA-A1: Unifying Understanding, Generation and Action for Robotic Manipulation
by: Cai, Junhao, et al.
Published: (2026)
by: Cai, Junhao, et al.
Published: (2026)
AC^2-VLA: Action-Context-Aware Adaptive Computation in Vision-Language-Action Models for Efficient Robotic Manipulation
by: Yu, Wenda, et al.
Published: (2026)
by: Yu, Wenda, et al.
Published: (2026)
BitVLA: 1-bit Vision-Language-Action Models for Robotics Manipulation
by: Wang, Hongyu, et al.
Published: (2025)
by: Wang, Hongyu, et al.
Published: (2025)
Gaze-Guided Task Decomposition for Imitation Learning in Robotic Manipulation
by: Takizawa, Ryo, et al.
Published: (2025)
by: Takizawa, Ryo, et al.
Published: (2025)
TGM-VLA: Task-Guided Mixup for Sampling-Efficient and Robust Robotic Manipulation
by: Pu, Fanqi, et al.
Published: (2026)
by: Pu, Fanqi, et al.
Published: (2026)
Early Detection of Human Handover Intentions in Human-Robot Collaboration: Comparing EEG, Gaze, and Hand Motion
by: Khanna, Parag, et al.
Published: (2025)
by: Khanna, Parag, et al.
Published: (2025)
Humanizing Robot Gaze Shifts: A Framework for Natural Gaze Shifts in Humanoid Robots
by: Wei, Jingchao, et al.
Published: (2026)
by: Wei, Jingchao, et al.
Published: (2026)
PrimitiveVLA: Learning Reusable Motion Primitives for Efficient and Generalizable Robotic Manipulation
by: Li, Yutai, et al.
Published: (2026)
by: Li, Yutai, et al.
Published: (2026)
ManualVLA: A Unified VLA Model for Chain-of-Thought Manual Generation and Robotic Manipulation
by: Gu, Chenyang, et al.
Published: (2025)
by: Gu, Chenyang, et al.
Published: (2025)
A Novel Robot Hand with Hoeckens Linkages and Soft Phalanges for Scooping and Self-Adaptive Grasping in Environmental Constraints
by: Guo, Wentao, et al.
Published: (2025)
by: Guo, Wentao, et al.
Published: (2025)
FedVLA: Federated Vision-Language-Action Learning with Dual Gating Mixture-of-Experts for Robotic Manipulation
by: Miao, Cui, et al.
Published: (2025)
by: Miao, Cui, et al.
Published: (2025)
ProgVLA: Progress-Aware Robot Manipulation Skill Learning
by: Kim, Seungsu, et al.
Published: (2026)
by: Kim, Seungsu, et al.
Published: (2026)
Move-Then-Operate: Behavioral Phasing for Human-Like Robotic Manipulation
by: Xu, Haoming, et al.
Published: (2026)
by: Xu, Haoming, et al.
Published: (2026)
Human Gaze and Head Rotation during Navigation, Exploration and Object Manipulation in Shared Environments with Robots
by: Schreiter, Tim, et al.
Published: (2024)
by: Schreiter, Tim, et al.
Published: (2024)
AnchorVLA4D: an Anchor-Based Spatial-Temporal Vision-Language-Action Model for Robotic Manipulation
by: Zhu, Juan, et al.
Published: (2026)
by: Zhu, Juan, et al.
Published: (2026)
VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning
by: Lu, Guanxing, et al.
Published: (2025)
by: Lu, Guanxing, et al.
Published: (2025)
Intent at a Glance: Gaze-Guided Robotic Manipulation via Foundation Models
by: Tay, Tracey Yee Hsin, et al.
Published: (2026)
by: Tay, Tracey Yee Hsin, et al.
Published: (2026)
ABot-M0: VLA Foundation Model for Robotic Manipulation with Action Manifold Learning
by: Yang, Yandan, et al.
Published: (2026)
by: Yang, Yandan, et al.
Published: (2026)
VLA Model-Expert Collaboration for Bi-directional Manipulation Learning
by: Xiang, Tian-Yu, et al.
Published: (2025)
by: Xiang, Tian-Yu, et al.
Published: (2025)
Learning Multimodal Confidence for Intention Recognition in Human-Robot Interaction
by: Zhao, Xiyuan, et al.
Published: (2024)
by: Zhao, Xiyuan, et al.
Published: (2024)
Predicting the Intention to Interact with a Service Robot:the Role of Gaze Cues
by: Arreghini, Simone, et al.
Published: (2024)
by: Arreghini, Simone, et al.
Published: (2024)
Interleave-VLA: Enhancing Robot Manipulation with Interleaved Image-Text Instructions
by: Fan, Cunxin, et al.
Published: (2025)
by: Fan, Cunxin, et al.
Published: (2025)
AtomVLA: Scalable Post-Training for Robotic Manipulation via Predictive Latent World Models
by: Sun, Xiaoquan, et al.
Published: (2026)
by: Sun, Xiaoquan, et al.
Published: (2026)
Non-Markovian Long-Horizon Robot Manipulation via Keyframe Chaining
by: Chen, Yipeng, et al.
Published: (2026)
by: Chen, Yipeng, et al.
Published: (2026)
ProcVLM: Learning Procedure-Grounded Progress Rewards for Robotic Manipulation
by: Feng, Youhe, et al.
Published: (2026)
by: Feng, Youhe, et al.
Published: (2026)
Compressor-VLA: Instruction-Guided Visual Token Compression for Efficient Robotic Manipulation
by: Gao, Juntao, et al.
Published: (2025)
by: Gao, Juntao, et al.
Published: (2025)
MimicDreamer: Aligning Human and Robot Demonstrations for Scalable VLA Training
by: Li, Haoyun, et al.
Published: (2025)
by: Li, Haoyun, et al.
Published: (2025)
DexCanvas: Bridging Human Demonstrations and Robot Learning for Dexterous Manipulation
by: Xu, Xinyue, et al.
Published: (2025)
by: Xu, Xinyue, et al.
Published: (2025)
TinyVLA: Towards Fast, Data-Efficient Vision-Language-Action Models for Robotic Manipulation
by: Wen, Junjie, et al.
Published: (2024)
by: Wen, Junjie, et al.
Published: (2024)
OmniVLA: Physically-Grounded Multimodal VLA with Unified Multi-Sensor Perception for Robotic Manipulation
by: Guo, Heyu, et al.
Published: (2025)
by: Guo, Heyu, et al.
Published: (2025)
Similar Items
-
Gaze-Guided Robotic Vascular Ultrasound Leveraging Human Intention Estimation
by: Bi, Yuan, et al.
Published: (2025) -
Learning Human-Aware Robot Policies for Adaptive Assistance
by: Qin, Jason, et al.
Published: (2024) -
FAVLA: A Force-Adaptive Fast-Slow VLA model for Contact-Rich Robotic Manipulation
by: Li, Yao, et al.
Published: (2026) -
Gaze-Based Intention Recognition for Human-Robot Collaboration
by: Belcamino, Valerio, et al.
Published: (2024) -
IntentionVLA: Generalizable and Efficient Embodied Intention Reasoning for Human-Robot Interaction
by: Chen, Yandu, et al.
Published: (2025)