Saved in:
| Main Authors: | Zhou, Xunlan, Chen, Xuanlin, Zhang, Shaowei, Wan, ShengHua, Hu, Xiaohai, Yuan, Lei, Zhan, De-chuan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2602.15872 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
UDON: Uncertainty-weighted Distributed Optimization for Multi-Robot Neural Implicit Mapping under Extreme Communication Constraints
by: Zhao, Hongrui, et al.
Published: (2025)
by: Zhao, Hongrui, et al.
Published: (2025)
TACO: Temporal Consensus Optimization for Continual Neural Mapping
by: Zhou, Xunlan, et al.
Published: (2026)
by: Zhou, Xunlan, et al.
Published: (2026)
VLATest: Testing and Evaluating Vision-Language-Action Models for Robotic Manipulation
by: Wang, Zhijie, et al.
Published: (2024)
by: Wang, Zhijie, et al.
Published: (2024)
LADEV: A Language-Driven Testing and Evaluation Platform for Vision-Language-Action Models in Robotic Manipulation
by: Wang, Zhijie, et al.
Published: (2024)
by: Wang, Zhijie, et al.
Published: (2024)
Learned Slip-Detection-Severity Framework using Tactile Deformation Field Feedback for Robotic Manipulation
by: Jawale, Neel, et al.
Published: (2024)
by: Jawale, Neel, et al.
Published: (2024)
Reflective Planning: Vision-Language Models for Multi-Stage Long-Horizon Robotic Manipulation
by: Feng, Yunhai, et al.
Published: (2025)
by: Feng, Yunhai, et al.
Published: (2025)
SG-VLA: Learning Spatially-Grounded Vision-Language-Action Models for Mobile Manipulation
by: Tu, Ruisen, et al.
Published: (2026)
by: Tu, Ruisen, et al.
Published: (2026)
Graph-Fused Vision-Language-Action for Policy Reasoning in Multi-Arm Robotic Manipulation
by: Li, Shunlei, et al.
Published: (2025)
by: Li, Shunlei, et al.
Published: (2025)
Stable Language Guidance for Vision-Language-Action Models
by: Zhan, Zhihao, et al.
Published: (2026)
by: Zhan, Zhihao, et al.
Published: (2026)
FG-CLTP: Fine-Grained Contrastive Language Tactile Pretraining for Robotic Manipulation
by: Ma, Wenxuan, et al.
Published: (2026)
by: Ma, Wenxuan, et al.
Published: (2026)
VTLA: Vision-Tactile-Language-Action Model with Preference Learning for Insertion Manipulation
by: Zhang, Chaofan, et al.
Published: (2025)
by: Zhang, Chaofan, et al.
Published: (2025)
3D Affordance Keypoint Detection for Robotic Manipulation
by: Liu, Zhiyang, et al.
Published: (2025)
by: Liu, Zhiyang, et al.
Published: (2025)
Play to the Score: Stage-Guided Dynamic Multi-Sensory Fusion for Robotic Manipulation
by: Feng, Ruoxuan, et al.
Published: (2024)
by: Feng, Ruoxuan, et al.
Published: (2024)
Gaze2Act: Gaze-Conditioned Vision-Language-Action Policies for Interactive Robot Manipulation
by: Zuo, Kuangji, et al.
Published: (2026)
by: Zuo, Kuangji, et al.
Published: (2026)
Robotic Manipulation is Vision-to-Geometry Mapping ($f(v) \rightarrow G$): Vision-Geometry Backbones over Language and Video Models
by: Song, Zijian, et al.
Published: (2026)
by: Song, Zijian, et al.
Published: (2026)
Task-oriented Robotic Manipulation with Vision Language Models
by: Guran, Nurhan Bulus, et al.
Published: (2024)
by: Guran, Nurhan Bulus, et al.
Published: (2024)
RoboFlamingo-Plus: Fusion of Depth and RGB Perception with Vision-Language Models for Enhanced Robotic Manipulation
by: Wang, Sheng
Published: (2025)
by: Wang, Sheng
Published: (2025)
CubeRobot: Grounding Language in Rubik's Cube Manipulation via Vision-Language Model
by: Wang, Feiyang, et al.
Published: (2025)
by: Wang, Feiyang, et al.
Published: (2025)
VLMPC: Vision-Language Model Predictive Control for Robotic Manipulation
by: Zhao, Wentao, et al.
Published: (2024)
by: Zhao, Wentao, et al.
Published: (2024)
Online Robot Navigation and Manipulation with Distilled Vision-Language Models
by: Liu, Kangcheng
Published: (2024)
by: Liu, Kangcheng
Published: (2024)
AHA: A Vision-Language-Model for Detecting and Reasoning Over Failures in Robotic Manipulation
by: Duan, Jiafei, et al.
Published: (2024)
by: Duan, Jiafei, et al.
Published: (2024)
VLBiMan: Vision-Language Anchored One-Shot Demonstration Enables Generalizable Bimanual Robotic Manipulation
by: Zhou, Huayi, et al.
Published: (2025)
by: Zhou, Huayi, et al.
Published: (2025)
When would Vision-Proprioception Policies Fail in Robotic Manipulation?
by: Lu, Jingxian, et al.
Published: (2026)
by: Lu, Jingxian, et al.
Published: (2026)
ResponsibleRobotBench: Benchmarking Responsible Robot Manipulation using Multi-modal Large Language Models
by: Zhang, Lei, et al.
Published: (2025)
by: Zhang, Lei, et al.
Published: (2025)
Green-VLA: Staged Vision-Language-Action Model for Generalist Robots
by: Apanasevich, I., et al.
Published: (2026)
by: Apanasevich, I., et al.
Published: (2026)
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors
by: Huang, Haifeng, et al.
Published: (2025)
by: Huang, Haifeng, et al.
Published: (2025)
BagelVLA: Enhancing Long-Horizon Manipulation via Interleaved Vision-Language-Action Generation
by: Hu, Yucheng, et al.
Published: (2026)
by: Hu, Yucheng, et al.
Published: (2026)
MoLe-VLA: Dynamic Layer-skipping Vision Language Action Model via Mixture-of-Layers for Efficient Robot Manipulation
by: Zhang, Rongyu, et al.
Published: (2025)
by: Zhang, Rongyu, et al.
Published: (2025)
Confusion-Aware In-Context-Learning for Vision-Language Models in Robotic Manipulation
by: He, Yayun, et al.
Published: (2026)
by: He, Yayun, et al.
Published: (2026)
Splat-MOVER: Multi-Stage, Open-Vocabulary Robotic Manipulation via Editable Gaussian Splatting
by: Shorinwa, Ola, et al.
Published: (2024)
by: Shorinwa, Ola, et al.
Published: (2024)
Data Scaling Laws in Imitation Learning for Robotic Manipulation
by: Hu, Yingdong, et al.
Published: (2024)
by: Hu, Yingdong, et al.
Published: (2024)
Interpretable Robotic Manipulation from Language
by: Zheng, Boyuan, et al.
Published: (2024)
by: Zheng, Boyuan, et al.
Published: (2024)
HERMES: Human-to-Robot Embodied Learning from Multi-Source Motion Data for Mobile Dexterous Manipulation
by: Yuan, Zhecheng, et al.
Published: (2025)
by: Yuan, Zhecheng, et al.
Published: (2025)
Manipulate-Anything: Automating Real-World Robots using Vision-Language Models
by: Duan, Jiafei, et al.
Published: (2024)
by: Duan, Jiafei, et al.
Published: (2024)
Robot Tape Manipulation for 3D Printing
by: Tushar, Nahid, et al.
Published: (2024)
by: Tushar, Nahid, et al.
Published: (2024)
Learning Manipulation Skills through Robot Chain-of-Thought with Sparse Failure Guidance
by: Zhang, Kaifeng, et al.
Published: (2024)
by: Zhang, Kaifeng, et al.
Published: (2024)
OHP-RL: Online Human Preference as Guidance in Reinforcement Learning for Robot Manipulation
by: Mo, Yunyang, et al.
Published: (2026)
by: Mo, Yunyang, et al.
Published: (2026)
VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation
by: Zhao, Wei, et al.
Published: (2025)
by: Zhao, Wei, et al.
Published: (2025)
VMTS: Vision-Assisted Teacher-Student Reinforcement Learning for Multi-Terrain Locomotion in Bipedal Robots
by: Chen, Fu, et al.
Published: (2025)
by: Chen, Fu, et al.
Published: (2025)
Empowering Embodied Manipulation: A Bimanual-Mobile Robot Manipulation Dataset for Household Tasks
by: Zhang, Tianle, et al.
Published: (2024)
by: Zhang, Tianle, et al.
Published: (2024)
Similar Items
-
UDON: Uncertainty-weighted Distributed Optimization for Multi-Robot Neural Implicit Mapping under Extreme Communication Constraints
by: Zhao, Hongrui, et al.
Published: (2025) -
TACO: Temporal Consensus Optimization for Continual Neural Mapping
by: Zhou, Xunlan, et al.
Published: (2026) -
VLATest: Testing and Evaluating Vision-Language-Action Models for Robotic Manipulation
by: Wang, Zhijie, et al.
Published: (2024) -
LADEV: A Language-Driven Testing and Evaluation Platform for Vision-Language-Action Models in Robotic Manipulation
by: Wang, Zhijie, et al.
Published: (2024) -
Learned Slip-Detection-Severity Framework using Tactile Deformation Field Feedback for Robotic Manipulation
by: Jawale, Neel, et al.
Published: (2024)