VLA-Touch: Enhancing Vision-Language-Action Models with Dual-Level Tactile Feedback
Fuente:
arXiv
Saved in:
| Main Authors: | Bi, Jianxin, Ma, Kevin Yuchen, Hao, Ce, Shou, Mike Zheng, Soh, Harold |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
EVOLVE-VLA: Test-Time Training from Environment Feedback for Vision-Language-Action Models
by: Bai, Zechen, et al.
Published: (2025)
by: Bai, Zechen, et al.
Published: (2025)
AT-VLA: Adaptive Tactile Injection for Enhanced Feedback Reaction in Vision-Language-Action Models
by: Li, Xiaoqi, et al.
Published: (2026)
by: Li, Xiaoqi, et al.
Published: (2026)
Tactile-VLA: Unlocking Vision-Language-Action Model's Physical Knowledge for Tactile Generalization
by: Huang, Jialei, et al.
Published: (2025)
by: Huang, Jialei, et al.
Published: (2025)
Imitation Learning with Limited Actions via Diffusion Planners and Deep Koopman Controllers
by: Bi, Jianxin, et al.
Published: (2024)
by: Bi, Jianxin, et al.
Published: (2024)
Semantic-Contact Fields for Category-Level Generalizable Tactile Tool Manipulation
by: Ma, Kevin Yuchen, et al.
Published: (2026)
by: Ma, Kevin Yuchen, et al.
Published: (2026)
Action Hallucination in Generative Vision-Language-Action Models
by: Soh, Harold, et al.
Published: (2026)
by: Soh, Harold, et al.
Published: (2026)
World-VLA-Loop: Closed-Loop Learning of Video World Model and VLA Policy
by: Liu, Xiaokang, et al.
Published: (2026)
by: Liu, Xiaokang, et al.
Published: (2026)
Diffusion Meets Options: Hierarchical Generative Skill Composition for Temporally-Extended Tasks
by: Feng, Zeyu, et al.
Published: (2024)
by: Feng, Zeyu, et al.
Published: (2024)
CoFreeVLA: Collision-Free Dual-Arm Manipulation via Vision-Language-Action Model and Risk Estimation
by: Zhai, Xuanran, et al.
Published: (2026)
by: Zhai, Xuanran, et al.
Published: (2026)
OpenVLA: An Open-Source Vision-Language-Action Model
by: Kim, Moo Jin, et al.
Published: (2024)
by: Kim, Moo Jin, et al.
Published: (2024)
SkillVLA: Tackling Combinatorial Diversity in Dual-Arm Manipulation via Skill Reuse
by: Zhai, Xuanran, et al.
Published: (2026)
by: Zhai, Xuanran, et al.
Published: (2026)
DeepThinkVLA: Enhancing Reasoning Capability of Vision-Language-Action Models
by: Yin, Cheng, et al.
Published: (2025)
by: Yin, Cheng, et al.
Published: (2025)
DropVLA: An Action-Level Backdoor Attack on Vision-Language-Action Models
by: Xu, Zonghuan, et al.
Published: (2025)
by: Xu, Zonghuan, et al.
Published: (2025)
HyperVLA: Efficient Inference in Vision-Language-Action Models via Hypernetworks
by: Xiong, Zheng, et al.
Published: (2025)
by: Xiong, Zheng, et al.
Published: (2025)
Demonstrating the Octopi-1.5 Visual-Tactile-Language Model
by: Yu, Samson, et al.
Published: (2025)
by: Yu, Samson, et al.
Published: (2025)
StereoVLA: Enhancing Vision-Language-Action Models with Stereo Vision
by: Deng, Shengliang, et al.
Published: (2025)
by: Deng, Shengliang, et al.
Published: (2025)
OmniVLA: An Omni-Modal Vision-Language-Action Model for Robot Navigation
by: Hirose, Noriaki, et al.
Published: (2025)
by: Hirose, Noriaki, et al.
Published: (2025)
SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics
by: Shukor, Mustafa, et al.
Published: (2025)
by: Shukor, Mustafa, et al.
Published: (2025)
DyQ-VLA: Temporal-Dynamic-Aware Quantization for Embodied Vision-Language-Action Models
by: Zheng, Zihao, et al.
Published: (2026)
by: Zheng, Zihao, et al.
Published: (2026)
RobustVLA: Robustness-Aware Reinforcement Post-Training for Vision-Language-Action Models
by: Zhang, Hongyin, et al.
Published: (2025)
by: Zhang, Hongyin, et al.
Published: (2025)
Tactile Modality Fusion for Vision-Language-Action Models
by: Morissette, Charlotte, et al.
Published: (2026)
by: Morissette, Charlotte, et al.
Published: (2026)
AsyncVLA: Asynchronous Flow Matching for Vision-Language-Action Models
by: Jiang, Yuhua, et al.
Published: (2025)
by: Jiang, Yuhua, et al.
Published: (2025)
Agentic-VLA: Efficient Online Adaptation for Vision-Language-Action Models
by: Jin, Ruofan, et al.
Published: (2026)
by: Jin, Ruofan, et al.
Published: (2026)
CRL-VLA: Continual Vision-Language-Action Learning
by: Zeng, Qixin, et al.
Published: (2026)
by: Zeng, Qixin, et al.
Published: (2026)
RationalVLA: A Rational Vision-Language-Action Model with Dual System
by: Song, Wenxuan, et al.
Published: (2025)
by: Song, Wenxuan, et al.
Published: (2025)
TacVLA: Contact-Aware Tactile Fusion for Robust Vision-Language-Action Manipulation
by: Zhang, Kaidi, et al.
Published: (2026)
by: Zhang, Kaidi, et al.
Published: (2026)
Large Language Models as Zero-Shot Human Models for Human-Robot Interaction
by: Zhang, Bowen, et al.
Published: (2023)
by: Zhang, Bowen, et al.
Published: (2023)
Octopi: Object Property Reasoning with Large Tactile-Language Models
by: Yu, Samson, et al.
Published: (2024)
by: Yu, Samson, et al.
Published: (2024)
TaF-VLA: Tactile-Force Alignment in Vision-Language-Action Models for Force-aware Manipulation
by: Huang, Yuzhe, et al.
Published: (2026)
by: Huang, Yuzhe, et al.
Published: (2026)
FlowTouch: View-Invariant Visuo-Tactile Prediction
by: Bien, Seongjin, et al.
Published: (2026)
by: Bien, Seongjin, et al.
Published: (2026)
TwinVLA: Data-Efficient Bimanual Manipulation with Twin Single-Arm Vision-Language-Action Models
by: Im, Hokyun, et al.
Published: (2025)
by: Im, Hokyun, et al.
Published: (2025)
ReMem-VLA: Empowering Vision-Language-Action Model with Memory via Dual-Level Recurrent Queries
by: Li, Hang, et al.
Published: (2026)
by: Li, Hang, et al.
Published: (2026)
DexTouch: Learning to Seek and Manipulate Objects with Tactile Dexterity
by: Lee, Kang-Won, et al.
Published: (2024)
by: Lee, Kang-Won, et al.
Published: (2024)
DISCO: Language-Guided Manipulation with Diffusion Policies and Constrained Inpainting
by: Hao, Ce, et al.
Published: (2024)
by: Hao, Ce, et al.
Published: (2024)
PointVLA: Injecting the 3D World into Vision-Language-Action Models
by: Li, Chengmeng, et al.
Published: (2025)
by: Li, Chengmeng, et al.
Published: (2025)
VLA-R1: Enhancing Reasoning in Vision-Language-Action Models
by: Ye, Angen, et al.
Published: (2025)
by: Ye, Angen, et al.
Published: (2025)
Abstracting Robot Manipulation Skills via Mixture-of-Experts Diffusion Policies
by: Hao, Ce, et al.
Published: (2026)
by: Hao, Ce, et al.
Published: (2026)
CoT-VLA: Visual Chain-of-Thought Reasoning for Vision-Language-Action Models
by: Zhao, Qingqing, et al.
Published: (2025)
by: Zhao, Qingqing, et al.
Published: (2025)
DyGRO-VLA: Cross-Task Scaling of Vision-Language-Action Models via Dynamic Grouped Residual Optimization
by: Lin, Sixu, et al.
Published: (2026)
by: Lin, Sixu, et al.
Published: (2026)
HapticVLA: Contact-Rich Manipulation via Vision-Language-Action Model without Inference-Time Tactile Sensing
by: Gubernatorov, Konstantin, et al.
Published: (2026)
by: Gubernatorov, Konstantin, et al.
Published: (2026)
Similar Items
-
EVOLVE-VLA: Test-Time Training from Environment Feedback for Vision-Language-Action Models
by: Bai, Zechen, et al.
Published: (2025) -
AT-VLA: Adaptive Tactile Injection for Enhanced Feedback Reaction in Vision-Language-Action Models
by: Li, Xiaoqi, et al.
Published: (2026) -
Tactile-VLA: Unlocking Vision-Language-Action Model's Physical Knowledge for Tactile Generalization
by: Huang, Jialei, et al.
Published: (2025) -
Imitation Learning with Limited Actions via Diffusion Planners and Deep Koopman Controllers
by: Bi, Jianxin, et al.
Published: (2024) -
Semantic-Contact Fields for Category-Level Generalizable Tactile Tool Manipulation
by: Ma, Kevin Yuchen, et al.
Published: (2026)