Touch begins where vision ends: Generalizable policies for contact-rich manipulation
Fuente:
arXiv
Saved in:
| Main Authors: | Zhao, Zifan, Haldar, Siddhant, Cui, Jinda, Pinto, Lerrel, Bhirangi, Raunaq |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
P3-PO: Prescriptive Point Priors for Visuo-Spatial Generalization of Robot Policies
by: Levy, Mara, et al.
Published: (2024)
by: Levy, Mara, et al.
Published: (2024)
Learning Precise, Contact-Rich Manipulation through Uncalibrated Tactile Skins
by: Pattabiraman, Venkatesh, et al.
Published: (2024)
by: Pattabiraman, Venkatesh, et al.
Published: (2024)
DynaMo: In-Domain Dynamics Pretraining for Visuo-Motor Control
by: Cui, Zichen Jeff, et al.
Published: (2024)
by: Cui, Zichen Jeff, et al.
Published: (2024)
EgoZero: Robot Learning from Smart Glasses
by: Liu, Vincent, et al.
Published: (2025)
by: Liu, Vincent, et al.
Published: (2025)
AnySkin: Plug-and-play Skin Sensing for Robotic Touch
by: Bhirangi, Raunaq, et al.
Published: (2024)
by: Bhirangi, Raunaq, et al.
Published: (2024)
eFlesh: Highly customizable Magnetic Touch Sensing using Cut-Cell Microstructures
by: Pattabiraman, Venkatesh, et al.
Published: (2025)
by: Pattabiraman, Venkatesh, et al.
Published: (2025)
Point Policy: Unifying Observations and Actions with Key Points for Robot Manipulation
by: Haldar, Siddhant, et al.
Published: (2025)
by: Haldar, Siddhant, et al.
Published: (2025)
Feel the Force: Contact-Driven Learning from Humans
by: Adeniji, Ademi, et al.
Published: (2025)
by: Adeniji, Ademi, et al.
Published: (2025)
GLOVER: Generalizable Open-Vocabulary Affordance Reasoning for Task-Oriented Grasping
by: Ma, Teli, et al.
Published: (2024)
by: Ma, Teli, et al.
Published: (2024)
CLIP-Fields: Weakly Supervised Semantic Fields for Robotic Memory
by: Shafiullah, Nur Muhammad Mahi, et al.
Published: (2022)
by: Shafiullah, Nur Muhammad Mahi, et al.
Published: (2022)
GenH2R: Learning Generalizable Human-to-Robot Handover via Scalable Simulation, Demonstration, and Imitation
by: Wang, Zifan, et al.
Published: (2024)
by: Wang, Zifan, et al.
Published: (2024)
Bridging the Human to Robot Dexterity Gap through Object-Oriented Rewards
by: Guzey, Irmak, et al.
Published: (2024)
by: Guzey, Irmak, et al.
Published: (2024)
TouchAnything: Diffusion-Guided 3D Reconstruction from Sparse Robot Touches
by: Gu, Langzhe, et al.
Published: (2026)
by: Gu, Langzhe, et al.
Published: (2026)
Cross-Sensor Touch Generation
by: Rodriguez, Samanta, et al.
Published: (2025)
by: Rodriguez, Samanta, et al.
Published: (2025)
Afford-X: Generalizable and Slim Affordance Reasoning for Task-oriented Manipulation
by: Zhu, Xiaomeng, et al.
Published: (2025)
by: Zhu, Xiaomeng, et al.
Published: (2025)
Object Pose Estimation through Dexterous Touch
by: Shahidzadeh, Amir-Hossein, et al.
Published: (2025)
by: Shahidzadeh, Amir-Hossein, et al.
Published: (2025)
BAKU: An Efficient Transformer for Multi-Task Policy Learning
by: Haldar, Siddhant, et al.
Published: (2024)
by: Haldar, Siddhant, et al.
Published: (2024)
A Touch, Vision, and Language Dataset for Multimodal Alignment
by: Fu, Letian, et al.
Published: (2024)
by: Fu, Letian, et al.
Published: (2024)
Binding Touch to Everything: Learning Unified Multimodal Tactile Representations
by: Yang, Fengyu, et al.
Published: (2024)
by: Yang, Fengyu, et al.
Published: (2024)
FetchBot: Learning Generalizable Object Fetching in Cluttered Scenes via Zero-Shot Sim2Real
by: Liu, Weiheng, et al.
Published: (2025)
by: Liu, Weiheng, et al.
Published: (2025)
PseudoTouch: Efficiently Imaging the Surface Feel of Objects for Robotic Manipulation
by: Röfer, Adrian, et al.
Published: (2024)
by: Röfer, Adrian, et al.
Published: (2024)
Touch-GS: Visual-Tactile Supervised 3D Gaussian Splatting
by: Swann, Aiden, et al.
Published: (2024)
by: Swann, Aiden, et al.
Published: (2024)
Towards Comprehensive Multimodal Perception: Introducing the Touch-Language-Vision Dataset
by: Cheng, Ning, et al.
Published: (2024)
by: Cheng, Ning, et al.
Published: (2024)
EgoTouch: On-Body Touch Input Using AR/VR Headset Cameras
by: Mollyn, Vimal, et al.
Published: (2025)
by: Mollyn, Vimal, et al.
Published: (2025)
Anyview: Generalizable Indoor 3D Object Detection with Variable Frames
by: Wu, Zhenyu, et al.
Published: (2023)
by: Wu, Zhenyu, et al.
Published: (2023)
End-to-End LiDAR optimization for 3D point cloud registration
by: Katyan, Siddhant, et al.
Published: (2026)
by: Katyan, Siddhant, et al.
Published: (2026)
Affordance-R1: Reinforcement Learning for Generalizable Affordance Reasoning in Multimodal Large Language Model
by: Wang, Hanqing, et al.
Published: (2025)
by: Wang, Hanqing, et al.
Published: (2025)
Efficient Training of Generalizable Visuomotor Policies via Control-Aware Augmentation
by: Zhao, Yinuo, et al.
Published: (2024)
by: Zhao, Yinuo, et al.
Published: (2024)
Vision-based robot manipulation of transparent liquid containers in a laboratory setting
by: Schober, Daniel, et al.
Published: (2024)
by: Schober, Daniel, et al.
Published: (2024)
Next Best Sense: Guiding Vision and Touch with FisherRF for 3D Gaussian Splatting
by: Strong, Matthew, et al.
Published: (2024)
by: Strong, Matthew, et al.
Published: (2024)
RLCNet: An end-to-end deep learning framework for simultaneous online calibration of LiDAR, RADAR, and Camera
by: Cholakkal, Hafeez Husain, et al.
Published: (2025)
by: Cholakkal, Hafeez Husain, et al.
Published: (2025)
TrajDiff: End-to-end Autonomous Driving without Perception Annotation
by: Gui, Xingtai, et al.
Published: (2025)
by: Gui, Xingtai, et al.
Published: (2025)
DriveGPT4: Interpretable End-to-end Autonomous Driving via Large Language Model
by: Xu, Zhenhua, et al.
Published: (2023)
by: Xu, Zhenhua, et al.
Published: (2023)
Bridge Thinking and Acting: Unleashing Physical Potential of VLM with Generalizable Action Expert
by: Liu, Mingyu, et al.
Published: (2025)
by: Liu, Mingyu, et al.
Published: (2025)
Generalizable Image Repair for Robust Visual Control
by: Sobolewski, Carson, et al.
Published: (2025)
by: Sobolewski, Carson, et al.
Published: (2025)
Towards Generalizable Robotic Manipulation in Dynamic Environments
by: Fang, Heng, et al.
Published: (2026)
by: Fang, Heng, et al.
Published: (2026)
Goal-conditioned dual-action imitation learning for dexterous dual-arm robot manipulation
by: Kim, Heecheol, et al.
Published: (2022)
by: Kim, Heecheol, et al.
Published: (2022)
MonoPlane: Exploiting Monocular Geometric Cues for Generalizable 3D Plane Reconstruction
by: Zhao, Wang, et al.
Published: (2024)
by: Zhao, Wang, et al.
Published: (2024)
StaMo: Unsupervised Learning of Generalizable Robot Motion from Compact State Representation
by: Liu, Mingyu, et al.
Published: (2025)
by: Liu, Mingyu, et al.
Published: (2025)
See What Matters: Differentiable Grid Sample Pruning for Generalizable Vision-Language-Action Model
by: Feng, Yixu, et al.
Published: (2026)
by: Feng, Yixu, et al.
Published: (2026)
Similar Items
-
P3-PO: Prescriptive Point Priors for Visuo-Spatial Generalization of Robot Policies
by: Levy, Mara, et al.
Published: (2024) -
Learning Precise, Contact-Rich Manipulation through Uncalibrated Tactile Skins
by: Pattabiraman, Venkatesh, et al.
Published: (2024) -
DynaMo: In-Domain Dynamics Pretraining for Visuo-Motor Control
by: Cui, Zichen Jeff, et al.
Published: (2024) -
EgoZero: Robot Learning from Smart Glasses
by: Liu, Vincent, et al.
Published: (2025) -
AnySkin: Plug-and-play Skin Sensing for Robotic Touch
by: Bhirangi, Raunaq, et al.
Published: (2024)