ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Zhiyuan, Zhao, Yongqiang, Luo, Shan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ViTaS: Visual Tactile Soft Fusion Contrastive Learning for Visuomotor Learning
by: Tian, Yufeng, et al.
Published: (2026)
by: Tian, Yufeng, et al.
Published: (2026)
ControlTac: Force- and Position-Controlled Tactile Data Augmentation with a Single Reference Image
by: Luo, Dongyu, et al.
Published: (2025)
by: Luo, Dongyu, et al.
Published: (2025)
ViTac-Tracing: Visual-Tactile Imitation Learning of Deformable Object Tracing
by: Zhao, Yongqiang, et al.
Published: (2026)
by: Zhao, Yongqiang, et al.
Published: (2026)
Learning to Feel the Future: DreamTacVLA for Contact-Rich Manipulation
by: Ye, Guo, et al.
Published: (2025)
by: Ye, Guo, et al.
Published: (2025)
CEDex: Cross-Embodiment Dexterous Grasp Generation at Scale from Human-like Contact Representations
by: Wu, Zhiyuan, et al.
Published: (2025)
by: Wu, Zhiyuan, et al.
Published: (2025)
Simultaneous Tactile-Visual Perception for Learning Multimodal Robot Manipulation
by: Li, Yuyang, et al.
Published: (2025)
by: Li, Yuyang, et al.
Published: (2025)
Transparent Fragments Contour Estimation via Visual-Tactile Fusion for Autonomous Reassembly
by: Lin, Qihao, et al.
Published: (2026)
by: Lin, Qihao, et al.
Published: (2026)
Collaborative Representation Learning for Alignment of Tactile, Language, and Vision Modalities
by: Zhou, Yiyun, et al.
Published: (2025)
by: Zhou, Yiyun, et al.
Published: (2025)
Binding Touch to Everything: Learning Unified Multimodal Tactile Representations
by: Yang, Fengyu, et al.
Published: (2024)
by: Yang, Fengyu, et al.
Published: (2024)
VAT: Vision Action Transformer by Unlocking Full Representation of ViT
by: Li, Wenhao, et al.
Published: (2025)
by: Li, Wenhao, et al.
Published: (2025)
AlignDrive: Aligned Lateral-Longitudinal Planning for End-to-End Autonomous Driving
by: Wu, Yanhao, et al.
Published: (2026)
by: Wu, Yanhao, et al.
Published: (2026)
Sensor-Invariant Tactile Representation
by: Gupta, Harsh, et al.
Published: (2025)
by: Gupta, Harsh, et al.
Published: (2025)
U-ViLAR: Uncertainty-Aware Visual Localization for Autonomous Driving via Differentiable Association and Registration
by: Li, Xiaofan, et al.
Published: (2025)
by: Li, Xiaofan, et al.
Published: (2025)
Touch-GS: Visual-Tactile Supervised 3D Gaussian Splatting
by: Swann, Aiden, et al.
Published: (2024)
by: Swann, Aiden, et al.
Published: (2024)
Symmetry-Aware Fusion of Vision and Tactile Sensing via Bilateral Force Priors for Robotic Manipulation
by: Lee, Wonju, et al.
Published: (2026)
by: Lee, Wonju, et al.
Published: (2026)
LiteViLNet: Lightweight Vision-LiDAR Fusion Network for Efficient Road Segmentation
by: Peng, Daojie, et al.
Published: (2026)
by: Peng, Daojie, et al.
Published: (2026)
ViT-VS: On the Applicability of Pretrained Vision Transformer Features for Generalizable Visual Servoing
by: Scherl, Alessandro, et al.
Published: (2025)
by: Scherl, Alessandro, et al.
Published: (2025)
Tactile Modality Fusion for Vision-Language-Action Models
by: Morissette, Charlotte, et al.
Published: (2026)
by: Morissette, Charlotte, et al.
Published: (2026)
Transferable Tactile Transformers for Representation Learning Across Diverse Sensors and Tasks
by: Zhao, Jialiang, et al.
Published: (2024)
by: Zhao, Jialiang, et al.
Published: (2024)
AnyTouch 2: General Optical Tactile Representation Learning For Dynamic Tactile Perception
by: Feng, Ruoxuan, et al.
Published: (2026)
by: Feng, Ruoxuan, et al.
Published: (2026)
ViTaSCOPE: Visuo-tactile Implicit Representation for In-hand Pose and Extrinsic Contact Estimation
by: Lee, Jayjun, et al.
Published: (2025)
by: Lee, Jayjun, et al.
Published: (2025)
OccCylindrical: Multi-Modal Fusion with Cylindrical Representation for 3D Semantic Occupancy Prediction
by: Ming, Zhenxing, et al.
Published: (2025)
by: Ming, Zhenxing, et al.
Published: (2025)
High-Precision Transformer-Based Visual Servoing for Humanoid Robots in Aligning Tiny Objects
by: Xue, Jialong, et al.
Published: (2025)
by: Xue, Jialong, et al.
Published: (2025)
Tacchi 2.0: A Low Computational Cost and Comprehensive Dynamic Contact Simulator for Vision-based Tactile Sensors
by: Sun, Yuhao, et al.
Published: (2025)
by: Sun, Yuhao, et al.
Published: (2025)
OccFusion: Multi-Sensor Fusion Framework for 3D Semantic Occupancy Prediction
by: Ming, Zhenxing, et al.
Published: (2024)
by: Ming, Zhenxing, et al.
Published: (2024)
Aligning Knowledge Graph with Visual Perception for Object-goal Navigation
by: Xu, Nuo, et al.
Published: (2024)
by: Xu, Nuo, et al.
Published: (2024)
VTAO-BiManip: Masked Visual-Tactile-Action Pre-training with Object Understanding for Bimanual Dexterous Manipulation
by: Sun, Zhengnan, et al.
Published: (2025)
by: Sun, Zhengnan, et al.
Published: (2025)
Fusion-Poly: A Polyhedral Framework Based on Spatial-Temporal Fusion for 3D Multi-Object Tracking
by: Wu, Xian, et al.
Published: (2026)
by: Wu, Xian, et al.
Published: (2026)
ViTacGen: Robotic Pushing with Vision-to-Touch Generation
by: Wu, Zhiyuan, et al.
Published: (2025)
by: Wu, Zhiyuan, et al.
Published: (2025)
FeelAnyForce: Estimating Contact Force Feedback from Tactile Sensation for Vision-Based Tactile Sensors
by: Shahidzadeh, Amir-Hossein, et al.
Published: (2024)
by: Shahidzadeh, Amir-Hossein, et al.
Published: (2024)
Task Success Prediction for Open-Vocabulary Manipulation Based on Multi-Level Aligned Representations
by: Goko, Miyu, et al.
Published: (2024)
by: Goko, Miyu, et al.
Published: (2024)
Visuo-Tactile Object Pose Estimation for a Multi-Finger Robot Hand with Low-Resolution In-Hand Tactile Sensing
by: Mack, Lukas, et al.
Published: (2025)
by: Mack, Lukas, et al.
Published: (2025)
Robotic Eye-in-hand Visual Servo Axially Aligning Nasopharyngeal Swabs with the Nasal Cavity
by: Lee, Peter Q., et al.
Published: (2024)
by: Lee, Peter Q., et al.
Published: (2024)
OCRA: Object-Centric Learning with 3D and Tactile Priors for Human-to-Robot Action Transfer
by: Wang, Kuanning, et al.
Published: (2026)
by: Wang, Kuanning, et al.
Published: (2026)
Latent Representations for Visual Proprioception in Inexpensive Robots
by: Sheikholeslami, Sahara, et al.
Published: (2025)
by: Sheikholeslami, Sahara, et al.
Published: (2025)
AcTExplore: Active Tactile Exploration of Unknown Objects
by: Shahidzadeh, Amir-Hossein, et al.
Published: (2023)
by: Shahidzadeh, Amir-Hossein, et al.
Published: (2023)
FastViDAR: Real-Time Omnidirectional Depth Estimation via Alternative Hierarchical Attention
by: Zhao, Hangtian, et al.
Published: (2025)
by: Zhao, Hangtian, et al.
Published: (2025)
MapGCLR: Geospatial Contrastive Learning of Representations for Online Vectorized HD Map Construction
by: Merkert, Jonas, et al.
Published: (2026)
by: Merkert, Jonas, et al.
Published: (2026)
Taccel: Scaling Up Vision-based Tactile Robotics via High-performance GPU Simulation
by: Li, Yuyang, et al.
Published: (2025)
by: Li, Yuyang, et al.
Published: (2025)
TLA: Tactile-Language-Action Model for Contact-Rich Manipulation
by: Hao, Peng, et al.
Published: (2025)
by: Hao, Peng, et al.
Published: (2025)
Similar Items
-
ViTaS: Visual Tactile Soft Fusion Contrastive Learning for Visuomotor Learning
by: Tian, Yufeng, et al.
Published: (2026) -
ControlTac: Force- and Position-Controlled Tactile Data Augmentation with a Single Reference Image
by: Luo, Dongyu, et al.
Published: (2025) -
ViTac-Tracing: Visual-Tactile Imitation Learning of Deformable Object Tracing
by: Zhao, Yongqiang, et al.
Published: (2026) -
Learning to Feel the Future: DreamTacVLA for Contact-Rich Manipulation
by: Ye, Guo, et al.
Published: (2025) -
CEDex: Cross-Embodiment Dexterous Grasp Generation at Scale from Human-like Contact Representations
by: Wu, Zhiyuan, et al.
Published: (2025)