Multimodal Visual-Tactile Representation Learning through Self-Supervised Contrastive Pre-Training
Fuente:
arXiv
Guardado en:
| Autores principales: | Dave, Vedant, Lygerakis, Fotios, Rueckert, Elmar |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
M2CURL: Sample-Efficient Multimodal Reinforcement Learning via Self-Supervised Representation Learning for Robotic Manipulation
por: Lygerakis, Fotios, et al.
Publicado: (2024)
por: Lygerakis, Fotios, et al.
Publicado: (2024)
ED-VAE: Entropy Decomposition of ELBO in Variational Autoencoders
por: Lygerakis, Fotios, et al.
Publicado: (2024)
por: Lygerakis, Fotios, et al.
Publicado: (2024)
ViTaPEs: Visuotactile Position Encodings for Cross-Modal Alignment in Multimodal Transformers
por: Lygerakis, Fotios, et al.
Publicado: (2025)
por: Lygerakis, Fotios, et al.
Publicado: (2025)
Integrating Human Expertise in Continuous Spaces: A Novel Interactive Bayesian Optimization Framework with Preference Expected Improvement
por: Feith, Nikolaus, et al.
Publicado: (2024)
por: Feith, Nikolaus, et al.
Publicado: (2024)
The Temporal Trap: Entanglement in Pre-Trained Visual Representations for Visuomotor Policy Learning
por: Tsagkas, Nikolaos, et al.
Publicado: (2025)
por: Tsagkas, Nikolaos, et al.
Publicado: (2025)
Reactive Diffusion Policy: Slow-Fast Visual-Tactile Policy Learning for Contact-Rich Manipulation
por: Xue, Han, et al.
Publicado: (2025)
por: Xue, Han, et al.
Publicado: (2025)
Tactile MNIST: Benchmarking Active Tactile Perception
por: Schneider, Tim, et al.
Publicado: (2025)
por: Schneider, Tim, et al.
Publicado: (2025)
MILES: Making Imitation Learning Easy with Self-Supervision
por: Papagiannis, Georgios, et al.
Publicado: (2024)
por: Papagiannis, Georgios, et al.
Publicado: (2024)
Learning Parameterized Skills from Demonstrations
por: Gupta, Vedant, et al.
Publicado: (2025)
por: Gupta, Vedant, et al.
Publicado: (2025)
Multimodal Human-Autonomous Agents Interaction Using Pre-Trained Language and Visual Foundation Models
por: Nwankwo, Linus, et al.
Publicado: (2024)
por: Nwankwo, Linus, et al.
Publicado: (2024)
VTouch++: A Multimodal Dataset with Vision-Based Tactile Enhancement for Bimanual Manipulation
por: Hua, Qianxi, et al.
Publicado: (2026)
por: Hua, Qianxi, et al.
Publicado: (2026)
SCIZOR: A Self-Supervised Approach to Data Curation for Large-Scale Imitation Learning
por: Zhang, Yu, et al.
Publicado: (2025)
por: Zhang, Yu, et al.
Publicado: (2025)
VendiRL: A Framework for Self-Supervised Reinforcement Learning of Diversely Diverse Skills
por: Lintunen, Erik M.
Publicado: (2025)
por: Lintunen, Erik M.
Publicado: (2025)
Unsupervised Learning of Efficient Exploration: Pre-training Adaptive Policies via Self-Imposed Goals
por: Pappalardo, Octavio
Publicado: (2026)
por: Pappalardo, Octavio
Publicado: (2026)
3D-ViTac: Learning Fine-Grained Manipulation with Visuo-Tactile Sensing
por: Huang, Binghao, et al.
Publicado: (2024)
por: Huang, Binghao, et al.
Publicado: (2024)
Touch in the Wild: Learning Fine-Grained Manipulation with a Portable Visuo-Tactile Gripper
por: Zhu, Xinyue, et al.
Publicado: (2025)
por: Zhu, Xinyue, et al.
Publicado: (2025)
PRTS: A Primitive Reasoning and Tasking System via Contrastive Representations
por: Zhang, Yang, et al.
Publicado: (2026)
por: Zhang, Yang, et al.
Publicado: (2026)
TADPO: Reinforcement Learning Goes Off-road
por: Wu, Zhouchonghao, et al.
Publicado: (2026)
por: Wu, Zhouchonghao, et al.
Publicado: (2026)
SPRINT: Scalable Policy Pre-Training via Language Instruction Relabeling
por: Zhang, Jesse, et al.
Publicado: (2023)
por: Zhang, Jesse, et al.
Publicado: (2023)
Tactile-based Object Retrieval From Granular Media
por: Xu, Jingxi, et al.
Publicado: (2024)
por: Xu, Jingxi, et al.
Publicado: (2024)
Tactile Estimation of Extrinsic Contact Patch for Stable Placement
por: Ota, Kei, et al.
Publicado: (2023)
por: Ota, Kei, et al.
Publicado: (2023)
Visual Forecasting as a Mid-level Representation for Avoidance
por: Yang, Hsuan-Kung, et al.
Publicado: (2023)
por: Yang, Hsuan-Kung, et al.
Publicado: (2023)
Action Conditioned Tactile Prediction: case study on slip prediction
por: Mandil, Willow, et al.
Publicado: (2022)
por: Mandil, Willow, et al.
Publicado: (2022)
AutoLoop: Fast Visual SLAM Fine-tuning through Agentic Curriculum Learning
por: Lahiany, Assaf, et al.
Publicado: (2025)
por: Lahiany, Assaf, et al.
Publicado: (2025)
Vintix II: Decision Pre-Trained Transformer is a Scalable In-Context Reinforcement Learner
por: Polubarov, Andrei, et al.
Publicado: (2026)
por: Polubarov, Andrei, et al.
Publicado: (2026)
Tightly-Coupled LiDAR-IMU-Leg Odometry with Online Learned Leg Kinematics Incorporating Foot Tactile Information
por: Okawara, Taku, et al.
Publicado: (2025)
por: Okawara, Taku, et al.
Publicado: (2025)
Towards Bio-Inspired Robotic Trajectory Planning via Self-Supervised RNN
por: Cibula, Miroslav, et al.
Publicado: (2025)
por: Cibula, Miroslav, et al.
Publicado: (2025)
Neural Lyapunov Function Approximation with Self-Supervised Reinforcement Learning
por: McCutcheon, Luc, et al.
Publicado: (2025)
por: McCutcheon, Luc, et al.
Publicado: (2025)
Don't Blind Your VLA: Aligning Visual Representations for OOD Generalization
por: Kachaev, Nikita, et al.
Publicado: (2025)
por: Kachaev, Nikita, et al.
Publicado: (2025)
How to Provably Improve Return Conditioned Supervised Learning?
por: Liu, Zhishuai, et al.
Publicado: (2025)
por: Liu, Zhishuai, et al.
Publicado: (2025)
Boundary-to-Region Supervision for Offline Safe Reinforcement Learning
por: Su, Huikang, et al.
Publicado: (2025)
por: Su, Huikang, et al.
Publicado: (2025)
SPLIT: Separating Physical-Contact via Latent Arithmetic in Image-Based Tactile Sensors
por: Amri, Wadhah Zai El, et al.
Publicado: (2026)
por: Amri, Wadhah Zai El, et al.
Publicado: (2026)
Contrastive Conceptor Activation Steering (COAST): Unlocking Vision-Language-Action Models through Hidden States
por: Miao, Miranda Muqing, et al.
Publicado: (2026)
por: Miao, Miranda Muqing, et al.
Publicado: (2026)
Premier-TACO is a Few-Shot Policy Learner: Pretraining Multitask Representation via Temporal Action-Driven Contrastive Loss
por: Zheng, Ruijie, et al.
Publicado: (2024)
por: Zheng, Ruijie, et al.
Publicado: (2024)
Sample Efficient Robot Learning in Supervised Effect Prediction Tasks
por: Eren, Mehmet Arda, et al.
Publicado: (2024)
por: Eren, Mehmet Arda, et al.
Publicado: (2024)
Tactile Memory with Soft Robot: Robust Object Insertion via Masked Encoding and Soft Wrist
por: Kamijo, Tatsuya, et al.
Publicado: (2026)
por: Kamijo, Tatsuya, et al.
Publicado: (2026)
Information-Theoretic Policy Pre-Training with Empowerment
por: Schneider, Moritz, et al.
Publicado: (2025)
por: Schneider, Moritz, et al.
Publicado: (2025)
FlexiTac: A Low-Cost, Open-Source, Scalable Tactile Sensing Solution for Robotic Systems
por: Huang, Binghao, et al.
Publicado: (2026)
por: Huang, Binghao, et al.
Publicado: (2026)
MOTO: Offline Pre-training to Online Fine-tuning for Model-based Robot Learning
por: Rafailov, Rafael, et al.
Publicado: (2024)
por: Rafailov, Rafael, et al.
Publicado: (2024)
Subgraph Gaussian Embedding Contrast for Self-Supervised Graph Representation Learning
por: Xie, Shifeng, et al.
Publicado: (2025)
por: Xie, Shifeng, et al.
Publicado: (2025)
Ejemplares similares
-
M2CURL: Sample-Efficient Multimodal Reinforcement Learning via Self-Supervised Representation Learning for Robotic Manipulation
por: Lygerakis, Fotios, et al.
Publicado: (2024) -
ED-VAE: Entropy Decomposition of ELBO in Variational Autoencoders
por: Lygerakis, Fotios, et al.
Publicado: (2024) -
ViTaPEs: Visuotactile Position Encodings for Cross-Modal Alignment in Multimodal Transformers
por: Lygerakis, Fotios, et al.
Publicado: (2025) -
Integrating Human Expertise in Continuous Spaces: A Novel Interactive Bayesian Optimization Framework with Preference Expected Improvement
por: Feith, Nikolaus, et al.
Publicado: (2024) -
The Temporal Trap: Entanglement in Pre-Trained Visual Representations for Visuomotor Policy Learning
por: Tsagkas, Nikolaos, et al.
Publicado: (2025)