A Touch, Vision, and Language Dataset for Multimodal Alignment
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Fu, Letian, Datta, Gaurav, Huang, Huang, Panitch, William Chung-Ho, Drake, Jaimyn, Ortiz, Joseph, Mukadam, Mustafa, Lambeta, Mike, Calandra, Roberto, Goldberg, Ken |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
In-Context Imitation Learning via Next-Token Prediction
von: Fu, Letian, et al.
Veröffentlicht: (2024)
von: Fu, Letian, et al.
Veröffentlicht: (2024)
Enhance Vision-based Tactile Sensors via Dynamic Illumination and Image Fusion
von: Redkin, Artemii, et al.
Veröffentlicht: (2025)
von: Redkin, Artemii, et al.
Veröffentlicht: (2025)
Learning Gentle Grasping Using Vision, Sound, and Touch
von: Nakahara, Ken, et al.
Veröffentlicht: (2025)
von: Nakahara, Ken, et al.
Veröffentlicht: (2025)
OTTER: A Vision-Language-Action Model with Text-Aware Visual Feature Extraction
von: Huang, Huang, et al.
Veröffentlicht: (2025)
von: Huang, Huang, et al.
Veröffentlicht: (2025)
Tactile Beyond Pixels: Multisensory Touch Representations for Robot Manipulation
von: Higuera, Carolina, et al.
Veröffentlicht: (2025)
von: Higuera, Carolina, et al.
Veröffentlicht: (2025)
Automating Deformable Gasket Assembly
von: Adebola, Simeon, et al.
Veröffentlicht: (2024)
von: Adebola, Simeon, et al.
Veröffentlicht: (2024)
From Simple to Complex Skills: The Case of In-Hand Object Reorientation
von: Qi, Haozhi, et al.
Veröffentlicht: (2025)
von: Qi, Haozhi, et al.
Veröffentlicht: (2025)
Digitizing Touch with an Artificial Multimodal Fingertip
von: Lambeta, Mike, et al.
Veröffentlicht: (2024)
von: Lambeta, Mike, et al.
Veröffentlicht: (2024)
FogROS2-Config: Optimizing Latency and Cost for Multi-Cloud Robot Applications
von: Chen, Kaiyuan, et al.
Veröffentlicht: (2023)
von: Chen, Kaiyuan, et al.
Veröffentlicht: (2023)
SuFIA: Language-Guided Augmented Dexterity for Robotic Surgical Assistants
von: Moghani, Masoud, et al.
Veröffentlicht: (2024)
von: Moghani, Masoud, et al.
Veröffentlicht: (2024)
The Alignment Ceiling: Objective Mismatch in Reinforcement Learning from Human Feedback
von: Lambert, Nathan, et al.
Veröffentlicht: (2023)
von: Lambert, Nathan, et al.
Veröffentlicht: (2023)
Touch100k: A Large-Scale Touch-Language-Vision Dataset for Touch-Centric Multimodal Representation
von: Cheng, Ning, et al.
Veröffentlicht: (2024)
von: Cheng, Ning, et al.
Veröffentlicht: (2024)
ORBIT-Surgical: An Open-Simulation Framework for Learning Surgical Augmented Dexterity
von: Yu, Qinxi, et al.
Veröffentlicht: (2024)
von: Yu, Qinxi, et al.
Veröffentlicht: (2024)
Sparsh: Self-supervised touch representations for vision-based tactile sensing
von: Higuera, Carolina, et al.
Veröffentlicht: (2024)
von: Higuera, Carolina, et al.
Veröffentlicht: (2024)
Towards Comprehensive Multimodal Perception: Introducing the Touch-Language-Vision Dataset
von: Cheng, Ning, et al.
Veröffentlicht: (2024)
von: Cheng, Ning, et al.
Veröffentlicht: (2024)
Using Fiber Optic Bundles to Miniaturize Vision-Based Tactile Sensors
von: Di, Julia, et al.
Veröffentlicht: (2024)
von: Di, Julia, et al.
Veröffentlicht: (2024)
Robo-DM: Data Management For Large Robot Datasets
von: Chen, Kaiyuan, et al.
Veröffentlicht: (2025)
von: Chen, Kaiyuan, et al.
Veröffentlicht: (2025)
Blox-Net: Generative Design-for-Robot-Assembly Using VLM Supervision, Physics Simulation, and a Robot with Reset
von: Goldberg, Andrew, et al.
Veröffentlicht: (2024)
von: Goldberg, Andrew, et al.
Veröffentlicht: (2024)
Self-supervised perception for tactile skin covered dexterous hands
von: Sharma, Akash, et al.
Veröffentlicht: (2025)
von: Sharma, Akash, et al.
Veröffentlicht: (2025)
The Feeling of Success: Does Touch Sensing Help Predict Grasp Outcomes?
von: Calandra, Roberto, et al.
Veröffentlicht: (2017)
von: Calandra, Roberto, et al.
Veröffentlicht: (2017)
Real2Render2Real: Scaling Robot Data Without Dynamics Simulation or Robot Hardware
von: Yu, Justin, et al.
Veröffentlicht: (2025)
von: Yu, Justin, et al.
Veröffentlicht: (2025)
SGANet: Semantic and Geometric Alignment for Multimodal Multi-view Anomaly Detection
von: Bai, Letian, et al.
Veröffentlicht: (2026)
von: Bai, Letian, et al.
Veröffentlicht: (2026)
3d Quantum Trace Map
von: Panitch, Samuel, et al.
Veröffentlicht: (2024)
von: Panitch, Samuel, et al.
Veröffentlicht: (2024)
Compatibility of quantum trace and UV-IR maps
von: Panitch, Samuel, et al.
Veröffentlicht: (2025)
von: Panitch, Samuel, et al.
Veröffentlicht: (2025)
Gaussian Process-Based Active Exploration Strategies in Vision and Touch
von: Choi, Ho Jin, et al.
Veröffentlicht: (2025)
von: Choi, Ho Jin, et al.
Veröffentlicht: (2025)
DexterityGen: Foundation Controller for Unprecedented Dexterity
von: Yin, Zhao-Heng, et al.
Veröffentlicht: (2025)
von: Yin, Zhao-Heng, et al.
Veröffentlicht: (2025)
Touching a Nerve: Neuroimmune Interactions in Asthma
von: James M. Kornfield, et al.
Veröffentlicht: (2025)
von: James M. Kornfield, et al.
Veröffentlicht: (2025)
TactAlign: Human-to-Robot Policy Transfer via Tactile Alignment
von: Wi, Youngsun, et al.
Veröffentlicht: (2026)
von: Wi, Youngsun, et al.
Veröffentlicht: (2026)
STITCH: Augmented Dexterity for Suture Throws Including Thread Coordination and Handoffs
von: Hari, Kush, et al.
Veröffentlicht: (2024)
von: Hari, Kush, et al.
Veröffentlicht: (2024)
VisGym: Diverse, Customizable, Scalable Environments for Multimodal Agents
von: Wang, Zirui, et al.
Veröffentlicht: (2026)
von: Wang, Zirui, et al.
Veröffentlicht: (2026)
TaskMet: Task-Driven Metric Learning for Model Learning
von: Bansal, Dishank, et al.
Veröffentlicht: (2023)
von: Bansal, Dishank, et al.
Veröffentlicht: (2023)
Implicit Multimodal Alignment: On the Generalization of Frozen LLMs to Multimodal Inputs
von: Shukor, Mustafa, et al.
Veröffentlicht: (2024)
von: Shukor, Mustafa, et al.
Veröffentlicht: (2024)
Optimising interpreter mediated dementia assessments
von: Naheed Mukadam, et al.
Veröffentlicht: (2025)
von: Naheed Mukadam, et al.
Veröffentlicht: (2025)
TouchFusion: Multimodal Wristband Sensing for Ubiquitous Touch Interactions
von: Whitmire, Eric, et al.
Veröffentlicht: (2026)
von: Whitmire, Eric, et al.
Veröffentlicht: (2026)
Vania Markarian, Universidad, revolución y dólares. Dos estudios sobre la Guerra Fría cultural en Uruguay durante los Sesenta, Montevideo, Penguin Random House, 2020, 262 páginas
von: Benedetta Calandra
Veröffentlicht: (2021)
von: Benedetta Calandra
Veröffentlicht: (2021)
¿Se puede medir un edificio con un barómetro ?
von: Alexander Calandra
Veröffentlicht: (2001)
von: Alexander Calandra
Veröffentlicht: (2001)
VLA-Touch: Enhancing Vision-Language-Action Models with Dual-Level Tactile Feedback
von: Bi, Jianxin, et al.
Veröffentlicht: (2025)
von: Bi, Jianxin, et al.
Veröffentlicht: (2025)
Lifelong LERF: Local 3D Semantic Inventory Monitoring Using FogROS2
von: Rashid, Adam, et al.
Veröffentlicht: (2024)
von: Rashid, Adam, et al.
Veröffentlicht: (2024)
The ARL Special Collections Initiative.
von: Hewitt, Joe A., et al.
Veröffentlicht: (2003)
von: Hewitt, Joe A., et al.
Veröffentlicht: (2003)
Effective Explanations for Belief-Desire-Intention Robots: When and What to Explain
von: Wang, Cong, et al.
Veröffentlicht: (2025)
von: Wang, Cong, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
In-Context Imitation Learning via Next-Token Prediction
von: Fu, Letian, et al.
Veröffentlicht: (2024) -
Enhance Vision-based Tactile Sensors via Dynamic Illumination and Image Fusion
von: Redkin, Artemii, et al.
Veröffentlicht: (2025) -
Learning Gentle Grasping Using Vision, Sound, and Touch
von: Nakahara, Ken, et al.
Veröffentlicht: (2025) -
OTTER: A Vision-Language-Action Model with Text-Aware Visual Feature Extraction
von: Huang, Huang, et al.
Veröffentlicht: (2025) -
Tactile Beyond Pixels: Multisensory Touch Representations for Robot Manipulation
von: Higuera, Carolina, et al.
Veröffentlicht: (2025)