Simultaneous Tactile-Visual Perception for Learning Multimodal Robot Manipulation
Fuente:
arXiv
Guardado en:
| Autores principales: | Li, Yuyang, Chen, Yinghan, Zhao, Zihang, Li, Puhao, Liu, Tengyu, Huang, Siyuan, Zhu, Yixin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Taccel: Scaling Up Vision-based Tactile Robotics via High-performance GPU Simulation
por: Li, Yuyang, et al.
Publicado: (2025)
por: Li, Yuyang, et al.
Publicado: (2025)
ManipTrans: Efficient Dexterous Bimanual Manipulation Transfer via Residual Learning
por: Li, Kailin, et al.
Publicado: (2025)
por: Li, Kailin, et al.
Publicado: (2025)
Ag2Manip: Learning Novel Manipulation Skills with Agent-Agnostic Visual and Action Representations
por: Li, Puhao, et al.
Publicado: (2024)
por: Li, Puhao, et al.
Publicado: (2024)
Grasp Multiple Objects with One Hand
por: Li, Yuyang, et al.
Publicado: (2023)
por: Li, Yuyang, et al.
Publicado: (2023)
Ag2x2: Robust Agent-Agnostic Visual Representations for Zero-Shot Bimanual Manipulation
por: Xiong, Ziyin, et al.
Publicado: (2025)
por: Xiong, Ziyin, et al.
Publicado: (2025)
GWM: Towards Scalable Gaussian World Models for Robotic Manipulation
por: Lu, Guanxing, et al.
Publicado: (2025)
por: Lu, Guanxing, et al.
Publicado: (2025)
AnySkill: Learning Open-Vocabulary Physical Skill for Interactive Agents
por: Cui, Jieming, et al.
Publicado: (2024)
por: Cui, Jieming, et al.
Publicado: (2024)
GROVE: A Generalized Reward for Learning Open-Vocabulary Physical Skill
por: Cui, Jieming, et al.
Publicado: (2025)
por: Cui, Jieming, et al.
Publicado: (2025)
Afford-X: Generalizable and Slim Affordance Reasoning for Task-oriented Manipulation
por: Zhu, Xiaomeng, et al.
Publicado: (2025)
por: Zhu, Xiaomeng, et al.
Publicado: (2025)
MetaScenes: Towards Automated Replica Creation for Real-world 3D Scans
por: Yu, Huangyue, et al.
Publicado: (2025)
por: Yu, Huangyue, et al.
Publicado: (2025)
Multimodal Perception System for Real Open Environment
por: Sha, Yuyang
Publicado: (2024)
por: Sha, Yuyang
Publicado: (2024)
RoboView-Bias: Benchmarking Visual Bias in Embodied Agents for Robotic Manipulation
por: Liu, Enguang, et al.
Publicado: (2025)
por: Liu, Enguang, et al.
Publicado: (2025)
StereoPolicy: Improving Robotic Manipulation Policies via Stereo Perception
por: Han, Evans, et al.
Publicado: (2026)
por: Han, Evans, et al.
Publicado: (2026)
ActiveUMI: Robotic Manipulation with Active Perception from Robot-Free Human Demonstrations
por: Zeng, Qiyuan, et al.
Publicado: (2025)
por: Zeng, Qiyuan, et al.
Publicado: (2025)
VitaTouch: Property-Aware Vision-Tactile-Language Model for Robotic Quality Inspection in Manufacturing
por: Zong, Junyi, et al.
Publicado: (2026)
por: Zong, Junyi, et al.
Publicado: (2026)
PhysPart: Physically Plausible Part Completion for Interactable Objects
por: Luo, Rundong, et al.
Publicado: (2024)
por: Luo, Rundong, et al.
Publicado: (2024)
Manipulation as in Simulation: Enabling Accurate Geometry Perception in Robots
por: Liu, Minghuan, et al.
Publicado: (2025)
por: Liu, Minghuan, et al.
Publicado: (2025)
Move as You Say, Interact as You Can: Language-guided Human Motion Generation with Scene Affordance
por: Wang, Zan, et al.
Publicado: (2024)
por: Wang, Zan, et al.
Publicado: (2024)
VTAO-BiManip: Masked Visual-Tactile-Action Pre-training with Object Understanding for Bimanual Dexterous Manipulation
por: Sun, Zhengnan, et al.
Publicado: (2025)
por: Sun, Zhengnan, et al.
Publicado: (2025)
RoboCerebra: A Large-scale Benchmark for Long-horizon Robotic Manipulation Evaluation
por: Han, Songhao, et al.
Publicado: (2025)
por: Han, Songhao, et al.
Publicado: (2025)
Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy
por: Mandil, Willow, et al.
Publicado: (2023)
por: Mandil, Willow, et al.
Publicado: (2023)
AR-VRM: Imitating Human Motions for Visual Robot Manipulation with Analogical Reasoning
por: Yang, Dejie, et al.
Publicado: (2025)
por: Yang, Dejie, et al.
Publicado: (2025)
OmniVLA: Physically-Grounded Multimodal VLA with Unified Multi-Sensor Perception for Robotic Manipulation
por: Guo, Heyu, et al.
Publicado: (2025)
por: Guo, Heyu, et al.
Publicado: (2025)
ClearDepth: Enhanced Stereo Perception of Transparent Objects for Robotic Manipulation
por: Bai, Kaixin, et al.
Publicado: (2024)
por: Bai, Kaixin, et al.
Publicado: (2024)
Ensuring Force Safety in Vision-Guided Robotic Manipulation via Implicit Tactile Calibration
por: Wei, Lai, et al.
Publicado: (2024)
por: Wei, Lai, et al.
Publicado: (2024)
UniAct: Unified Motion Generation and Action Streaming for Humanoid Robots
por: Jiang, Nan, et al.
Publicado: (2025)
por: Jiang, Nan, et al.
Publicado: (2025)
Binding Touch to Everything: Learning Unified Multimodal Tactile Representations
por: Yang, Fengyu, et al.
Publicado: (2024)
por: Yang, Fengyu, et al.
Publicado: (2024)
Spatially Visual Perception for End-to-End Robotic Learning
por: Davies, Travis, et al.
Publicado: (2024)
por: Davies, Travis, et al.
Publicado: (2024)
TLA: Tactile-Language-Action Model for Contact-Rich Manipulation
por: Hao, Peng, et al.
Publicado: (2025)
por: Hao, Peng, et al.
Publicado: (2025)
Distracted Robot: How Visual Clutter Undermine Robotic Manipulation
por: Rasouli, Amir, et al.
Publicado: (2025)
por: Rasouli, Amir, et al.
Publicado: (2025)
Symmetry-Aware Fusion of Vision and Tactile Sensing via Bilateral Force Priors for Robotic Manipulation
por: Lee, Wonju, et al.
Publicado: (2026)
por: Lee, Wonju, et al.
Publicado: (2026)
SceneVerse: Scaling 3D Vision-Language Learning for Grounded Scene Understanding
por: Jia, Baoxiong, et al.
Publicado: (2024)
por: Jia, Baoxiong, et al.
Publicado: (2024)
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors
por: Huang, Haifeng, et al.
Publicado: (2025)
por: Huang, Haifeng, et al.
Publicado: (2025)
Think Small, Act Big: Primitive Prompt Learning for Lifelong Robot Manipulation
por: Yao, Yuanqi, et al.
Publicado: (2025)
por: Yao, Yuanqi, et al.
Publicado: (2025)
Reduced-order Neural Modeling with Differentiable Simulation for High-Detail Tactile Perception
por: Guo, Yuhu, et al.
Publicado: (2026)
por: Guo, Yuhu, et al.
Publicado: (2026)
Diagnose, Correct, and Learn from Manipulation Failures via Visual Symbols
por: Zeng, Xianchao, et al.
Publicado: (2025)
por: Zeng, Xianchao, et al.
Publicado: (2025)
Imagine2touch: Predictive Tactile Sensing for Robotic Manipulation using Efficient Low-Dimensional Signals
por: Ayad, Abdallah, et al.
Publicado: (2024)
por: Ayad, Abdallah, et al.
Publicado: (2024)
Learning to See and Act: Task-Aware Virtual View Exploration for Robotic Manipulation
por: Bai, Yongjie, et al.
Publicado: (2025)
por: Bai, Yongjie, et al.
Publicado: (2025)
Object-Centric Instruction Augmentation for Robotic Manipulation
por: Wen, Junjie, et al.
Publicado: (2024)
por: Wen, Junjie, et al.
Publicado: (2024)
Genie Envisioner: A Unified World Foundation Platform for Robotic Manipulation
por: Liao, Yue, et al.
Publicado: (2025)
por: Liao, Yue, et al.
Publicado: (2025)
Ejemplares similares
-
Taccel: Scaling Up Vision-based Tactile Robotics via High-performance GPU Simulation
por: Li, Yuyang, et al.
Publicado: (2025) -
ManipTrans: Efficient Dexterous Bimanual Manipulation Transfer via Residual Learning
por: Li, Kailin, et al.
Publicado: (2025) -
Ag2Manip: Learning Novel Manipulation Skills with Agent-Agnostic Visual and Action Representations
por: Li, Puhao, et al.
Publicado: (2024) -
Grasp Multiple Objects with One Hand
por: Li, Yuyang, et al.
Publicado: (2023) -
Ag2x2: Robust Agent-Agnostic Visual Representations for Zero-Shot Bimanual Manipulation
por: Xiong, Ziyin, et al.
Publicado: (2025)