What Foundation Models can Bring for Robot Learning in Manipulation : A Survey
Fuente:
arXiv
Guardado en:
| Autores principales: | Li, Dingzhe, Jin, Yixiang, Sun, Yuhao, A, Yong, Yu, Hongze, Shi, Jun, Hao, Xiaoshuai, Hao, Peng, Liu, Huaping, Li, Xiang, Li, Xinde, Sun, Fuchun, Zhang, Jianwei, Fang, Bin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
EHC-MM: Embodied Holistic Control for Mobile Manipulation
por: Wang, Jiawen, et al.
Publicado: (2024)
por: Wang, Jiawen, et al.
Publicado: (2024)
Smooth Computation without Input Delay: Robust Tube-Based Model Predictive Control for Robot Manipulator Planning
por: Luo, Yu, et al.
Publicado: (2024)
por: Luo, Yu, et al.
Publicado: (2024)
TLA: Tactile-Language-Action Model for Contact-Rich Manipulation
por: Hao, Peng, et al.
Publicado: (2025)
por: Hao, Peng, et al.
Publicado: (2025)
What Really Matters for Robust Multi-Sensor HD Map Construction?
por: Hao, Xiaoshuai, et al.
Publicado: (2025)
por: Hao, Xiaoshuai, et al.
Publicado: (2025)
ASGrasp: Generalizable Transparent Object Reconstruction and 6-DoF Grasp Detection from RGB-D Active Stereo Camera
por: Shi, Jun, et al.
Publicado: (2024)
por: Shi, Jun, et al.
Publicado: (2024)
DAM-VLA: A Dynamic Action Model-Based Vision-Language-Action Framework for Robot Manipulation
por: Peng, Xiongfeng, et al.
Publicado: (2026)
por: Peng, Xiongfeng, et al.
Publicado: (2026)
A Robotic Prosthetic Hand for Computer Mouse Operations
por: Ziming Chen, et al.
Publicado: (2025)
por: Ziming Chen, et al.
Publicado: (2025)
Artificial Skin Based on Visuo‐Tactile Sensing for 3D Shape Reconstruction: Material, Method, and Evaluation
por: Shixin Zhang, et al.
Publicado: (2024)
por: Shixin Zhang, et al.
Publicado: (2024)
Evolving the Complete Muscle: Efficient Morphology-Control Co-design for Musculoskeletal Locomotion
por: Sun, Lidong, et al.
Publicado: (2026)
por: Sun, Lidong, et al.
Publicado: (2026)
Simulation of Optical Tactile Sensors Supporting Slip and Rotation using Path Tracing and IMPM
por: Shen, Zirong, et al.
Publicado: (2024)
por: Shen, Zirong, et al.
Publicado: (2024)
VERM: Leveraging Foundation Models to Create a Virtual Eye for Efficient 3D Robotic Manipulation
por: Chen, Yixiang, et al.
Publicado: (2025)
por: Chen, Yixiang, et al.
Publicado: (2025)
Soft Contact Simulation and Manipulation Learning of Deformable Objects with Vision-based Tactile Sensor
por: Shan, Jianhua, et al.
Publicado: (2024)
por: Shan, Jianhua, et al.
Publicado: (2024)
Transforming Monolithic Foundation Models into Embodied Multi-Agent Architectures for Human-Robot Collaboration
por: Sun, Nan, et al.
Publicado: (2025)
por: Sun, Nan, et al.
Publicado: (2025)
Small Scale Data-Free Knowledge Distillation
por: Liu, He, et al.
Publicado: (2024)
por: Liu, He, et al.
Publicado: (2024)
Multi-segment Soft Robot Control via Deep Koopman-based Model Predictive Control
por: Lv, Lei, et al.
Publicado: (2025)
por: Lv, Lei, et al.
Publicado: (2025)
FlowDreamer: A RGB-D World Model with Flow-based Motion Representations for Robot Manipulation
por: Guo, Jun, et al.
Publicado: (2025)
por: Guo, Jun, et al.
Publicado: (2025)
MapDistill: Boosting Efficient Camera-based HD Map Construction via Camera-LiDAR Fusion Model Distillation
por: Hao, Xiaoshuai, et al.
Publicado: (2024)
por: Hao, Xiaoshuai, et al.
Publicado: (2024)
Pollen‐Biochar‐Based Tactile‐Pain Dual‐Function Sensors for Intelligent Robotics
por: Longwei Li, et al.
Publicado: (2025)
por: Longwei Li, et al.
Publicado: (2025)
Force Measurement Technology of Vision‐Based Tactile Sensor
por: Bin Fang, et al.
Publicado: (2024)
por: Bin Fang, et al.
Publicado: (2024)
Tacchi 2.0: A Low Computational Cost and Comprehensive Dynamic Contact Simulator for Vision-based Tactile Sensors
por: Sun, Yuhao, et al.
Publicado: (2025)
por: Sun, Yuhao, et al.
Publicado: (2025)
Observe Then Act: Asynchronous Active Vision-Action Model for Robotic Manipulation
por: Wang, Guokang, et al.
Publicado: (2024)
por: Wang, Guokang, et al.
Publicado: (2024)
The Multi-Round Diagnostic RAG Framework for Emulating Clinical Reasoning
por: Sun, Penglei, et al.
Publicado: (2025)
por: Sun, Penglei, et al.
Publicado: (2025)
Spatial-Temporal Graph Diffusion Policy with Kinematic Modeling for Bimanual Robotic Manipulation
por: Lv, Qi, et al.
Publicado: (2025)
por: Lv, Qi, et al.
Publicado: (2025)
What Will the Delivery Robots Bring Us Tomorrow?
por: Ronja Kaiser, et al.
Publicado: (2024)
por: Ronja Kaiser, et al.
Publicado: (2024)
Chain-of-Description: What I can understand, I can put into words
por: Guo, Jiaxin, et al.
Publicado: (2025)
por: Guo, Jiaxin, et al.
Publicado: (2025)
Type II Singularities of Lagrangian Mean Curvature Flow with Zero Maslov Class
por: Li, Xiang, et al.
Publicado: (2024)
por: Li, Xiang, et al.
Publicado: (2024)
Tacchi: A Pluggable and Low Computational Cost Elastomer Deformation Simulator for Optical Tactile Sensors
por: Chen, Zixi, et al.
Publicado: (2023)
por: Chen, Zixi, et al.
Publicado: (2023)
A Physical Model-Guided Framework for Underwater Image Enhancement and Depth Estimation
por: Du, Dazhao, et al.
Publicado: (2024)
por: Du, Dazhao, et al.
Publicado: (2024)
A Step Toward World Models: A Survey on Robotic Manipulation
por: Zhang, Peng-Fei, et al.
Publicado: (2025)
por: Zhang, Peng-Fei, et al.
Publicado: (2025)
Design and Analysis of an Origami Robot With Tensegrity Characteristics
por: Peng Ni, et al.
Publicado: (2026)
por: Peng Ni, et al.
Publicado: (2026)
When Vision Meets Touch: A Contemporary Review for Visuotactile Sensors from the Signal Processing Perspective
por: Li, Shoujie, et al.
Publicado: (2024)
por: Li, Shoujie, et al.
Publicado: (2024)
VTLA: Vision-Tactile-Language-Action Model with Preference Learning for Insertion Manipulation
por: Zhang, Chaofan, et al.
Publicado: (2025)
por: Zhang, Chaofan, et al.
Publicado: (2025)
Enhancing Cognition and Explainability of Multimodal Foundation Models with Self-Synthesized Data
por: Shi, Yucheng, et al.
Publicado: (2025)
por: Shi, Yucheng, et al.
Publicado: (2025)
Towards Generalizable Robotic Manipulation in Dynamic Environments
por: Fang, Heng, et al.
Publicado: (2026)
por: Fang, Heng, et al.
Publicado: (2026)
A Survey of Robotic Navigation and Manipulation with Physics Simulators in the Era of Embodied AI
por: Wong, Lik Hang Kenny, et al.
Publicado: (2025)
por: Wong, Lik Hang Kenny, et al.
Publicado: (2025)
CAGE: Causal Attention Enables Data-Efficient Generalizable Robotic Manipulation
por: Xia, Shangning, et al.
Publicado: (2024)
por: Xia, Shangning, et al.
Publicado: (2024)
AI-Generated Text Detection and Classification Based on BERT Deep Learning Algorithm
por: Wang, Hao, et al.
Publicado: (2024)
por: Wang, Hao, et al.
Publicado: (2024)
Siamese Foundation Models for Crystal Structure Prediction
por: Wu, Liming, et al.
Publicado: (2025)
por: Wu, Liming, et al.
Publicado: (2025)
IIANet: An Intra- and Inter-Modality Attention Network for Audio-Visual Speech Separation
por: Li, Kai, et al.
Publicado: (2023)
por: Li, Kai, et al.
Publicado: (2023)
M4-BLIP: Advancing Multi-Modal Media Manipulation Detection through Face-Enhanced Local Analysis
por: Wu, Hang, et al.
Publicado: (2025)
por: Wu, Hang, et al.
Publicado: (2025)
Ejemplares similares
-
EHC-MM: Embodied Holistic Control for Mobile Manipulation
por: Wang, Jiawen, et al.
Publicado: (2024) -
Smooth Computation without Input Delay: Robust Tube-Based Model Predictive Control for Robot Manipulator Planning
por: Luo, Yu, et al.
Publicado: (2024) -
TLA: Tactile-Language-Action Model for Contact-Rich Manipulation
por: Hao, Peng, et al.
Publicado: (2025) -
What Really Matters for Robust Multi-Sensor HD Map Construction?
por: Hao, Xiaoshuai, et al.
Publicado: (2025) -
ASGrasp: Generalizable Transparent Object Reconstruction and 6-DoF Grasp Detection from RGB-D Active Stereo Camera
por: Shi, Jun, et al.
Publicado: (2024)