DynaMo: In-Domain Dynamics Pretraining for Visuo-Motor Control
Fuente:
arXiv
Salvato in:
| Autori principali: | Cui, Zichen Jeff, Pan, Hengkai, Iyer, Aadhithya, Haldar, Siddhant, Pinto, Lerrel |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
P3-PO: Prescriptive Point Priors for Visuo-Spatial Generalization of Robot Policies
di: Levy, Mara, et al.
Pubblicazione: (2024)
di: Levy, Mara, et al.
Pubblicazione: (2024)
Touch begins where vision ends: Generalizable policies for contact-rich manipulation
di: Zhao, Zifan, et al.
Pubblicazione: (2025)
di: Zhao, Zifan, et al.
Pubblicazione: (2025)
Adapting Pretrained ViTs with Convolution Injector for Visuo-Motor Control
di: Hwang, Dongyoon, et al.
Pubblicazione: (2024)
di: Hwang, Dongyoon, et al.
Pubblicazione: (2024)
MoDem-V2: Visuo-Motor World Models for Real-World Robot Manipulation
di: Lancaster, Patrick, et al.
Pubblicazione: (2023)
di: Lancaster, Patrick, et al.
Pubblicazione: (2023)
Point Policy: Unifying Observations and Actions with Key Points for Robot Manipulation
di: Haldar, Siddhant, et al.
Pubblicazione: (2025)
di: Haldar, Siddhant, et al.
Pubblicazione: (2025)
OPEN TEACH: A Versatile Teleoperation System for Robotic Manipulation
di: Iyer, Aadhithya, et al.
Pubblicazione: (2024)
di: Iyer, Aadhithya, et al.
Pubblicazione: (2024)
Object-Centric Action-Enhanced Representations for Robot Visuo-Motor Policy Learning
di: Giannakakis, Nikos, et al.
Pubblicazione: (2025)
di: Giannakakis, Nikos, et al.
Pubblicazione: (2025)
V-HOP: Visuo-Haptic 6D Object Pose Tracking
di: Li, Hongyu, et al.
Pubblicazione: (2025)
di: Li, Hongyu, et al.
Pubblicazione: (2025)
DynaRend: Learning 3D Dynamics via Masked Future Rendering for Robotic Manipulation
di: Tian, Jingyi, et al.
Pubblicazione: (2025)
di: Tian, Jingyi, et al.
Pubblicazione: (2025)
BAKU: An Efficient Transformer for Multi-Task Policy Learning
di: Haldar, Siddhant, et al.
Pubblicazione: (2024)
di: Haldar, Siddhant, et al.
Pubblicazione: (2024)
Learning Precise, Contact-Rich Manipulation through Uncalibrated Tactile Skins
di: Pattabiraman, Venkatesh, et al.
Pubblicazione: (2024)
di: Pattabiraman, Venkatesh, et al.
Pubblicazione: (2024)
OK-Robot: What Really Matters in Integrating Open-Knowledge Models for Robotics
di: Liu, Peiqi, et al.
Pubblicazione: (2024)
di: Liu, Peiqi, et al.
Pubblicazione: (2024)
DINO-WM: World Models on Pre-trained Visual Features enable Zero-shot Planning
di: Zhou, Gaoyue, et al.
Pubblicazione: (2024)
di: Zhou, Gaoyue, et al.
Pubblicazione: (2024)
CLIP-Fields: Weakly Supervised Semantic Fields for Robotic Memory
di: Shafiullah, Nur Muhammad Mahi, et al.
Pubblicazione: (2022)
di: Shafiullah, Nur Muhammad Mahi, et al.
Pubblicazione: (2022)
ViPRA: Video Prediction for Robot Actions
di: Routray, Sandeep, et al.
Pubblicazione: (2025)
di: Routray, Sandeep, et al.
Pubblicazione: (2025)
DynaNav: Dynamic Feature and Layer Selection for Efficient Visual Navigation
di: Wang, Jiahui, et al.
Pubblicazione: (2025)
di: Wang, Jiahui, et al.
Pubblicazione: (2025)
Online-Adaptive Anomaly Detection for Defect Identification in Aircraft Assembly
di: Shete, Siddhant, et al.
Pubblicazione: (2024)
di: Shete, Siddhant, et al.
Pubblicazione: (2024)
RUKA: Rethinking the Design of Humanoid Hands with Learning
di: Zorin, Anya, et al.
Pubblicazione: (2025)
di: Zorin, Anya, et al.
Pubblicazione: (2025)
Hand-Object Interaction Pretraining from Videos
di: Singh, Himanshu Gaurav, et al.
Pubblicazione: (2024)
di: Singh, Himanshu Gaurav, et al.
Pubblicazione: (2024)
ArtReg: Visuo-Tactile based Pose Tracking and Manipulation of Unseen Articulated Objects
di: Murali, Prajval Kumar, et al.
Pubblicazione: (2025)
di: Murali, Prajval Kumar, et al.
Pubblicazione: (2025)
ViTaSCOPE: Visuo-tactile Implicit Representation for In-hand Pose and Extrinsic Contact Estimation
di: Lee, Jayjun, et al.
Pubblicazione: (2025)
di: Lee, Jayjun, et al.
Pubblicazione: (2025)
AnyTouch: Learning Unified Static-Dynamic Representation across Multiple Visuo-tactile Sensors
di: Feng, Ruoxuan, et al.
Pubblicazione: (2025)
di: Feng, Ruoxuan, et al.
Pubblicazione: (2025)
SmartPretrain: Model-Agnostic and Dataset-Agnostic Representation Learning for Motion Prediction
di: Zhou, Yang, et al.
Pubblicazione: (2024)
di: Zhou, Yang, et al.
Pubblicazione: (2024)
WAM-Diff: A Masked Diffusion VLA Framework with MoE and Online Reinforcement Learning for Autonomous Driving
di: Xu, Mingwang, et al.
Pubblicazione: (2025)
di: Xu, Mingwang, et al.
Pubblicazione: (2025)
D2E: Scaling Vision-Action Pretraining on Desktop Data for Transfer to Embodied AI
di: Choi, Suhwan, et al.
Pubblicazione: (2025)
di: Choi, Suhwan, et al.
Pubblicazione: (2025)
Understanding Particles From Video: Property Estimation of Granular Materials via Visuo-Haptic Learning
di: Zhang, Zeqing, et al.
Pubblicazione: (2024)
di: Zhang, Zeqing, et al.
Pubblicazione: (2024)
RDT2: Exploring the Scaling Limit of UMI Data Towards Zero-Shot Cross-Embodiment Generalization
di: Liu, Songming, et al.
Pubblicazione: (2026)
di: Liu, Songming, et al.
Pubblicazione: (2026)
Vidar: Embodied Video Diffusion Model for Generalist Manipulation
di: Feng, Yao, et al.
Pubblicazione: (2025)
di: Feng, Yao, et al.
Pubblicazione: (2025)
DynaVINS++: Robust Visual-Inertial State Estimator in Dynamic Environments by Adaptive Truncated Least Squares and Stable State Recovery
di: Song, Seungwon, et al.
Pubblicazione: (2024)
di: Song, Seungwon, et al.
Pubblicazione: (2024)
ChatVLA-2: Vision-Language-Action Model with Open-World Embodied Reasoning from Pretrained Knowledge
di: Zhou, Zhongyi, et al.
Pubblicazione: (2025)
di: Zhou, Zhongyi, et al.
Pubblicazione: (2025)
Visuo-Tactile Object Pose Estimation for a Multi-Finger Robot Hand with Low-Resolution In-Hand Tactile Sensing
di: Mack, Lukas, et al.
Pubblicazione: (2025)
di: Mack, Lukas, et al.
Pubblicazione: (2025)
Open-vocabulary Mobile Manipulation in Unseen Dynamic Environments with 3D Semantic Maps
di: Qiu, Dicong, et al.
Pubblicazione: (2024)
di: Qiu, Dicong, et al.
Pubblicazione: (2024)
Rapid Motor Adaptation for Robotic Manipulator Arms
di: Liang, Yichao, et al.
Pubblicazione: (2023)
di: Liang, Yichao, et al.
Pubblicazione: (2023)
MoRight: Motion Control Done Right
di: Liu, Shaowei, et al.
Pubblicazione: (2026)
di: Liu, Shaowei, et al.
Pubblicazione: (2026)
RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation
di: Liu, Songming, et al.
Pubblicazione: (2024)
di: Liu, Songming, et al.
Pubblicazione: (2024)
Disentangled Robot Learning via Separate Forward and Inverse Dynamics Pretraining
di: Zhang, Wenyao, et al.
Pubblicazione: (2026)
di: Zhang, Wenyao, et al.
Pubblicazione: (2026)
EgoZero: Robot Learning from Smart Glasses
di: Liu, Vincent, et al.
Pubblicazione: (2025)
di: Liu, Vincent, et al.
Pubblicazione: (2025)
OMEGA: Efficient Occlusion-Aware Navigation for Air-Ground Robot in Dynamic Environments via State Space Model
di: Wang, Junming, et al.
Pubblicazione: (2024)
di: Wang, Junming, et al.
Pubblicazione: (2024)
Vision-and-Language Navigation Generative Pretrained Transformer
di: Hanlin, Wen
Pubblicazione: (2024)
di: Hanlin, Wen
Pubblicazione: (2024)
AnyTouch 2: General Optical Tactile Representation Learning For Dynamic Tactile Perception
di: Feng, Ruoxuan, et al.
Pubblicazione: (2026)
di: Feng, Ruoxuan, et al.
Pubblicazione: (2026)
Documenti analoghi
-
P3-PO: Prescriptive Point Priors for Visuo-Spatial Generalization of Robot Policies
di: Levy, Mara, et al.
Pubblicazione: (2024) -
Touch begins where vision ends: Generalizable policies for contact-rich manipulation
di: Zhao, Zifan, et al.
Pubblicazione: (2025) -
Adapting Pretrained ViTs with Convolution Injector for Visuo-Motor Control
di: Hwang, Dongyoon, et al.
Pubblicazione: (2024) -
MoDem-V2: Visuo-Motor World Models for Real-World Robot Manipulation
di: Lancaster, Patrick, et al.
Pubblicazione: (2023) -
Point Policy: Unifying Observations and Actions with Key Points for Robot Manipulation
di: Haldar, Siddhant, et al.
Pubblicazione: (2025)