RoboGolf: Mastering Real-World Minigolf with a Reflective Multi-Modality Vision-Language Model
Fuente:
arXiv
Saved in:
| Main Authors: | Zhou, Hantao, Ji, Tianying, Sommerhalder, Lukas, Goerner, Michael, Hendrich, Norman, Zhang, Jianwei, Sun, Fuchun, Xu, Huazhe |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Smooth Computation without Input Delay: Robust Tube-Based Model Predictive Control for Robot Manipulator Planning
by: Luo, Yu, et al.
Published: (2024)
by: Luo, Yu, et al.
Published: (2024)
RoboTracer: Mastering Spatial Trace with Reasoning in Vision-Language Models for Robotics
by: Zhou, Enshen, et al.
Published: (2025)
by: Zhou, Enshen, et al.
Published: (2025)
Seizing Serendipity: Exploiting the Value of Past Success in Off-Policy Actor-Critic
by: Ji, Tianying, et al.
Published: (2023)
by: Ji, Tianying, et al.
Published: (2023)
Offline-Boosted Actor-Critic: Adaptively Blending Optimal Historical Behaviors in Deep Off-Policy RL
by: Luo, Yu, et al.
Published: (2024)
by: Luo, Yu, et al.
Published: (2024)
OMPO: A Unified Framework for RL under Policy and Dynamics Shifts
by: Luo, Yu, et al.
Published: (2024)
by: Luo, Yu, et al.
Published: (2024)
RoboDuet: Learning a Cooperative Policy for Whole-body Legged Loco-Manipulation
by: Pan, Guoping, et al.
Published: (2024)
by: Pan, Guoping, et al.
Published: (2024)
RecipeMasterLLM: Revisiting RoboEarth in the Era of Large Language Models
by: Bozcuoglu, Asil Kaan, et al.
Published: (2025)
by: Bozcuoglu, Asil Kaan, et al.
Published: (2025)
Beyond Action Residuals: Real-World Robot Policy Steering via Bottleneck Latent Reinforcement Learning
by: Yu, Dongjie, et al.
Published: (2026)
by: Yu, Dongjie, et al.
Published: (2026)
RoboStream: Weaving Spatio-Temporal Reasoning with Memory in Vision-Language Models for Robotics
by: Huang, Yuzhi, et al.
Published: (2026)
by: Huang, Yuzhi, et al.
Published: (2026)
The Cambridge RoboMaster: An Agile Multi-Robot Research Platform
by: Blumenkamp, Jan, et al.
Published: (2024)
by: Blumenkamp, Jan, et al.
Published: (2024)
RoboArena: Distributed Real-World Evaluation of Generalist Robot Policies
by: Atreya, Pranav, et al.
Published: (2025)
by: Atreya, Pranav, et al.
Published: (2025)
RoboReward: General-Purpose Vision-Language Reward Models for Robotics
by: Lee, Tony, et al.
Published: (2026)
by: Lee, Tony, et al.
Published: (2026)
RoboReflect: A Robotic Reflective Reasoning Framework for Grasping Ambiguous-Condition Objects
by: Luo, Zhen, et al.
Published: (2025)
by: Luo, Zhen, et al.
Published: (2025)
Robo-ABC: Affordance Generalization Beyond Categories via Semantic Correspondence for Robot Manipulation
by: Ju, Yuanchen, et al.
Published: (2024)
by: Ju, Yuanchen, et al.
Published: (2024)
RoboWheel: A Data Engine from Real-World Human Demonstrations for Cross-Embodiment Robotic Learning
by: Zhang, Yuhong, et al.
Published: (2025)
by: Zhang, Yuhong, et al.
Published: (2025)
RoboHorizon: An LLM-Assisted Multi-View World Model for Long-Horizon Robotic Manipulation
by: Chen, Zixuan, et al.
Published: (2025)
by: Chen, Zixuan, et al.
Published: (2025)
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors
by: Huang, Haifeng, et al.
Published: (2025)
by: Huang, Haifeng, et al.
Published: (2025)
Subequivariant Reinforcement Learning in 3D Multi-Entity Physical Environments
by: Chen, Runfa, et al.
Published: (2024)
by: Chen, Runfa, et al.
Published: (2024)
CrayonRobo: Object-Centric Prompt-Driven Vision-Language-Action Model for Robotic Manipulation
by: Li, Xiaoqi, et al.
Published: (2025)
by: Li, Xiaoqi, et al.
Published: (2025)
RoboNurse-VLA: Robotic Scrub Nurse System based on Vision-Language-Action Model
by: Li, Shunlei, et al.
Published: (2024)
by: Li, Shunlei, et al.
Published: (2024)
RoboDreamer: Learning Compositional World Models for Robot Imagination
by: Zhou, Siyuan, et al.
Published: (2024)
by: Zhou, Siyuan, et al.
Published: (2024)
RoboChallenge: Large-scale Real-robot Evaluation of Embodied Policies
by: Yakefu, Adina, et al.
Published: (2025)
by: Yakefu, Adina, et al.
Published: (2025)
RoboTron-Sim: Improving Real-World Driving via Simulated Hard-Case
by: Xiao, Baihui, et al.
Published: (2025)
by: Xiao, Baihui, et al.
Published: (2025)
DA-MMP: Learning Coordinated and Accurate Throwing with Dynamics-Aware Motion Manifold Primitives
by: Chu, Chi, et al.
Published: (2025)
by: Chu, Chi, et al.
Published: (2025)
DTactive: A Vision-Based Tactile Sensor with Active Surface
by: Xu, Jikai, et al.
Published: (2024)
by: Xu, Jikai, et al.
Published: (2024)
RoboFlow4D: A Lightweight Flow World Model Toward Real-Time Flow-Guided Robotic Manipulation
by: Lin, Sixu, et al.
Published: (2026)
by: Lin, Sixu, et al.
Published: (2026)
Learning Multi-Modal Whole-Body Control for Real-World Humanoid Robots
by: Dugar, Pranay, et al.
Published: (2024)
by: Dugar, Pranay, et al.
Published: (2024)
RoboSubtaskNet: Temporal Sub-task Segmentation for Human-to-Robot Skill Transfer in Real-World Environments
by: Sharma, Dharmendra, et al.
Published: (2026)
by: Sharma, Dharmendra, et al.
Published: (2026)
Multi-segment Soft Robot Control via Deep Koopman-based Model Predictive Control
by: Lv, Lei, et al.
Published: (2025)
by: Lv, Lei, et al.
Published: (2025)
RoboLLM: Robotic Vision Tasks Grounded on Multimodal Large Language Models
by: Long, Zijun, et al.
Published: (2023)
by: Long, Zijun, et al.
Published: (2023)
CollabVLA: Self-Reflective Vision-Language-Action Model Dreaming Together with Human
by: Sun, Nan, et al.
Published: (2025)
by: Sun, Nan, et al.
Published: (2025)
RL-100: Performant Robotic Manipulation with Real-World Reinforcement Learning
by: Lei, Kun, et al.
Published: (2025)
by: Lei, Kun, et al.
Published: (2025)
RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models
by: Kwok, Jacky, et al.
Published: (2025)
by: Kwok, Jacky, et al.
Published: (2025)
RoboScape: Physics-informed Embodied World Model
by: Shang, Yu, et al.
Published: (2025)
by: Shang, Yu, et al.
Published: (2025)
Failure-Aware RL: Reliable Offline-to-Online Reinforcement Learning with Self-Recovery for Real-World Manipulation
by: Li, Huanyu, et al.
Published: (2026)
by: Li, Huanyu, et al.
Published: (2026)
Advancing Audio-Visual Navigation Through Multi-Agent Collaboration in 3D Environments
by: Zhang, Hailong, et al.
Published: (2025)
by: Zhang, Hailong, et al.
Published: (2025)
RoboRefer: Towards Spatial Referring with Reasoning in Vision-Language Models for Robotics
by: Zhou, Enshen, et al.
Published: (2025)
by: Zhou, Enshen, et al.
Published: (2025)
RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
by: Yuan, Wentao, et al.
Published: (2024)
by: Yuan, Wentao, et al.
Published: (2024)
On Robustness of Vision-Language-Action Model against Multi-Modal Perturbations
by: Guo, Jianing, et al.
Published: (2025)
by: Guo, Jianing, et al.
Published: (2025)
From Production Logistics to Smart Manufacturing: The Vision for a New RoboCup Industrial League
by: Dissanayaka, Supun, et al.
Published: (2025)
by: Dissanayaka, Supun, et al.
Published: (2025)
Similar Items
-
Smooth Computation without Input Delay: Robust Tube-Based Model Predictive Control for Robot Manipulator Planning
by: Luo, Yu, et al.
Published: (2024) -
RoboTracer: Mastering Spatial Trace with Reasoning in Vision-Language Models for Robotics
by: Zhou, Enshen, et al.
Published: (2025) -
Seizing Serendipity: Exploiting the Value of Past Success in Off-Policy Actor-Critic
by: Ji, Tianying, et al.
Published: (2023) -
Offline-Boosted Actor-Critic: Adaptively Blending Optimal Historical Behaviors in Deep Off-Policy RL
by: Luo, Yu, et al.
Published: (2024) -
OMPO: A Unified Framework for RL under Policy and Dynamics Shifts
by: Luo, Yu, et al.
Published: (2024)