RoboArena: Distributed Real-World Evaluation of Generalist Robot Policies
Fuente:
arXiv
Saved in:
| Main Authors: | Atreya, Pranav, Pertsch, Karl, Lee, Tony, Kim, Moo Jin, Jain, Arhan, Kuramshin, Artur, Eppner, Clemens, Neary, Cyrus, Hu, Edward, Ramos, Fabio, Tremblay, Jonathan, Arora, Kanav, Ellis, Kirsty, Macesanu, Luca, Villasevil, Marcel Torne, Leonard, Matthew, Cho, Meedeum, Aslan, Ozgur, Dass, Shivin, Wang, Jie, Reger, William, Yuan, Xingfang, Yang, Xuning, Gupta, Abhishek, Jayaraman, Dinesh, Berseth, Glen, Daniilidis, Kostas, Martin-Martin, Roberto, Lee, Youngwoon, Liang, Percy, Finn, Chelsea, Levine, Sergey |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Improving Pre-Trained Vision-Language-Action Policies with Model-Based Search
by: Neary, Cyrus, et al.
Published: (2025)
by: Neary, Cyrus, et al.
Published: (2025)
COLLAGE: Adaptive Fusion-based Retrieval for Augmented Policy Learning
by: Kumar, Sateesh, et al.
Published: (2025)
by: Kumar, Sateesh, et al.
Published: (2025)
Model-Based Runtime Monitoring with Interactive Imitation Learning
by: Liu, Huihan, et al.
Published: (2023)
by: Liu, Huihan, et al.
Published: (2023)
UniSkill: Imitating Human Videos via Cross-Embodiment Skill Representations
by: Kim, Hanjung, et al.
Published: (2025)
by: Kim, Hanjung, et al.
Published: (2025)
Learning to Look: Seeking Information for Decision Making via Policy Factorization
by: Dass, Shivin, et al.
Published: (2024)
by: Dass, Shivin, et al.
Published: (2024)
DataMIL: Selecting Data for Robot Imitation Learning with Datamodels
by: Dass, Shivin, et al.
Published: (2025)
by: Dass, Shivin, et al.
Published: (2025)
Mash, Spread, Slice! Learning to Manipulate Object States via Visual Spatial Progress
by: Mandikal, Priyanka, et al.
Published: (2025)
by: Mandikal, Priyanka, et al.
Published: (2025)
Model-based Offline Reinforcement Learning with Lower Expectile Q-Learning
by: Park, Kwanyoung, et al.
Published: (2024)
by: Park, Kwanyoung, et al.
Published: (2024)
DreamSmooth: Improving Model-based Reinforcement Learning via Reward Smoothing
by: Lee, Vint, et al.
Published: (2023)
by: Lee, Vint, et al.
Published: (2023)
RoboReward: General-Purpose Vision-Language Reward Models for Robotics
by: Lee, Tony, et al.
Published: (2026)
by: Lee, Tony, et al.
Published: (2026)
PolaRiS: Scalable Real-to-Sim Evaluations for Generalist Robot Policies
by: Jain, Arhan, et al.
Published: (2025)
by: Jain, Arhan, et al.
Published: (2025)
Robot Policy Evaluation for Sim-to-Real Transfer: A Benchmarking Perspective
by: Yang, Xuning, et al.
Published: (2025)
by: Yang, Xuning, et al.
Published: (2025)
Chunk-Guided Q-Learning
by: Song, Gwanwoo, et al.
Published: (2026)
by: Song, Gwanwoo, et al.
Published: (2026)
TLDR: Unsupervised Goal-Conditioned RL via Temporal Distance-Aware Representations
by: Bae, Junik, et al.
Published: (2024)
by: Bae, Junik, et al.
Published: (2024)
Physical oceanography during Le Suroît cruise Topogulf 2
by: Arhan, Michel
Published: (2010)
by: Arhan, Michel
Published: (2010)
Bathymetry along track line of L ATALANTE cruise 35A3CITHER3_1
by: Arhan, Michel
Published: (2006)
by: Arhan, Michel
Published: (2006)
Bathymetry along track line of L ATALANTE cruise 35A3CITHER3_2
by: Arhan, Michel
Published: (2006)
by: Arhan, Michel
Published: (2006)
Physical oceanography during Le Suroît cruise Topogulf 3
by: Arhan, Michel
Published: (2010)
by: Arhan, Michel
Published: (2010)
ARM-FM: Automated Reward Machines via Foundation Models for Compositional Reinforcement Learning
by: Castanyer, Roger Creus, et al.
Published: (2025)
by: Castanyer, Roger Creus, et al.
Published: (2025)
TeleMoMa: A Modular and Versatile Teleoperation System for Mobile Manipulation
by: Dass, Shivin, et al.
Published: (2024)
by: Dass, Shivin, et al.
Published: (2024)
AutoEval: Autonomous Evaluation of Generalist Robot Manipulation Policies in the Real World
by: Zhou, Zhiyuan, et al.
Published: (2025)
by: Zhou, Zhiyuan, et al.
Published: (2025)
Latent Policy Steering through One-Step Flow Policies
by: Im, Hokyun, et al.
Published: (2026)
by: Im, Hokyun, et al.
Published: (2026)
Scalable Offline Model-Based RL with Action Chunks
by: Park, Kwanyoung, et al.
Published: (2025)
by: Park, Kwanyoung, et al.
Published: (2025)
CAVER: Curious Audiovisual Exploring Robot
by: Macesanu, Luca, et al.
Published: (2025)
by: Macesanu, Luca, et al.
Published: (2025)
Counseling trainees’ academic burnout, meaningful work, and career choice satisfaction: A resilience framework
by: Byeolbee Um, et al.
Published: (2024)
by: Byeolbee Um, et al.
Published: (2024)
Is Exploration or Optimization the Problem for Deep Reinforcement Learning?
by: Berseth, Glen
Published: (2025)
by: Berseth, Glen
Published: (2025)
Tokenization of Real Estate Assets Using Blockchain
by: Joshi, Shashank, et al.
Published: (2024)
by: Joshi, Shashank, et al.
Published: (2024)
Interpretable and Perceptually-Aligned Music Similarity with Pretrained Embeddings
by: Vohra, Arhan, et al.
Published: (2026)
by: Vohra, Arhan, et al.
Published: (2026)
Robot Learning with Super-Linear Scaling
by: Torne, Marcel, et al.
Published: (2024)
by: Torne, Marcel, et al.
Published: (2024)
Mixed-Initiative Dialog for Human-Robot Collaborative Manipulation
by: Yu, Albert, et al.
Published: (2025)
by: Yu, Albert, et al.
Published: (2025)
RoboLab: A High-Fidelity Simulation Benchmark for Analysis of Task Generalist Policies
by: Yang, Xuning, et al.
Published: (2026)
by: Yang, Xuning, et al.
Published: (2026)
TwinVLA: Data-Efficient Bimanual Manipulation with Twin Single-Arm Vision-Language-Action Models
by: Im, Hokyun, et al.
Published: (2025)
by: Im, Hokyun, et al.
Published: (2025)
Kognitive Robotik wird zum Wendepunkt
by: David Reger
Published: (2025)
by: David Reger
Published: (2025)
Solvability of a Hadamard Fractional Boundary Value Problem at Resonance on Infinite Domain
by: Xingfang Feng, et al.
Published: (2024)
by: Xingfang Feng, et al.
Published: (2024)
Updated Theoretical Resource Assessment for CONUS Riverine Hydrokinetic Energy
by: Kim, Jeongin, et al.
Published: (2024)
by: Kim, Jeongin, et al.
Published: (2024)
HumanoidBench: Simulated Humanoid Benchmark for Whole-Body Locomotion and Manipulation
by: Sferrazza, Carmelo, et al.
Published: (2024)
by: Sferrazza, Carmelo, et al.
Published: (2024)
Crafting In-context Examples according to LMs' Parametric Knowledge
by: Lee, Yoonsang, et al.
Published: (2023)
by: Lee, Yoonsang, et al.
Published: (2023)
Exploring Radio Pulsars with New Technologies
by: Torne, Pablo
Published: (2017)
by: Torne, Pablo
Published: (2017)
Improved Bounds on Diffsequences with Gaps in Powers of 2
by: Talwar, Kanav, et al.
Published: (2025)
by: Talwar, Kanav, et al.
Published: (2025)
A novel tool for risk assessment, screening, diagnosis, assessment, and therapy in postpartum depression
by: Prachi Sharma, et al.
Published: (2024)
by: Prachi Sharma, et al.
Published: (2024)
Similar Items
-
Improving Pre-Trained Vision-Language-Action Policies with Model-Based Search
by: Neary, Cyrus, et al.
Published: (2025) -
COLLAGE: Adaptive Fusion-based Retrieval for Augmented Policy Learning
by: Kumar, Sateesh, et al.
Published: (2025) -
Model-Based Runtime Monitoring with Interactive Imitation Learning
by: Liu, Huihan, et al.
Published: (2023) -
UniSkill: Imitating Human Videos via Cross-Embodiment Skill Representations
by: Kim, Hanjung, et al.
Published: (2025) -
Learning to Look: Seeking Information for Decision Making via Policy Factorization
by: Dass, Shivin, et al.
Published: (2024)