THE COLOSSEUM: A Benchmark for Evaluating Generalization for Robotic Manipulation
Fuente:
arXiv
Saved in:
| Main Authors: | Pumacay, Wilbert, Singh, Ishika, Duan, Jiafei, Krishna, Ranjay, Thomason, Jesse, Fox, Dieter |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
by: Yuan, Wentao, et al.
Published: (2024)
by: Yuan, Wentao, et al.
Published: (2024)
PEEK: Guiding and Minimal Image Representations for Zero-Shot Generalization of Robot Manipulation Policies
by: Zhang, Jesse, et al.
Published: (2025)
by: Zhang, Jesse, et al.
Published: (2025)
SAM2Act: Integrating Visual Foundation Model with A Memory Architecture for Robotic Manipulation
by: Fang, Haoquan, et al.
Published: (2025)
by: Fang, Haoquan, et al.
Published: (2025)
TOPReward: Token Probabilities as Hidden Zero-Shot Rewards for Robotics
by: Chen, Shirui, et al.
Published: (2026)
by: Chen, Shirui, et al.
Published: (2026)
Manipulate-Anything: Automating Real-World Robots using Vision-Language Models
by: Duan, Jiafei, et al.
Published: (2024)
by: Duan, Jiafei, et al.
Published: (2024)
RoboEval: Where Robotic Manipulation Meets Structured and Scalable Evaluation
by: Wang, Yi Ru, et al.
Published: (2025)
by: Wang, Yi Ru, et al.
Published: (2025)
AHA: A Vision-Language-Model for Detecting and Reasoning Over Failures in Robotic Manipulation
by: Duan, Jiafei, et al.
Published: (2024)
by: Duan, Jiafei, et al.
Published: (2024)
Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection
by: Anwar, Abrar, et al.
Published: (2025)
by: Anwar, Abrar, et al.
Published: (2025)
M3PT: A Transformer for Multimodal, Multi-Party Social Signal Prediction with Person-aware Blockwise Attention
by: Tang, Yiming, et al.
Published: (2025)
by: Tang, Yiming, et al.
Published: (2025)
PerAct2: Benchmarking and Learning for Robotic Bimanual Manipulation Tasks
by: Grotz, Markus, et al.
Published: (2024)
by: Grotz, Markus, et al.
Published: (2024)
Language Models can Infer Action Semantics for Symbolic Planners from Environment Feedback
by: Zhu, Wang, et al.
Published: (2024)
by: Zhu, Wang, et al.
Published: (2024)
TwoStep: Multi-agent Task Planning using Classical Planners and Large Language Models
by: Bai, David, et al.
Published: (2024)
by: Bai, David, et al.
Published: (2024)
DexMachina: Functional Retargeting for Bimanual Dexterous Manipulation
by: Mandi, Zhao, et al.
Published: (2025)
by: Mandi, Zhao, et al.
Published: (2025)
A Taxonomy for Evaluating Generalist Robot Manipulation Policies
by: Gao, Jensen, et al.
Published: (2025)
by: Gao, Jensen, et al.
Published: (2025)
Zero-Shot Visual Generalization in Robot Manipulation
by: Batra, Sumeet, et al.
Published: (2025)
by: Batra, Sumeet, et al.
Published: (2025)
λ: A Benchmark for Data-Efficiency in Long-Horizon Indoor Mobile Manipulation Robotics
by: Jaafar, Ahmed, et al.
Published: (2024)
by: Jaafar, Ahmed, et al.
Published: (2024)
Robometer: Scaling General-Purpose Robotic Reward Models via Trajectory Comparisons
by: Liang, Anthony, et al.
Published: (2026)
by: Liang, Anthony, et al.
Published: (2026)
Neural Robot Dynamics
by: Xu, Jie, et al.
Published: (2025)
by: Xu, Jie, et al.
Published: (2025)
Contrast Sets for Evaluating Language-Guided Robot Policies
by: Anwar, Abrar, et al.
Published: (2024)
by: Anwar, Abrar, et al.
Published: (2024)
Benchmarking Reinforcement Learning Methods for Dexterous Robotic Manipulation with a Three-Fingered Gripper
by: Cutler, Elizabeth, et al.
Published: (2024)
by: Cutler, Elizabeth, et al.
Published: (2024)
Efficient Data Collection for Robotic Manipulation via Compositional Generalization
by: Gao, Jensen, et al.
Published: (2024)
by: Gao, Jensen, et al.
Published: (2024)
Learning Neuro-symbolic Programs for Language Guided Robot Manipulation
by: Kalithasan, Namasivayam, et al.
Published: (2022)
by: Kalithasan, Namasivayam, et al.
Published: (2022)
MolmoSpaces: A Large-Scale Open Ecosystem for Robot Navigation and Manipulation
by: Kim, Yejin, et al.
Published: (2026)
by: Kim, Yejin, et al.
Published: (2026)
Unsupervised Skill Discovery for Robotic Manipulation through Automatic Task Generation
by: Jansonnie, Paul, et al.
Published: (2024)
by: Jansonnie, Paul, et al.
Published: (2024)
Embodied-R1: Reinforced Embodied Reasoning for General Robotic Manipulation
by: Yuan, Yifu, et al.
Published: (2025)
by: Yuan, Yifu, et al.
Published: (2025)
MSG: Multi-Stream Generative Policies for Sample-Efficient Robotic Manipulation
by: von Hartz, Jan Ole, et al.
Published: (2025)
by: von Hartz, Jan Ole, et al.
Published: (2025)
Dexterity from Smart Lenses: Multi-Fingered Robot Manipulation with In-the-Wild Human Demonstrations
by: Guzey, Irmak, et al.
Published: (2025)
by: Guzey, Irmak, et al.
Published: (2025)
Geometric Red-Teaming for Robotic Manipulation
by: Goel, Divyam, et al.
Published: (2025)
by: Goel, Divyam, et al.
Published: (2025)
Autoregressive Action Sequence Learning for Robotic Manipulation
by: Zhang, Xinyu, et al.
Published: (2024)
by: Zhang, Xinyu, et al.
Published: (2024)
Adaptive Diffusion Policy Optimization for Robotic Manipulation
by: Jiang, Huiyun, et al.
Published: (2025)
by: Jiang, Huiyun, et al.
Published: (2025)
Humanoid Manipulation Interface: Humanoid Whole-Body Manipulation from Robot-Free Demonstrations
by: Nai, Ruiqian, et al.
Published: (2026)
by: Nai, Ruiqian, et al.
Published: (2026)
Is Diversity All You Need for Scalable Robotic Manipulation?
by: Shi, Modi, et al.
Published: (2025)
by: Shi, Modi, et al.
Published: (2025)
Learning to Transfer Human Hand Skills for Robot Manipulations
by: Park, Sungjae, et al.
Published: (2025)
by: Park, Sungjae, et al.
Published: (2025)
Robotic Manipulation Datasets for Offline Compositional Reinforcement Learning
by: Hussing, Marcel, et al.
Published: (2023)
by: Hussing, Marcel, et al.
Published: (2023)
Memory, Benchmark & Robots: A Benchmark for Solving Complex Tasks with Reinforcement Learning
by: Cherepanov, Egor, et al.
Published: (2025)
by: Cherepanov, Egor, et al.
Published: (2025)
One-Shot Imitation Learning with Invariance Matching for Robotic Manipulation
by: Zhang, Xinyu, et al.
Published: (2024)
by: Zhang, Xinyu, et al.
Published: (2024)
Collective Intelligence for 2D Push Manipulations with Mobile Robots
by: Kuroki, So, et al.
Published: (2022)
by: Kuroki, So, et al.
Published: (2022)
From Seeing to Doing: Bridging Reasoning and Decision for Robotic Manipulation
by: Yuan, Yifu, et al.
Published: (2025)
by: Yuan, Yifu, et al.
Published: (2025)
SPIRE: Synergistic Planning, Imitation, and Reinforcement Learning for Long-Horizon Manipulation
by: Zhou, Zihan, et al.
Published: (2024)
by: Zhou, Zihan, et al.
Published: (2024)
BLAZER: Bootstrapping LLM-based Manipulation Agents with Zero-Shot Data Generation
by: Das, Rocktim Jyoti, et al.
Published: (2025)
by: Das, Rocktim Jyoti, et al.
Published: (2025)
Similar Items
-
RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
by: Yuan, Wentao, et al.
Published: (2024) -
PEEK: Guiding and Minimal Image Representations for Zero-Shot Generalization of Robot Manipulation Policies
by: Zhang, Jesse, et al.
Published: (2025) -
SAM2Act: Integrating Visual Foundation Model with A Memory Architecture for Robotic Manipulation
by: Fang, Haoquan, et al.
Published: (2025) -
TOPReward: Token Probabilities as Hidden Zero-Shot Rewards for Robotics
by: Chen, Shirui, et al.
Published: (2026) -
Manipulate-Anything: Automating Real-World Robots using Vision-Language Models
by: Duan, Jiafei, et al.
Published: (2024)