Robot Air Hockey: A Manipulation Testbed for Robot Learning with Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Chuck, Caleb, Qi, Carl, Munje, Michael J., Li, Shuozhe, Rudolph, Max, Shi, Chang, Agarwal, Siddhant, Sikchi, Harshit, Peri, Abhinav, Dayal, Sarthak, Kuo, Evan, Mehta, Kavan, Wang, Anthony, Stone, Peter, Zhang, Amy, Niekum, Scott |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Dual Approach to Imitation Learning from Observations with Offline Datasets
by: Sikchi, Harshit, et al.
Published: (2024)
by: Sikchi, Harshit, et al.
Published: (2024)
RLZero: Direct Policy Inference from Language Without In-Domain Supervision
by: Sikchi, Harshit, et al.
Published: (2024)
by: Sikchi, Harshit, et al.
Published: (2024)
An Optimal Discriminator Weighted Imitation Perspective for Reinforcement Learning
by: Xu, Haoran, et al.
Published: (2025)
by: Xu, Haoran, et al.
Published: (2025)
Proto Successor Measure: Representing the Behavior Space of an RL Agent
by: Agarwal, Siddhant, et al.
Published: (2024)
by: Agarwal, Siddhant, et al.
Published: (2024)
Regularized Latent Dynamics Prediction is a Strong Baseline For Behavioral Foundation Models
by: Jajoo, Pranaya, et al.
Published: (2026)
by: Jajoo, Pranaya, et al.
Published: (2026)
Dual RL: Unification and New Methods for Reinforcement and Imitation Learning
by: Sikchi, Harshit, et al.
Published: (2023)
by: Sikchi, Harshit, et al.
Published: (2023)
Exploiting Local Dynamics Regularity for Reusable Skills in Offline Hierarchical RL
by: Dayal, Sarthak, et al.
Published: (2026)
by: Dayal, Sarthak, et al.
Published: (2026)
Null Counterfactual Factor Interactions for Goal-Conditioned Reinforcement Learning
by: Chuck, Caleb, et al.
Published: (2025)
by: Chuck, Caleb, et al.
Published: (2025)
Learning Action-based Representations Using Invariance
by: Rudolph, Max, et al.
Published: (2024)
by: Rudolph, Max, et al.
Published: (2024)
SMORE: Score Models for Offline Goal-Conditioned Reinforcement Learning
by: Sikchi, Harshit, et al.
Published: (2023)
by: Sikchi, Harshit, et al.
Published: (2023)
Contrastive Preference Learning: Learning from Human Feedback without RL
by: Hejna, Joey, et al.
Published: (2023)
by: Hejna, Joey, et al.
Published: (2023)
A Retrospective on the Robot Air Hockey Challenge: Benchmarking Robust, Reliable, and Safe Learning Techniques for Real-world Robotics
by: Liu, Puze, et al.
Published: (2024)
by: Liu, Puze, et al.
Published: (2024)
Distilling Contact Planning for Fast Trajectory Optimization in Robot Air Hockey
by: Jankowski, Julius, et al.
Published: (2024)
by: Jankowski, Julius, et al.
Published: (2024)
Granger Causal Interaction Skill Chains
by: Chuck, Caleb, et al.
Published: (2023)
by: Chuck, Caleb, et al.
Published: (2023)
Reinforcement Learning Within the Classical Robotics Stack: A Case Study in Robot Soccer
by: Labiosa, Adam, et al.
Published: (2024)
by: Labiosa, Adam, et al.
Published: (2024)
De Selma a Montgomery
by: Stone, Chuck
Published: (1965)
by: Stone, Chuck
Published: (1965)
Learning to Play Air Hockey with Model-Based Deep Reinforcement Learning
by: Orsula, Andrej
Published: (2024)
by: Orsula, Andrej
Published: (2024)
Scaling Laws for Reward Model Overoptimization in Direct Alignment Algorithms
by: Rafailov, Rafael, et al.
Published: (2024)
by: Rafailov, Rafael, et al.
Published: (2024)
Automated Discovery of Functional Actual Causes in Complex Environments
by: Chuck, Caleb, et al.
Published: (2024)
by: Chuck, Caleb, et al.
Published: (2024)
SkiLD: Unsupervised Skill Discovery Guided by Factor Interactions
by: Wang, Zizhao, et al.
Published: (2024)
by: Wang, Zizhao, et al.
Published: (2024)
Waypoint-Based Reinforcement Learning for Robot Manipulation Tasks
by: Mehta, Shaunak A., et al.
Published: (2024)
by: Mehta, Shaunak A., et al.
Published: (2024)
Point Policy: Unifying Observations and Actions with Key Points for Robot Manipulation
by: Haldar, Siddhant, et al.
Published: (2025)
by: Haldar, Siddhant, et al.
Published: (2025)
Fast Adaptation with Behavioral Foundation Models
by: Sikchi, Harshit, et al.
Published: (2025)
by: Sikchi, Harshit, et al.
Published: (2025)
CREStE: Scalable Mapless Navigation with Internet Scale Priors and Counterfactual Guidance
by: Zhang, Arthur, et al.
Published: (2025)
by: Zhang, Arthur, et al.
Published: (2025)
CLASS: Contrastive Learning via Action Sequence Supervision for Robot Manipulation
by: Lee, Sung-Wook, et al.
Published: (2025)
by: Lee, Sung-Wook, et al.
Published: (2025)
P3-PO: Prescriptive Point Priors for Visuo-Spatial Generalization of Robot Policies
by: Levy, Mara, et al.
Published: (2024)
by: Levy, Mara, et al.
Published: (2024)
Learning Robotic Manipulation Policies from Point Clouds with Conditional Flow Matching
by: Chisari, Eugenio, et al.
Published: (2024)
by: Chisari, Eugenio, et al.
Published: (2024)
Non-conflicting Energy Minimization in Reinforcement Learning based Robot Control
by: Peri, Skand, et al.
Published: (2025)
by: Peri, Skand, et al.
Published: (2025)
SocialNav-SUB: Benchmarking VLMs for Scene Understanding in Social Robot Navigation
by: Munje, Michael J., et al.
Published: (2025)
by: Munje, Michael J., et al.
Published: (2025)
From Robotics to Sepsis Treatment: Offline RL via Geometric Pessimism
by: Wanjari, Sarthak
Published: (2026)
by: Wanjari, Sarthak
Published: (2026)
OPEN TEACH: A Versatile Teleoperation System for Robotic Manipulation
by: Iyer, Aadhithya, et al.
Published: (2024)
by: Iyer, Aadhithya, et al.
Published: (2024)
1 Modular Parallel Manipulator for Long-Term Soft Robotic Data Collection
by: Chin, Kiyn, et al.
Published: (2024)
by: Chin, Kiyn, et al.
Published: (2024)
Constraining Gaussian Process Implicit Surfaces for Robot Manipulation via Dataset Refinement
by: Kumar, Abhinav, et al.
Published: (2024)
by: Kumar, Abhinav, et al.
Published: (2024)
CogniSQL-R1-Zero: Lightweight Reinforced Reasoning for Efficient SQL Generation
by: Gajjar, Kushal, et al.
Published: (2025)
by: Gajjar, Kushal, et al.
Published: (2025)
Utilizing Inpainting for Keypoint Detection for Vision-Based Control of Robotic Manipulators
by: Chatterjee, Sreejani, et al.
Published: (2026)
by: Chatterjee, Sreejani, et al.
Published: (2026)
Image-Based Roadmaps for Vision-Only Planning and Control of Robotic Manipulators
by: Chatterjee, Sreejani, et al.
Published: (2025)
by: Chatterjee, Sreejani, et al.
Published: (2025)
Bayesian Optimization for Sample-Efficient Policy Improvement in Robotic Manipulation
by: Röfer, Adrian, et al.
Published: (2024)
by: Röfer, Adrian, et al.
Published: (2024)
Autoregressive Action Sequence Learning for Robotic Manipulation
by: Zhang, Xinyu, et al.
Published: (2024)
by: Zhang, Xinyu, et al.
Published: (2024)
Learning Efficient Robotic Garment Manipulation with Standardization
by: Zhou, Changshi, et al.
Published: (2025)
by: Zhou, Changshi, et al.
Published: (2025)
Lifelong Language-Conditioned Robotic Manipulation Learning
by: Wang, Xudong, et al.
Published: (2026)
by: Wang, Xudong, et al.
Published: (2026)
Similar Items
-
A Dual Approach to Imitation Learning from Observations with Offline Datasets
by: Sikchi, Harshit, et al.
Published: (2024) -
RLZero: Direct Policy Inference from Language Without In-Domain Supervision
by: Sikchi, Harshit, et al.
Published: (2024) -
An Optimal Discriminator Weighted Imitation Perspective for Reinforcement Learning
by: Xu, Haoran, et al.
Published: (2025) -
Proto Successor Measure: Representing the Behavior Space of an RL Agent
by: Agarwal, Siddhant, et al.
Published: (2024) -
Regularized Latent Dynamics Prediction is a Strong Baseline For Behavioral Foundation Models
by: Jajoo, Pranaya, et al.
Published: (2026)