Automatic Environment Shaping is the Next Frontier in RL
Fuente:
arXiv
Saved in:
| Main Authors: | Park, Younghyo, Margolis, Gabriel B., Agrawal, Pulkit |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Bridging the Sim-to-Real Gap for Athletic Loco-Manipulation
by: Fey, Nolan, et al.
Published: (2025)
by: Fey, Nolan, et al.
Published: (2025)
SoftMimic: Learning Compliant Whole-body Control from Examples
by: Margolis, Gabriel B., et al.
Published: (2025)
by: Margolis, Gabriel B., et al.
Published: (2025)
Learning Force Control for Legged Manipulation
by: Portela, Tifanny, et al.
Published: (2024)
by: Portela, Tifanny, et al.
Published: (2024)
Tune to Learn: How Controller Gains Shape Robot Policy Learning
by: Bronars, Antonia, et al.
Published: (2026)
by: Bronars, Antonia, et al.
Published: (2026)
ORSO: Accelerating Reward Design via Online Reward Selection and Policy Optimization
by: Zhang, Chen Bo Calvin, et al.
Published: (2024)
by: Zhang, Chen Bo Calvin, et al.
Published: (2024)
ROER: Regularized Optimal Experience Replay
by: Li, Changling, et al.
Published: (2024)
by: Li, Changling, et al.
Published: (2024)
Aligning Robot and Human Representations
by: Bobu, Andreea, et al.
Published: (2023)
by: Bobu, Andreea, et al.
Published: (2023)
Vegetable Peeling: A Case Study in Constrained Dexterous Manipulation
by: Chen, Tao, et al.
Published: (2024)
by: Chen, Tao, et al.
Published: (2024)
Reconciling Reality through Simulation: A Real-to-Sim-to-Real Approach for Robust Manipulation
by: Torne, Marcel, et al.
Published: (2024)
by: Torne, Marcel, et al.
Published: (2024)
Few-Shot Task Learning through Inverse Generative Modeling
by: Netanyahu, Aviv, et al.
Published: (2024)
by: Netanyahu, Aviv, et al.
Published: (2024)
Robot Learning with Super-Linear Scaling
by: Torne, Marcel, et al.
Published: (2024)
by: Torne, Marcel, et al.
Published: (2024)
Confounding Robust Continuous Control via Automatic Reward Shaping
by: Juliani, Mateo, et al.
Published: (2026)
by: Juliani, Mateo, et al.
Published: (2026)
METRA: Scalable Unsupervised RL with Metric-Aware Abstraction
by: Park, Seohong, et al.
Published: (2023)
by: Park, Seohong, et al.
Published: (2023)
HIQL: Offline Goal-Conditioned RL with Latent States as Actions
by: Park, Seohong, et al.
Published: (2023)
by: Park, Seohong, et al.
Published: (2023)
Embodied Red Teaming for Auditing Robotic Foundation Models
by: Karnik, Sathwik, et al.
Published: (2024)
by: Karnik, Sathwik, et al.
Published: (2024)
From Imitation to Refinement -- Residual RL for Precise Assembly
by: Ankile, Lars, et al.
Published: (2024)
by: Ankile, Lars, et al.
Published: (2024)
SuReNav: Superpixel Graph-based Constraint Relaxation for Navigation in Over-constrained Environments
by: Koh, Keonyoung, et al.
Published: (2026)
by: Koh, Keonyoung, et al.
Published: (2026)
Language-Conditioned Offline RL for Multi-Robot Navigation
by: Morad, Steven, et al.
Published: (2024)
by: Morad, Steven, et al.
Published: (2024)
Sample-efficient and Scalable Exploration in Continuous-Time RL
by: Iten, Klemens, et al.
Published: (2025)
by: Iten, Klemens, et al.
Published: (2025)
Investigating Memory in Model-Free RL with POPGym Arcade
by: Wang, Zekang, et al.
Published: (2025)
by: Wang, Zekang, et al.
Published: (2025)
DexHub and DART: Towards Internet Scale Robot Data Collection
by: Park, Younghyo, et al.
Published: (2024)
by: Park, Younghyo, et al.
Published: (2024)
GRAM: Generalization in Deep RL with a Robust Adaptation Module
by: Queeney, James, et al.
Published: (2024)
by: Queeney, James, et al.
Published: (2024)
First Order Model-Based RL through Decoupled Backpropagation
by: Amigo, Joseph, et al.
Published: (2025)
by: Amigo, Joseph, et al.
Published: (2025)
CaRL: Learning Scalable Planning Policies with Simple Rewards
by: Jaeger, Bernhard, et al.
Published: (2025)
by: Jaeger, Bernhard, et al.
Published: (2025)
Dual RL: Unification and New Methods for Reinforcement and Imitation Learning
by: Sikchi, Harshit, et al.
Published: (2023)
by: Sikchi, Harshit, et al.
Published: (2023)
Generalised Linear Models in Deep Bayesian RL with Learnable Basis Functions
by: You, Jingyang, et al.
Published: (2025)
by: You, Jingyang, et al.
Published: (2025)
Enhancing RL Generalizability in Robotics through SHAP Analysis of Algorithms and Hyperparameters
by: Kong, Lingxiao, et al.
Published: (2026)
by: Kong, Lingxiao, et al.
Published: (2026)
Posterior Behavioral Cloning: Pretraining BC Policies for Efficient RL Finetuning
by: Wagenmaker, Andrew, et al.
Published: (2025)
by: Wagenmaker, Andrew, et al.
Published: (2025)
RL-100: Performant Robotic Manipulation with Real-World Reinforcement Learning
by: Lei, Kun, et al.
Published: (2025)
by: Lei, Kun, et al.
Published: (2025)
DEAS: DEtached value learning with Action Sequence for Scalable Offline RL
by: Kim, Changyeon, et al.
Published: (2025)
by: Kim, Changyeon, et al.
Published: (2025)
ARC-RL: A Reinforcement Learning Playground Inspired by ARC Raiders
by: Romeo, Carlo, et al.
Published: (2026)
by: Romeo, Carlo, et al.
Published: (2026)
Performance Comparison of Deep RL Algorithms for Mixed Traffic Cooperative Lane-Changing
by: Yao, Xue, et al.
Published: (2024)
by: Yao, Xue, et al.
Published: (2024)
RACER: Epistemic Risk-Sensitive RL Enables Fast Driving with Fewer Crashes
by: Stachowicz, Kyle, et al.
Published: (2024)
by: Stachowicz, Kyle, et al.
Published: (2024)
CtRL-Sim: Reactive and Controllable Driving Agents with Offline Reinforcement Learning
by: Rowe, Luke, et al.
Published: (2024)
by: Rowe, Luke, et al.
Published: (2024)
A Review of Online Diffusion Policy RL Algorithms for Scalable Robotic Control
by: Choi, Wonhyeok, et al.
Published: (2026)
by: Choi, Wonhyeok, et al.
Published: (2026)
Efficient Online RL Fine Tuning with Offline Pre-trained Policy Only
by: Xiao, Wei, et al.
Published: (2025)
by: Xiao, Wei, et al.
Published: (2025)
ELMUR: External Layer Memory with Update/Rewrite for Long-Horizon RL Problems
by: Cherepanov, Egor, et al.
Published: (2025)
by: Cherepanov, Egor, et al.
Published: (2025)
ReFORM: Reflected Flows for On-support Offline RL via Noise Manipulation
by: Zhang, Songyuan, et al.
Published: (2026)
by: Zhang, Songyuan, et al.
Published: (2026)
Learning Human-Like RL Agents Through Trajectory Optimization With Action Quantization
by: Guo, Jian-Ting, et al.
Published: (2025)
by: Guo, Jian-Ting, et al.
Published: (2025)
TWISTED-RL: Hierarchical Skilled Agents for Knot-Tying without Human Demonstrations
by: Freund, Guy, et al.
Published: (2026)
by: Freund, Guy, et al.
Published: (2026)
Similar Items
-
Bridging the Sim-to-Real Gap for Athletic Loco-Manipulation
by: Fey, Nolan, et al.
Published: (2025) -
SoftMimic: Learning Compliant Whole-body Control from Examples
by: Margolis, Gabriel B., et al.
Published: (2025) -
Learning Force Control for Legged Manipulation
by: Portela, Tifanny, et al.
Published: (2024) -
Tune to Learn: How Controller Gains Shape Robot Policy Learning
by: Bronars, Antonia, et al.
Published: (2026) -
ORSO: Accelerating Reward Design via Online Reward Selection and Policy Optimization
by: Zhang, Chen Bo Calvin, et al.
Published: (2024)