First Order Model-Based RL through Decoupled Backpropagation
Fuente:
arXiv
Saved in:
| Main Authors: | Amigo, Joseph, Khorrambakht, Rooholla, Chane-Sane, Elliot, Mansard, Nicolas, Righetti, Ludovic |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Coupled Local and Global World Models for Efficient First Order RL
by: Amigo, Joseph, et al.
Published: (2026)
by: Amigo, Joseph, et al.
Published: (2026)
SoloParkour: Constrained Reinforcement Learning for Visual Locomotion from Privileged Experience
by: Chane-Sane, Elliot, et al.
Published: (2024)
by: Chane-Sane, Elliot, et al.
Published: (2024)
Reinforcement Learning from Wild Animal Videos
by: Chane-Sane, Elliot, et al.
Published: (2024)
by: Chane-Sane, Elliot, et al.
Published: (2024)
CaT: Constraints as Terminations for Legged Locomotion Reinforcement Learning
by: Chane-Sane, Elliot, et al.
Published: (2024)
by: Chane-Sane, Elliot, et al.
Published: (2024)
ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards
by: Li, Fanxing, et al.
Published: (2025)
by: Li, Fanxing, et al.
Published: (2025)
Enhancing RL Generalizability in Robotics through SHAP Analysis of Algorithms and Hyperparameters
by: Kong, Lingxiao, et al.
Published: (2026)
by: Kong, Lingxiao, et al.
Published: (2026)
An Introduction to Zero-Order Optimization Techniques for Robotics
by: Jordana, Armand, et al.
Published: (2025)
by: Jordana, Armand, et al.
Published: (2025)
MaxInfoRL: Boosting exploration in reinforcement learning through information gain maximization
by: Sukhija, Bhavya, et al.
Published: (2024)
by: Sukhija, Bhavya, et al.
Published: (2024)
Investigating Memory in Model-Free RL with POPGym Arcade
by: Wang, Zekang, et al.
Published: (2025)
by: Wang, Zekang, et al.
Published: (2025)
SPAARS: Safer RL Policy Alignment through Abstract Exploration and Refined Exploitation of Action Space
by: K, Swaminathan S, et al.
Published: (2026)
by: K, Swaminathan S, et al.
Published: (2026)
Decoupled Q-Chunking
by: Li, Qiyang, et al.
Published: (2025)
by: Li, Qiyang, et al.
Published: (2025)
WorldPlanner: Monte Carlo Tree Search and MPC with Action-Conditioned Visual World Models
by: Khorrambakht, R., et al.
Published: (2025)
by: Khorrambakht, R., et al.
Published: (2025)
Generalised Linear Models in Deep Bayesian RL with Learnable Basis Functions
by: You, Jingyang, et al.
Published: (2025)
by: You, Jingyang, et al.
Published: (2025)
Aligning Text-to-Image Diffusion Models with Reward Backpropagation
by: Prabhudesai, Mihir, et al.
Published: (2023)
by: Prabhudesai, Mihir, et al.
Published: (2023)
Cost Function Estimation Using Inverse Reinforcement Learning with Minimal Observations
by: Mehrdad, Sarmad, et al.
Published: (2025)
by: Mehrdad, Sarmad, et al.
Published: (2025)
RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
by: Wang, Yufei, et al.
Published: (2024)
by: Wang, Yufei, et al.
Published: (2024)
Automatic Environment Shaping is the Next Frontier in RL
by: Park, Younghyo, et al.
Published: (2024)
by: Park, Younghyo, et al.
Published: (2024)
Q-Guided Stein Variational Model Predictive Control via RL-informed Policy Prior
by: Cai, Shizhe, et al.
Published: (2025)
by: Cai, Shizhe, et al.
Published: (2025)
Offline Learning of Controllable Diverse Behaviors
by: Petitbois, Mathieu, et al.
Published: (2025)
by: Petitbois, Mathieu, et al.
Published: (2025)
Sample-efficient and Scalable Exploration in Continuous-Time RL
by: Iten, Klemens, et al.
Published: (2025)
by: Iten, Klemens, et al.
Published: (2025)
METRA: Scalable Unsupervised RL with Metric-Aware Abstraction
by: Park, Seohong, et al.
Published: (2023)
by: Park, Seohong, et al.
Published: (2023)
Language-Conditioned Offline RL for Multi-Robot Navigation
by: Morad, Steven, et al.
Published: (2024)
by: Morad, Steven, et al.
Published: (2024)
Decoupled Action Expert: Confining Task Knowledge to the Conditioning Pathway
by: Zhou, Jian, et al.
Published: (2025)
by: Zhou, Jian, et al.
Published: (2025)
CaRL: Learning Scalable Planning Policies with Simple Rewards
by: Jaeger, Bernhard, et al.
Published: (2025)
by: Jaeger, Bernhard, et al.
Published: (2025)
GRAM: Generalization in Deep RL with a Robust Adaptation Module
by: Queeney, James, et al.
Published: (2024)
by: Queeney, James, et al.
Published: (2024)
Dual RL: Unification and New Methods for Reinforcement and Imitation Learning
by: Sikchi, Harshit, et al.
Published: (2023)
by: Sikchi, Harshit, et al.
Published: (2023)
HIQL: Offline Goal-Conditioned RL with Latent States as Actions
by: Park, Seohong, et al.
Published: (2023)
by: Park, Seohong, et al.
Published: (2023)
Navigation with QPHIL: Quantizing Planner for Hierarchical Implicit Q-Learning
by: Canesse, Alexi, et al.
Published: (2024)
by: Canesse, Alexi, et al.
Published: (2024)
Select before Act: Spatially Decoupled Action Repetition for Continuous Control
by: Nie, Buqing, et al.
Published: (2025)
by: Nie, Buqing, et al.
Published: (2025)
Posterior Behavioral Cloning: Pretraining BC Policies for Efficient RL Finetuning
by: Wagenmaker, Andrew, et al.
Published: (2025)
by: Wagenmaker, Andrew, et al.
Published: (2025)
RL-100: Performant Robotic Manipulation with Real-World Reinforcement Learning
by: Lei, Kun, et al.
Published: (2025)
by: Lei, Kun, et al.
Published: (2025)
DEAS: DEtached value learning with Action Sequence for Scalable Offline RL
by: Kim, Changyeon, et al.
Published: (2025)
by: Kim, Changyeon, et al.
Published: (2025)
ARC-RL: A Reinforcement Learning Playground Inspired by ARC Raiders
by: Romeo, Carlo, et al.
Published: (2026)
by: Romeo, Carlo, et al.
Published: (2026)
DISC: Decoupling Instruction from State-Conditioned Control via Policy Generation
by: Ren, Hanxiang, et al.
Published: (2026)
by: Ren, Hanxiang, et al.
Published: (2026)
Efficient Online RL Fine Tuning with Offline Pre-trained Policy Only
by: Xiao, Wei, et al.
Published: (2025)
by: Xiao, Wei, et al.
Published: (2025)
ELMUR: External Layer Memory with Update/Rewrite for Long-Horizon RL Problems
by: Cherepanov, Egor, et al.
Published: (2025)
by: Cherepanov, Egor, et al.
Published: (2025)
Learning Human-Like RL Agents Through Trajectory Optimization With Action Quantization
by: Guo, Jian-Ting, et al.
Published: (2025)
by: Guo, Jian-Ting, et al.
Published: (2025)
Performance Comparison of Deep RL Algorithms for Mixed Traffic Cooperative Lane-Changing
by: Yao, Xue, et al.
Published: (2024)
by: Yao, Xue, et al.
Published: (2024)
A Review of Online Diffusion Policy RL Algorithms for Scalable Robotic Control
by: Choi, Wonhyeok, et al.
Published: (2026)
by: Choi, Wonhyeok, et al.
Published: (2026)
RACER: Epistemic Risk-Sensitive RL Enables Fast Driving with Fewer Crashes
by: Stachowicz, Kyle, et al.
Published: (2024)
by: Stachowicz, Kyle, et al.
Published: (2024)
Similar Items
-
Coupled Local and Global World Models for Efficient First Order RL
by: Amigo, Joseph, et al.
Published: (2026) -
SoloParkour: Constrained Reinforcement Learning for Visual Locomotion from Privileged Experience
by: Chane-Sane, Elliot, et al.
Published: (2024) -
Reinforcement Learning from Wild Animal Videos
by: Chane-Sane, Elliot, et al.
Published: (2024) -
CaT: Constraints as Terminations for Legged Locomotion Reinforcement Learning
by: Chane-Sane, Elliot, et al.
Published: (2024) -
ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards
by: Li, Fanxing, et al.
Published: (2025)