Distilling Reinforcement Learning Policies for Interpretable Robot Locomotion: Gradient Boosting Machines and Symbolic Regression
Fuente:
arXiv
Saved in:
| Main Authors: | Acero, Fernando, Li, Zhibin |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning
by: Xu, Charles, et al.
Published: (2024)
by: Xu, Charles, et al.
Published: (2024)
Continual Policy Distillation of Reinforcement Learning-based Controllers for Soft Robotic In-Hand Manipulation
by: Li, Lanpei, et al.
Published: (2024)
by: Li, Lanpei, et al.
Published: (2024)
Enhancing Hardware Fault Tolerance in Machines with Reinforcement Learning Policy Gradient Algorithms
by: Schoepp, Sheila, et al.
Published: (2024)
by: Schoepp, Sheila, et al.
Published: (2024)
Rethinking Policy Diversity in Ensemble Policy Gradient in Large-Scale Reinforcement Learning
by: Shitanda, Naoki, et al.
Published: (2026)
by: Shitanda, Naoki, et al.
Published: (2026)
ManyQuadrupeds: Learning a Single Locomotion Policy for Diverse Quadruped Robots
by: Shafiee, Milad, et al.
Published: (2023)
by: Shafiee, Milad, et al.
Published: (2023)
Exploiting Hybrid Policy in Reinforcement Learning for Interpretable Temporal Logic Manipulation
by: Zhang, Hao, et al.
Published: (2024)
by: Zhang, Hao, et al.
Published: (2024)
Tiny Reinforcement Learning for Quadruped Locomotion using Decision Transformers
by: Akgün, Orhan Eren, et al.
Published: (2024)
by: Akgün, Orhan Eren, et al.
Published: (2024)
Towards Embodiment Scaling Laws in Robot Locomotion
by: Ai, Bo, et al.
Published: (2025)
by: Ai, Bo, et al.
Published: (2025)
Robot Policy Transfer with Online Demonstrations: An Active Reinforcement Learning Approach
by: Hou, Muhan, et al.
Published: (2025)
by: Hou, Muhan, et al.
Published: (2025)
Multi-Agent Reinforcement Learning for Unmanned Aerial Vehicle Coordination by Multi-Critic Policy Gradient Optimization
by: Alon, Yoav, et al.
Published: (2020)
by: Alon, Yoav, et al.
Published: (2020)
Not Only Rewards But Also Constraints: Applications on Legged Robot Locomotion
by: Kim, Yunho, et al.
Published: (2023)
by: Kim, Yunho, et al.
Published: (2023)
RobotKeyframing: Learning Locomotion with High-Level Objectives via Mixture of Dense and Sparse Rewards
by: Zargarbashi, Fatemeh, et al.
Published: (2024)
by: Zargarbashi, Fatemeh, et al.
Published: (2024)
Integrating Learning-Based Manipulation and Physics-Based Locomotion for Whole-Body Badminton Robot Control
by: Wang, Haochen, et al.
Published: (2025)
by: Wang, Haochen, et al.
Published: (2025)
Reinforcement Learning via Auxiliary Task Distillation
by: Harish, Abhinav Narayan, et al.
Published: (2024)
by: Harish, Abhinav Narayan, et al.
Published: (2024)
Symbolic Imitation Learning: From Black-Box to Explainable Driving Policies
by: Sharifi, Iman, et al.
Published: (2023)
by: Sharifi, Iman, et al.
Published: (2023)
Symmetry-Guided Memory Augmentation for Efficient Locomotion Learning
by: Bao, Kaixi, et al.
Published: (2025)
by: Bao, Kaixi, et al.
Published: (2025)
Leveraging Analytic Gradients in Provably Safe Reinforcement Learning
by: Walter, Tim, et al.
Published: (2025)
by: Walter, Tim, et al.
Published: (2025)
Variational Distillation of Diffusion Policies into Mixture of Experts
by: Zhou, Hongyi, et al.
Published: (2024)
by: Zhou, Hongyi, et al.
Published: (2024)
Towards Interpretable Foundation Models of Robot Behavior: A Task Specific Policy Generation Approach
by: Sheidlower, Isaac, et al.
Published: (2024)
by: Sheidlower, Isaac, et al.
Published: (2024)
Distilling On-device Language Models for Robot Planning with Minimal Human Intervention
by: Ravichandran, Zachary, et al.
Published: (2025)
by: Ravichandran, Zachary, et al.
Published: (2025)
Reinforcement Learning Within the Classical Robotics Stack: A Case Study in Robot Soccer
by: Labiosa, Adam, et al.
Published: (2024)
by: Labiosa, Adam, et al.
Published: (2024)
Adaptive Advantage-Guided Policy Regularization for Offline Reinforcement Learning
by: Liu, Tenglong, et al.
Published: (2024)
by: Liu, Tenglong, et al.
Published: (2024)
Interpretable Robotic Friction Learning via Symbolic Regression
by: Scholl, Philipp, et al.
Published: (2025)
by: Scholl, Philipp, et al.
Published: (2025)
Does "Do Differentiable Simulators Give Better Policy Gradients?'' Give Better Policy Gradients?
by: Onoda, Ku, et al.
Published: (2026)
by: Onoda, Ku, et al.
Published: (2026)
Learning Sim-to-Real Humanoid Locomotion in 15 Minutes
by: Seo, Younggyo, et al.
Published: (2025)
by: Seo, Younggyo, et al.
Published: (2025)
Unsupervised Skill Discovery as Exploration for Learning Agile Locomotion
by: Rho, Seungeun, et al.
Published: (2025)
by: Rho, Seungeun, et al.
Published: (2025)
Neuro-Symbolic Imitation Learning: Discovering Symbolic Abstractions for Skill Learning
by: Keller, Leon, et al.
Published: (2025)
by: Keller, Leon, et al.
Published: (2025)
Robot Policy Learning with Temporal Optimal Transport Reward
by: Fu, Yuwei, et al.
Published: (2024)
by: Fu, Yuwei, et al.
Published: (2024)
Body Transformer: Leveraging Robot Embodiment for Policy Learning
by: Sferrazza, Carmelo, et al.
Published: (2024)
by: Sferrazza, Carmelo, et al.
Published: (2024)
Uncertainty-Aware Robotic World Model Makes Offline Model-Based Reinforcement Learning Work on Real Robots
by: Li, Chenhao, et al.
Published: (2025)
by: Li, Chenhao, et al.
Published: (2025)
ReinforceGen: Hybrid Skill Policies with Automated Data Generation and Reinforcement Learning
by: Zhou, Zihan, et al.
Published: (2025)
by: Zhou, Zihan, et al.
Published: (2025)
Mutual Information Tracks Policy Coherence in Reinforcement Learning
by: Reid, Cameron, et al.
Published: (2025)
by: Reid, Cameron, et al.
Published: (2025)
Random Network Distillation Based Deep Reinforcement Learning for AGV Path Planning
by: Yin, Huilin, et al.
Published: (2024)
by: Yin, Huilin, et al.
Published: (2024)
Tool-as-Interface: Learning Robot Policies from Observing Human Tool Use
by: Chen, Haonan, et al.
Published: (2025)
by: Chen, Haonan, et al.
Published: (2025)
Mollification Effects of Policy Gradient Methods
by: Wang, Tao, et al.
Published: (2024)
by: Wang, Tao, et al.
Published: (2024)
Research on Autonomous Robots Navigation based on Reinforcement Learning
by: Wang, Zixiang, et al.
Published: (2024)
by: Wang, Zixiang, et al.
Published: (2024)
BeamDojo: Learning Agile Humanoid Locomotion on Sparse Footholds
by: Wang, Huayi, et al.
Published: (2025)
by: Wang, Huayi, et al.
Published: (2025)
Robotic Manipulation Datasets for Offline Compositional Reinforcement Learning
by: Hussing, Marcel, et al.
Published: (2023)
by: Hussing, Marcel, et al.
Published: (2023)
Safe Reinforcement Learning in a Simulated Robotic Arm
by: Kovač, Luka, et al.
Published: (2023)
by: Kovač, Luka, et al.
Published: (2023)
Human-Aware Robot Navigation via Reinforcement Learning with Hindsight Experience Replay and Curriculum Learning
by: Li, Keyu, et al.
Published: (2021)
by: Li, Keyu, et al.
Published: (2021)
Similar Items
-
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning
by: Xu, Charles, et al.
Published: (2024) -
Continual Policy Distillation of Reinforcement Learning-based Controllers for Soft Robotic In-Hand Manipulation
by: Li, Lanpei, et al.
Published: (2024) -
Enhancing Hardware Fault Tolerance in Machines with Reinforcement Learning Policy Gradient Algorithms
by: Schoepp, Sheila, et al.
Published: (2024) -
Rethinking Policy Diversity in Ensemble Policy Gradient in Large-Scale Reinforcement Learning
by: Shitanda, Naoki, et al.
Published: (2026) -
ManyQuadrupeds: Learning a Single Locomotion Policy for Diverse Quadruped Robots
by: Shafiee, Milad, et al.
Published: (2023)