Decoupling Task and Behavior: A Two-Stage Reward Curriculum in Reinforcement Learning for Robotics
Fuente:
arXiv
Saved in:
| Main Authors: | Freitag, Kilian, Åkesson, Knut, Chehreghani, Morteza Haghir |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Curriculum Reinforcement Learning for Complex Reward Functions
by: Freitag, Kilian, et al.
Published: (2024)
by: Freitag, Kilian, et al.
Published: (2024)
Tactical Decision Making for Autonomous Trucks by Deep Reinforcement Learning with Total Cost of Operation Based Reward
by: Pathare, Deepthi, et al.
Published: (2024)
by: Pathare, Deepthi, et al.
Published: (2024)
Two-Stage Learned Decomposition for Scalable Routing on Multigraphs
by: Rydin, Filip, et al.
Published: (2026)
by: Rydin, Filip, et al.
Published: (2026)
Hierarchical Correlation Clustering and Tree Preserving Embedding
by: Chehreghani, Morteza Haghir, et al.
Published: (2020)
by: Chehreghani, Morteza Haghir, et al.
Published: (2020)
Shift of Pairwise Similarities for Data Clustering
by: Chehreghani, Morteza Haghir
Published: (2021)
by: Chehreghani, Morteza Haghir
Published: (2021)
Correlation Clustering with Active Learning of Pairwise Similarities
by: Aronsson, Linus, et al.
Published: (2023)
by: Aronsson, Linus, et al.
Published: (2023)
Non-Myopic Active Feature Acquisition via Pathwise Policy Gradients
by: Aronsson, Linus, et al.
Published: (2026)
by: Aronsson, Linus, et al.
Published: (2026)
Information-Theoretic Active Correlation Clustering
by: Aronsson, Linus, et al.
Published: (2024)
by: Aronsson, Linus, et al.
Published: (2024)
Adaptive Prior Selection in Gaussian Process Bandits with Thompson Sampling
by: Sandberg, Jack, et al.
Published: (2025)
by: Sandberg, Jack, et al.
Published: (2025)
Multi-Objective Reinforcement Learning for Tactical Decision Making for Trucks in Highway Traffic
by: Pathare, Deepthi, et al.
Published: (2026)
by: Pathare, Deepthi, et al.
Published: (2026)
Passive and Active Learning of Driver Behavior from Electric Vehicles
by: Comuni, Federica, et al.
Published: (2022)
by: Comuni, Federica, et al.
Published: (2022)
An Efficient Local Search Approach for Polarized Community Discovery in Signed Networks
by: Aronsson, Linus, et al.
Published: (2025)
by: Aronsson, Linus, et al.
Published: (2025)
Utilizing Reinforcement Learning for de novo Drug Design
by: Svensson, Hampus Gummesson, et al.
Published: (2023)
by: Svensson, Hampus Gummesson, et al.
Published: (2023)
A Survey on Active Feature Acquisition Strategies
by: Aronsson, Linus, et al.
Published: (2025)
by: Aronsson, Linus, et al.
Published: (2025)
Diversity-Aware Reinforcement Learning for de novo Drug Design
by: Svensson, Hampus Gummesson, et al.
Published: (2024)
by: Svensson, Hampus Gummesson, et al.
Published: (2024)
Cost-Efficient Online Decision Making: A Combinatorial Multi-Armed Bandit Approach
by: Rahbar, Arman, et al.
Published: (2023)
by: Rahbar, Arman, et al.
Published: (2023)
On-Robot Reinforcement Learning with Goal-Contrastive Rewards
by: Biza, Ondrej, et al.
Published: (2024)
by: Biza, Ondrej, et al.
Published: (2024)
On the Complexity of Optimal Graph Rewiring for Oversmoothing and Oversquashing in Graph Neural Networks
by: Chehreghani, Mostafa Haghir
Published: (2026)
by: Chehreghani, Mostafa Haghir
Published: (2026)
Generalization in Deep Reinforcement Learning for Robotic Navigation by Reward Shaping
by: Miranda, Victor R. F., et al.
Published: (2022)
by: Miranda, Victor R. F., et al.
Published: (2022)
Bayesian Analysis of Combinatorial Gaussian Process Bandits
by: Sandberg, Jack, et al.
Published: (2023)
by: Sandberg, Jack, et al.
Published: (2023)
Efficient Online Decision Tree Learning with Active Feature Acquisition
by: Rahbar, Arman, et al.
Published: (2023)
by: Rahbar, Arman, et al.
Published: (2023)
DrS: Learning Reusable Dense Rewards for Multi-Stage Tasks
by: Mu, Tongzhou, et al.
Published: (2024)
by: Mu, Tongzhou, et al.
Published: (2024)
Constraints as Rewards: Reinforcement Learning for Robots without Reward Functions
by: Ishihara, Yu, et al.
Published: (2025)
by: Ishihara, Yu, et al.
Published: (2025)
Robotic Skill Diversification via Active Mutation of Reward Functions in Reinforcement Learning During a Liquid Pouring Task
by: van Buuren, Jannick, et al.
Published: (2025)
by: van Buuren, Jannick, et al.
Published: (2025)
Average-Reward Maximum Entropy Reinforcement Learning for Underactuated Double Pendulum Tasks
by: Choe, Jean Seong Bjorn, et al.
Published: (2024)
by: Choe, Jean Seong Bjorn, et al.
Published: (2024)
Self-Supervised Curriculum Generation for Autonomous Reinforcement Learning without Task-Specific Knowledge
by: Lee, Sang-Hyun, et al.
Published: (2023)
by: Lee, Sang-Hyun, et al.
Published: (2023)
Cold-Start Active Correlation Clustering
by: Aronsson, Linus, et al.
Published: (2025)
by: Aronsson, Linus, et al.
Published: (2025)
Interactive Trajectory Planning with Learning-based Distributionally Robust Model Predictive Control and Markov Systems
by: Börve, Erik, et al.
Published: (2026)
by: Börve, Erik, et al.
Published: (2026)
REBEL: Reward Regularization-Based Approach for Robotic Reinforcement Learning from Human Feedback
by: Chakraborty, Souradip, et al.
Published: (2023)
by: Chakraborty, Souradip, et al.
Published: (2023)
Diverse Mini-Batch Selection in Reinforcement Learning for Efficient Chemical Exploration in de novo Drug Design
by: Svensson, Hampus Gummesson, et al.
Published: (2025)
by: Svensson, Hampus Gummesson, et al.
Published: (2025)
Learning a High-quality Robotic Wiping Policy Using Systematic Reward Analysis and Visual-Language Model Based Curriculum
by: Liu, Yihong, et al.
Published: (2025)
by: Liu, Yihong, et al.
Published: (2025)
Future-Oriented Navigation: Dynamic Obstacle Avoidance with One-Shot Energy-Based Multimodal Motion Prediction
by: Zhang, Ze, et al.
Published: (2025)
by: Zhang, Ze, et al.
Published: (2025)
Analysing the Behaviour of Tree-Based Neural Networks in Regression Tasks
by: Samoaa, Peter, et al.
Published: (2024)
by: Samoaa, Peter, et al.
Published: (2024)
A Benchmark Dataset for Graph Regression with Homogeneous and Multi-Relational Variants
by: Samoaa, Peter, et al.
Published: (2025)
by: Samoaa, Peter, et al.
Published: (2025)
STRIDE: Automating Reward Design, Deep Reinforcement Learning Training and Feedback Optimization in Humanoid Robotics Locomotion
by: Wu, Zhenwei, et al.
Published: (2025)
by: Wu, Zhenwei, et al.
Published: (2025)
Online Learning Models for Vehicle Usage Prediction During COVID-19
by: Lindroth, Tobias, et al.
Published: (2022)
by: Lindroth, Tobias, et al.
Published: (2022)
Quantum Deep Reinforcement Learning for Robot Navigation Tasks
by: Hohenfeld, Hans, et al.
Published: (2022)
by: Hohenfeld, Hans, et al.
Published: (2022)
Tree Ensembles for Contextual Bandits
by: Nilsson, Hannes, et al.
Published: (2024)
by: Nilsson, Hannes, et al.
Published: (2024)
Graph Neural Networks in Multi-Omics Cancer Research: A Structured Survey
by: Zohari, Payam, et al.
Published: (2025)
by: Zohari, Payam, et al.
Published: (2025)
Solving Robotics Tasks with Prior Demonstration via Exploration-Efficient Deep Reinforcement Learning
by: Shen, Chengyandan, et al.
Published: (2025)
by: Shen, Chengyandan, et al.
Published: (2025)
Similar Items
-
Curriculum Reinforcement Learning for Complex Reward Functions
by: Freitag, Kilian, et al.
Published: (2024) -
Tactical Decision Making for Autonomous Trucks by Deep Reinforcement Learning with Total Cost of Operation Based Reward
by: Pathare, Deepthi, et al.
Published: (2024) -
Two-Stage Learned Decomposition for Scalable Routing on Multigraphs
by: Rydin, Filip, et al.
Published: (2026) -
Hierarchical Correlation Clustering and Tree Preserving Embedding
by: Chehreghani, Morteza Haghir, et al.
Published: (2020) -
Shift of Pairwise Similarities for Data Clustering
by: Chehreghani, Morteza Haghir
Published: (2021)