Adapting the Behavior of Reinforcement Learning Agents to Changing Action Spaces and Reward Functions
Fuente:
arXiv
Saved in:
| Main Authors: | de la Rosa, Raul, Dusparic, Ivana, Cardozo, Nicolas |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TREX: Trajectory Explanations for Multi-Objective Reinforcement Learning
by: Rajapakse, Dilina, et al.
Published: (2026)
by: Rajapakse, Dilina, et al.
Published: (2026)
ACTER: Diverse and Actionable Counterfactual Sequences for Explaining and Diagnosing RL Policies
by: Gajcin, Jasmina, et al.
Published: (2024)
by: Gajcin, Jasmina, et al.
Published: (2024)
Optimistic Exploration for Risk-Averse Constrained Reinforcement Learning
by: McCarthy, James, et al.
Published: (2025)
by: McCarthy, James, et al.
Published: (2025)
Continual Reinforcement Learning for Cyber-Physical Systems: Lessons Learned and Open Challenges
by: Nolle, Kim N., et al.
Published: (2025)
by: Nolle, Kim N., et al.
Published: (2025)
Multi-Objective Deep Reinforcement Learning for Optimisation in Autonomous Systems
by: Rosero, Juan C., et al.
Published: (2024)
by: Rosero, Juan C., et al.
Published: (2024)
Drama: Mamba-Enabled Model-Based Reinforcement Learning Is Sample and Parameter Efficient
by: Wang, Wenlong, et al.
Published: (2024)
by: Wang, Wenlong, et al.
Published: (2024)
Redefining Counterfactual Explanations for Reinforcement Learning: Overview, Challenges and Opportunities
by: Gajcin, Jasmina, et al.
Published: (2022)
by: Gajcin, Jasmina, et al.
Published: (2022)
Semifactual Explanations for Reinforcement Learning
by: Gajcin, Jasmina, et al.
Published: (2024)
by: Gajcin, Jasmina, et al.
Published: (2024)
Penalizing Infeasible Actions and Reward Scaling in Reinforcement Learning with Offline Data
by: Kim, Jeonghye, et al.
Published: (2025)
by: Kim, Jeonghye, et al.
Published: (2025)
Constraints as Rewards: Reinforcement Learning for Robots without Reward Functions
by: Ishihara, Yu, et al.
Published: (2025)
by: Ishihara, Yu, et al.
Published: (2025)
In-Context Reinforcement Learning for Variable Action Spaces
by: Sinii, Viacheslav, et al.
Published: (2023)
by: Sinii, Viacheslav, et al.
Published: (2023)
Provably Adaptive Average Reward Reinforcement Learning for Metric Spaces
by: Kar, Avik, et al.
Published: (2024)
by: Kar, Avik, et al.
Published: (2024)
Model-based Reinforcement Learning for Parameterized Action Spaces
by: Zhang, Renhao, et al.
Published: (2024)
by: Zhang, Renhao, et al.
Published: (2024)
On the Geometry of Reinforcement Learning in Continuous State and Action Spaces
by: Tiwari, Saket, et al.
Published: (2022)
by: Tiwari, Saket, et al.
Published: (2022)
Generalizing Behavior via Inverse Reinforcement Learning with Closed-Form Reward Centroids
by: Lazzati, Filippo, et al.
Published: (2025)
by: Lazzati, Filippo, et al.
Published: (2025)
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors
by: Ke, Jingyang, et al.
Published: (2025)
by: Ke, Jingyang, et al.
Published: (2025)
Model-Based Reinforcement Learning in Discrete-Action Non-Markovian Reward Decision Processes
by: Trapasso, Alessandro, et al.
Published: (2025)
by: Trapasso, Alessandro, et al.
Published: (2025)
Context-Aware Model-Based Reinforcement Learning for Autonomous Racing
by: Moustafa, Emran Yasser, et al.
Published: (2025)
by: Moustafa, Emran Yasser, et al.
Published: (2025)
Geometry of Neural Reinforcement Learning in Continuous State and Action Spaces
by: Tiwari, Saket, et al.
Published: (2025)
by: Tiwari, Saket, et al.
Published: (2025)
AI Alignment with Changing and Influenceable Reward Functions
by: Carroll, Micah, et al.
Published: (2024)
by: Carroll, Micah, et al.
Published: (2024)
Unsupervised Zero-Shot Reinforcement Learning via Functional Reward Encodings
by: Frans, Kevin, et al.
Published: (2024)
by: Frans, Kevin, et al.
Published: (2024)
Towards Adapting Reinforcement Learning Agents to New Tasks: Insights from Q-Values
by: Ramaswamy, Ashwin, et al.
Published: (2024)
by: Ramaswamy, Ashwin, et al.
Published: (2024)
Balancing Multiple Objectives in Urban Traffic Control with Reinforcement Learning from AI Feedback
by: Zhao, Chenyang, et al.
Published: (2026)
by: Zhao, Chenyang, et al.
Published: (2026)
Learning to Explore with Parameter-Space Noise: A Deep Dive into Parameter-Space Noise for Reinforcement Learning with Verifiable Rewards
by: Bai, Bizhe, et al.
Published: (2026)
by: Bai, Bizhe, et al.
Published: (2026)
Centralized Reward Agent for Knowledge Sharing and Transfer in Multi-Task Reinforcement Learning
by: Ma, Haozhe, et al.
Published: (2024)
by: Ma, Haozhe, et al.
Published: (2024)
Transfer Reinforcement Learning in Heterogeneous Action Spaces using Subgoal Mapping
by: Sivakumar, Kavinayan P., et al.
Published: (2024)
by: Sivakumar, Kavinayan P., et al.
Published: (2024)
An Advantage-based Optimization Method for Reinforcement Learning in Large Action Space
by: Lin, Hai, et al.
Published: (2024)
by: Lin, Hai, et al.
Published: (2024)
Robust Deep Reinforcement Learning with Adaptive Adversarial Perturbations in Action Space
by: Liu, Qianmei, et al.
Published: (2024)
by: Liu, Qianmei, et al.
Published: (2024)
Proximal Action Replacement for Behavior Cloning Actor-Critic in Offline Reinforcement Learning
by: Dong, Jinzong, et al.
Published: (2026)
by: Dong, Jinzong, et al.
Published: (2026)
BCR-DRL: Behavior- and Context-aware Reward for Deep Reinforcement Learning in Human-AI Coordination
by: Hao, Xin, et al.
Published: (2024)
by: Hao, Xin, et al.
Published: (2024)
Tiered Reward: Designing Rewards for Specification and Fast Learning of Desired Behavior
by: Zhou, Zhiyuan, et al.
Published: (2022)
by: Zhou, Zhiyuan, et al.
Published: (2022)
RLVMR: Reinforcement Learning with Verifiable Meta-Reasoning Rewards for Robust Long-Horizon Agents
by: Zhang, Zijing, et al.
Published: (2025)
by: Zhang, Zijing, et al.
Published: (2025)
Breaking the Grid: Distance-Guided Reinforcement Learning in Large Discrete Action Spaces
by: Hoppe, Heiko, et al.
Published: (2026)
by: Hoppe, Heiko, et al.
Published: (2026)
Discretizing Continuous Action Space with Unimodal Probability Distributions for On-Policy Reinforcement Learning
by: Zhu, Yuanyang, et al.
Published: (2024)
by: Zhu, Yuanyang, et al.
Published: (2024)
A Review of Reward Functions for Reinforcement Learning in the context of Autonomous Driving
by: Abouelazm, Ahmed, et al.
Published: (2024)
by: Abouelazm, Ahmed, et al.
Published: (2024)
Reinforcement Learning with Symbolic Reward Machines
by: Krug, Thomas, et al.
Published: (2026)
by: Krug, Thomas, et al.
Published: (2026)
Kernel-Based Function Approximation for Average Reward Reinforcement Learning: An Optimist No-Regret Algorithm
by: Vakili, Sattar, et al.
Published: (2024)
by: Vakili, Sattar, et al.
Published: (2024)
Reinforcement Learning with Exogenous States and Rewards
by: Trimponias, George, et al.
Published: (2023)
by: Trimponias, George, et al.
Published: (2023)
Reinforcement Learning with Stochastic Reward Machines
by: Corazza, Jan, et al.
Published: (2025)
by: Corazza, Jan, et al.
Published: (2025)
Offline Reinforcement Learning with Imputed Rewards
by: Romeo, Carlo, et al.
Published: (2024)
by: Romeo, Carlo, et al.
Published: (2024)
Similar Items
-
TREX: Trajectory Explanations for Multi-Objective Reinforcement Learning
by: Rajapakse, Dilina, et al.
Published: (2026) -
ACTER: Diverse and Actionable Counterfactual Sequences for Explaining and Diagnosing RL Policies
by: Gajcin, Jasmina, et al.
Published: (2024) -
Optimistic Exploration for Risk-Averse Constrained Reinforcement Learning
by: McCarthy, James, et al.
Published: (2025) -
Continual Reinforcement Learning for Cyber-Physical Systems: Lessons Learned and Open Challenges
by: Nolle, Kim N., et al.
Published: (2025) -
Multi-Objective Deep Reinforcement Learning for Optimisation in Autonomous Systems
by: Rosero, Juan C., et al.
Published: (2024)