Towards Adapting Reinforcement Learning Agents to New Tasks: Insights from Q-Values
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ramaswamy, Ashwin, Senanayake, Ransalu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The Role of Predictive Uncertainty and Diversity in Embodied AI and Robot Learning
von: Senanayake, Ransalu
Veröffentlicht: (2024)
von: Senanayake, Ransalu
Veröffentlicht: (2024)
RiskQ: Risk-sensitive Multi-Agent Reinforcement Learning Value Factorization
von: Shen, Siqi, et al.
Veröffentlicht: (2023)
von: Shen, Siqi, et al.
Veröffentlicht: (2023)
Modeling Multi-Objective Tradeoffs with Monotonic Utility Functions
von: Chen, Edward, et al.
Veröffentlicht: (2024)
von: Chen, Edward, et al.
Veröffentlicht: (2024)
Adapting the Behavior of Reinforcement Learning Agents to Changing Action Spaces and Reward Functions
von: de la Rosa, Raul, et al.
Veröffentlicht: (2026)
von: de la Rosa, Raul, et al.
Veröffentlicht: (2026)
Exploiting Estimation Bias in Clipped Double Q-Learning for Continous Control Reinforcement Learning Tasks
von: Turcato, Niccolò, et al.
Veröffentlicht: (2024)
von: Turcato, Niccolò, et al.
Veröffentlicht: (2024)
Expert Q-learning: Deep Reinforcement Learning with Coarse State Values from Offline Expert Examples
von: Meng, Li, et al.
Veröffentlicht: (2021)
von: Meng, Li, et al.
Veröffentlicht: (2021)
Solving New Tasks by Adapting Internet Video Knowledge
von: Luo, Calvin, et al.
Veröffentlicht: (2025)
von: Luo, Calvin, et al.
Veröffentlicht: (2025)
Pessimistic Value Iteration for Multi-Task Data Sharing in Offline Reinforcement Learning
von: Bai, Chenjia, et al.
Veröffentlicht: (2024)
von: Bai, Chenjia, et al.
Veröffentlicht: (2024)
LLM Routing as Reasoning: A MaxSAT View
von: Nguyen, Son, et al.
Veröffentlicht: (2026)
von: Nguyen, Son, et al.
Veröffentlicht: (2026)
Towards Safe Reinforcement Learning via Constraining Conditional Value-at-Risk
von: Ying, Chengyang, et al.
Veröffentlicht: (2022)
von: Ying, Chengyang, et al.
Veröffentlicht: (2022)
Is there Value in Reinforcement Learning?
von: Fox, Lior, et al.
Veröffentlicht: (2025)
von: Fox, Lior, et al.
Veröffentlicht: (2025)
Towards Principled Unsupervised Multi-Agent Reinforcement Learning
von: Zamboni, Riccardo, et al.
Veröffentlicht: (2025)
von: Zamboni, Riccardo, et al.
Veröffentlicht: (2025)
Reinforcement Learning Interventions on Boundedly Rational Human Agents in Frictionful Tasks
von: Nofshin, Eura, et al.
Veröffentlicht: (2024)
von: Nofshin, Eura, et al.
Veröffentlicht: (2024)
PAC Bench: Do Foundation Models Understand Prerequisites for Executing Manipulation Policies?
von: Gundawar, Atharva, et al.
Veröffentlicht: (2025)
von: Gundawar, Atharva, et al.
Veröffentlicht: (2025)
AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations
von: Verma, Gaurav, et al.
Veröffentlicht: (2024)
von: Verma, Gaurav, et al.
Veröffentlicht: (2024)
Centralized Reward Agent for Knowledge Sharing and Transfer in Multi-Task Reinforcement Learning
von: Ma, Haozhe, et al.
Veröffentlicht: (2024)
von: Ma, Haozhe, et al.
Veröffentlicht: (2024)
In-Context Compositional Q-Learning for Offline Reinforcement Learning
von: Xu, Qiushui, et al.
Veröffentlicht: (2025)
von: Xu, Qiushui, et al.
Veröffentlicht: (2025)
Imagination-Limited Q-Learning for Offline Reinforcement Learning
von: Liu, Wenhui, et al.
Veröffentlicht: (2025)
von: Liu, Wenhui, et al.
Veröffentlicht: (2025)
Mildly Conservative Q-Learning for Offline Reinforcement Learning
von: Lyu, Jiafei, et al.
Veröffentlicht: (2022)
von: Lyu, Jiafei, et al.
Veröffentlicht: (2022)
Decoding Generalization from Memorization in Deep Neural Networks
von: Ketha, Simran, et al.
Veröffentlicht: (2025)
von: Ketha, Simran, et al.
Veröffentlicht: (2025)
Failures Are Fated, But Can Be Faded: Characterizing and Mitigating Unwanted Behaviors in Large-Scale Vision and Language Models
von: Sagar, Som, et al.
Veröffentlicht: (2024)
von: Sagar, Som, et al.
Veröffentlicht: (2024)
LLM-Assisted Red Teaming of Diffusion Models through "Failures Are Fated, But Can Be Faded"
von: Sagar, Som, et al.
Veröffentlicht: (2024)
von: Sagar, Som, et al.
Veröffentlicht: (2024)
CUPID in the Model Zoo: Online Matchmaking for Selecting Your Dream LLM
von: Nguyen, Son, et al.
Veröffentlicht: (2026)
von: Nguyen, Son, et al.
Veröffentlicht: (2026)
SafeAdapt: Provably Safe Policy Updates in Deep Reinforcement Learning
von: Anisimov, Maksim, et al.
Veröffentlicht: (2026)
von: Anisimov, Maksim, et al.
Veröffentlicht: (2026)
Continual Reinforcement Learning via Autoencoder-Driven Task and New Environment Recognition
von: Erden, Zeki Doruk, et al.
Veröffentlicht: (2025)
von: Erden, Zeki Doruk, et al.
Veröffentlicht: (2025)
Agent Q: Advanced Reasoning and Learning for Autonomous AI Agents
von: Putta, Pranav, et al.
Veröffentlicht: (2024)
von: Putta, Pranav, et al.
Veröffentlicht: (2024)
Logical Specifications-guided Dynamic Task Sampling for Reinforcement Learning Agents
von: Shukla, Yash, et al.
Veröffentlicht: (2024)
von: Shukla, Yash, et al.
Veröffentlicht: (2024)
Adaptive Neighborhood-Constrained Q Learning for Offline Reinforcement Learning
von: Mao, Yixiu, et al.
Veröffentlicht: (2025)
von: Mao, Yixiu, et al.
Veröffentlicht: (2025)
Bellman Unbiasedness: Toward Provably Efficient Distributional Reinforcement Learning with General Value Function Approximation
von: Cho, Taehyun, et al.
Veröffentlicht: (2024)
von: Cho, Taehyun, et al.
Veröffentlicht: (2024)
Task Scheduling & Forgetting in Multi-Task Reinforcement Learning
von: Speckmann, Marc, et al.
Veröffentlicht: (2025)
von: Speckmann, Marc, et al.
Veröffentlicht: (2025)
Explainable Concept Generation through Vision-Language Preference Learning for Understanding Neural Networks' Internal Representations
von: Taparia, Aditya, et al.
Veröffentlicht: (2024)
von: Taparia, Aditya, et al.
Veröffentlicht: (2024)
Exclusively Penalized Q-learning for Offline Reinforcement Learning
von: Yeom, Junghyuk, et al.
Veröffentlicht: (2024)
von: Yeom, Junghyuk, et al.
Veröffentlicht: (2024)
Residual Q-Learning: Offline and Online Policy Customization without Value
von: Li, Chenran, et al.
Veröffentlicht: (2023)
von: Li, Chenran, et al.
Veröffentlicht: (2023)
Towards Reinforcement Learning from Neural Feedback: Mapping fNIRS Signals to Agent Performance
von: Santaniello, Julia, et al.
Veröffentlicht: (2025)
von: Santaniello, Julia, et al.
Veröffentlicht: (2025)
The Composite Task Challenge for Cooperative Multi-Agent Reinforcement Learning
von: Li, Yurui, et al.
Veröffentlicht: (2025)
von: Li, Yurui, et al.
Veröffentlicht: (2025)
Model-based Offline Reinforcement Learning with Lower Expectile Q-Learning
von: Park, Kwanyoung, et al.
Veröffentlicht: (2024)
von: Park, Kwanyoung, et al.
Veröffentlicht: (2024)
Fast Value Tracking for Deep Reinforcement Learning
von: Shih, Frank, et al.
Veröffentlicht: (2024)
von: Shih, Frank, et al.
Veröffentlicht: (2024)
Value-Distributional Model-Based Reinforcement Learning
von: Luis, Carlos E., et al.
Veröffentlicht: (2023)
von: Luis, Carlos E., et al.
Veröffentlicht: (2023)
Reinforcement Learning via Value Gradient Flow
von: Xu, Haoran, et al.
Veröffentlicht: (2026)
von: Xu, Haoran, et al.
Veröffentlicht: (2026)
Task Specific Sharpness Aware O-RAN Resource Management using Multi Agent Reinforcement Learning
von: Lotfi, Fatemeh, et al.
Veröffentlicht: (2025)
von: Lotfi, Fatemeh, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
The Role of Predictive Uncertainty and Diversity in Embodied AI and Robot Learning
von: Senanayake, Ransalu
Veröffentlicht: (2024) -
RiskQ: Risk-sensitive Multi-Agent Reinforcement Learning Value Factorization
von: Shen, Siqi, et al.
Veröffentlicht: (2023) -
Modeling Multi-Objective Tradeoffs with Monotonic Utility Functions
von: Chen, Edward, et al.
Veröffentlicht: (2024) -
Adapting the Behavior of Reinforcement Learning Agents to Changing Action Spaces and Reward Functions
von: de la Rosa, Raul, et al.
Veröffentlicht: (2026) -
Exploiting Estimation Bias in Clipped Double Q-Learning for Continous Control Reinforcement Learning Tasks
von: Turcato, Niccolò, et al.
Veröffentlicht: (2024)