Learning Rate-Free Reinforcement Learning: A Case for Model Selection with Non-Stationary Objectives
Fuente:
arXiv
Saved in:
| Main Authors: | Afshar, Aida, Pacchiano, Aldo |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Improved Training Mechanism for Reinforcement Learning via Online Model Selection
by: Afshar, Aida, et al.
Published: (2025)
by: Afshar, Aida, et al.
Published: (2025)
State-free Reinforcement Learning
by: Chen, Mingyu, et al.
Published: (2024)
by: Chen, Mingyu, et al.
Published: (2024)
Bayesian Online Model Selection
by: Afshar, Aida, et al.
Published: (2026)
by: Afshar, Aida, et al.
Published: (2026)
Second Order Bounds for Contextual Bandits with Function Approximation
by: Pacchiano, Aldo
Published: (2024)
by: Pacchiano, Aldo
Published: (2024)
Data-Driven Online Model Selection With Regret Guarantees
by: Pacchiano, Aldo, et al.
Published: (2023)
by: Pacchiano, Aldo, et al.
Published: (2023)
In-Context Learning for Pure Exploration
by: Russo, Alessio, et al.
Published: (2025)
by: Russo, Alessio, et al.
Published: (2025)
DeLF: Designing Learning Environments with Foundation Models
by: Afshar, Aida, et al.
Published: (2024)
by: Afshar, Aida, et al.
Published: (2024)
Adaptive Exploration for Multi-Reward Multi-Policy Evaluation
by: Russo, Alessio, et al.
Published: (2025)
by: Russo, Alessio, et al.
Published: (2025)
Is Prior-Free Black-Box Non-Stationary Reinforcement Learning Feasible?
by: Gerogiannis, Argyrios, et al.
Published: (2024)
by: Gerogiannis, Argyrios, et al.
Published: (2024)
In-Context Learning for Pure Exploration in Continuous Spaces
by: Russo, Alessio, et al.
Published: (2026)
by: Russo, Alessio, et al.
Published: (2026)
Provable Interactive Learning with Hindsight Instruction Feedback
by: Misra, Dipendra, et al.
Published: (2024)
by: Misra, Dipendra, et al.
Published: (2024)
Inverse Reinforcement Learning from Non-Stationary Learning Agents
by: Sivakumar, Kavinayan P., et al.
Published: (2024)
by: Sivakumar, Kavinayan P., et al.
Published: (2024)
Online Reinforcement Learning in Non-Stationary Context-Driven Environments
by: Hamadanian, Pouya, et al.
Published: (2023)
by: Hamadanian, Pouya, et al.
Published: (2023)
Multiple-policy Evaluation via Density Estimation
by: Chen, Yilei, et al.
Published: (2024)
by: Chen, Yilei, et al.
Published: (2024)
Experiment Planning with Function Approximation
by: Pacchiano, Aldo, et al.
Published: (2024)
by: Pacchiano, Aldo, et al.
Published: (2024)
A Theoretical Framework for Partially Observed Reward-States in RLHF
by: Kausik, Chinmaya, et al.
Published: (2024)
by: Kausik, Chinmaya, et al.
Published: (2024)
Constrained Reinforcement Learning with Average Reward Objective: Model-Based and Model-Free Algorithms
by: Aggarwal, Vaneet, et al.
Published: (2024)
by: Aggarwal, Vaneet, et al.
Published: (2024)
Tracking Drift: Variation-Aware Entropy Scheduling for Non-Stationary Reinforcement Learning
by: Wang, Tongxi, et al.
Published: (2026)
by: Wang, Tongxi, et al.
Published: (2026)
ORSO: Accelerating Reward Design via Online Reward Selection and Policy Optimization
by: Zhang, Chen Bo Calvin, et al.
Published: (2024)
by: Zhang, Chen Bo Calvin, et al.
Published: (2024)
TIMRL: A Novel Meta-Reinforcement Learning Framework for Non-Stationary and Multi-Task Environments
by: Qi, Chenyang, et al.
Published: (2025)
by: Qi, Chenyang, et al.
Published: (2025)
Reinforcement Learning in Switching Non-Stationary Markov Decision Processes: Algorithms and Convergence Analysis
by: Amiri, Mohsen, et al.
Published: (2025)
by: Amiri, Mohsen, et al.
Published: (2025)
Polychromic Objectives for Reinforcement Learning
by: Hamid, Jubayer Ibn, et al.
Published: (2025)
by: Hamid, Jubayer Ibn, et al.
Published: (2025)
In-Context Learning for Non-Stationary MIMO Equalization
by: Jiang, Jiachen, et al.
Published: (2025)
by: Jiang, Jiachen, et al.
Published: (2025)
The Good, the Bad, and the Sampled: a No-Regret Approach to Safe Online Classification
by: Baharav, Tavor Z., et al.
Published: (2025)
by: Baharav, Tavor Z., et al.
Published: (2025)
The Max-Min Formulation of Multi-Objective Reinforcement Learning: From Theory to a Model-Free Algorithm
by: Park, Giseung, et al.
Published: (2024)
by: Park, Giseung, et al.
Published: (2024)
Optimistic Reinforcement Learning with Quantile Objectives
by: Alipour-Vaezi, Mohammad, et al.
Published: (2025)
by: Alipour-Vaezi, Mohammad, et al.
Published: (2025)
Automated Reinforcement Learning: An Overview
by: Afshar, Reza Refaei, et al.
Published: (2022)
by: Afshar, Reza Refaei, et al.
Published: (2022)
ScheduleFree+: Scaling Learning-Rate-Free & Schedule-Free Learning to Large Language Models
by: Defazio, Aaron
Published: (2026)
by: Defazio, Aaron
Published: (2026)
Improving Intrinsic Exploration by Creating Stationary Objectives
by: Castanyer, Roger Creus, et al.
Published: (2023)
by: Castanyer, Roger Creus, et al.
Published: (2023)
Reinforcement Learning with Non-Cumulative Objective
by: Cui, Wei, et al.
Published: (2023)
by: Cui, Wei, et al.
Published: (2023)
Multi-Label Transfer Learning in Non-Stationary Data Streams
by: Du, Honghui, et al.
Published: (2025)
by: Du, Honghui, et al.
Published: (2025)
Principled Fine-tuning of LLMs from User-Edits: A Medley of Preference, Supervision, and Reward
by: Misra, Dipendra, et al.
Published: (2026)
by: Misra, Dipendra, et al.
Published: (2026)
Demonstration Guided Multi-Objective Reinforcement Learning
by: Lu, Junlin, et al.
Published: (2024)
by: Lu, Junlin, et al.
Published: (2024)
Reinforcement Learning with $ω$-Regular Objectives and Constraints
by: Wagner, Dominik, et al.
Published: (2025)
by: Wagner, Dominik, et al.
Published: (2025)
Multi-Objective Reinforcement Learning for Water Management
by: Osika, Zuzanna, et al.
Published: (2025)
by: Osika, Zuzanna, et al.
Published: (2025)
Non-Stationary Learning of Neural Networks with Automatic Soft Parameter Reset
by: Galashov, Alexandre, et al.
Published: (2024)
by: Galashov, Alexandre, et al.
Published: (2024)
MARLINE: Multi-Source Mapping Transfer Learning for Non-Stationary Environments
by: Du, Honghui, et al.
Published: (2025)
by: Du, Honghui, et al.
Published: (2025)
Active Preference Optimization for Sample Efficient RLHF
by: Das, Nirjhar, et al.
Published: (2024)
by: Das, Nirjhar, et al.
Published: (2024)
TREX: Trajectory Explanations for Multi-Objective Reinforcement Learning
by: Rajapakse, Dilina, et al.
Published: (2026)
by: Rajapakse, Dilina, et al.
Published: (2026)
Interpretability by Design for Efficient Multi-Objective Reinforcement Learning
by: Xia, Qiyue, et al.
Published: (2025)
by: Xia, Qiyue, et al.
Published: (2025)
Similar Items
-
Improved Training Mechanism for Reinforcement Learning via Online Model Selection
by: Afshar, Aida, et al.
Published: (2025) -
State-free Reinforcement Learning
by: Chen, Mingyu, et al.
Published: (2024) -
Bayesian Online Model Selection
by: Afshar, Aida, et al.
Published: (2026) -
Second Order Bounds for Contextual Bandits with Function Approximation
by: Pacchiano, Aldo
Published: (2024) -
Data-Driven Online Model Selection With Regret Guarantees
by: Pacchiano, Aldo, et al.
Published: (2023)