Dynamic Learning Rate for Deep Reinforcement Learning: A Bandit Approach
Fuente:
arXiv
Saved in:
| Main Authors: | Donâncio, Henrique, Barrier, Antoine, South, Leah F., Forbes, Florence |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Pump Scheduling Problem: A Real-World Scenario for Reinforcement Learning
by: Donâncio, Henrique, et al.
Published: (2022)
by: Donâncio, Henrique, et al.
Published: (2022)
Learning Rate Optimization for Deep Neural Networks Using Lipschitz Bandits
by: Priyanka, Padma, et al.
Published: (2024)
by: Priyanka, Padma, et al.
Published: (2024)
Adaptive Action Duration with Contextual Bandits for Deep Reinforcement Learning in Dynamic Environments
by: Verma, Abhishek, et al.
Published: (2025)
by: Verma, Abhishek, et al.
Published: (2025)
Uniform Last-Iterate Guarantee for Bandits and Reinforcement Learning
by: Liu, Junyan, et al.
Published: (2024)
by: Liu, Junyan, et al.
Published: (2024)
Restless Bandits with Individual Penalty Constraints: Near-Optimal Indices and Deep Reinforcement Learning
by: Zamir, Nida, et al.
Published: (2026)
by: Zamir, Nida, et al.
Published: (2026)
PSAT: Pediatric Segmentation Approaches via Adult Augmentations and Transfer Learning
by: Kirscher, Tristan, et al.
Published: (2025)
by: Kirscher, Tristan, et al.
Published: (2025)
The Polynomial Stein Discrepancy for Assessing Moment Convergence
by: Srinivasan, Narayan, et al.
Published: (2024)
by: Srinivasan, Narayan, et al.
Published: (2024)
LLMs Are In-Context Bandit Reinforcement Learners
by: Monea, Giovanni, et al.
Published: (2024)
by: Monea, Giovanni, et al.
Published: (2024)
Deep Reinforcement Learning based Triggering Function for Early Classifiers of Time Series
by: Renault, Aurélien, et al.
Published: (2025)
by: Renault, Aurélien, et al.
Published: (2025)
Multi-Objective Adaptive Rate Limiting in Microservices Using Deep Reinforcement Learning
by: Lyu, Ning, et al.
Published: (2025)
by: Lyu, Ning, et al.
Published: (2025)
Deep Reinforcement Learning: A Convex Optimization Approach
by: Gattami, Ather
Published: (2024)
by: Gattami, Ather
Published: (2024)
Networked Restless Multi-Arm Bandits with Reinforcement Learning
by: Zhang, Hanmo, et al.
Published: (2025)
by: Zhang, Hanmo, et al.
Published: (2025)
Contextual Rollout Bandits for Reinforcement Learning with Verifiable Rewards
by: Lu, Xiaodong, et al.
Published: (2026)
by: Lu, Xiaodong, et al.
Published: (2026)
Pure Exploration for a Good Policy in Reinforcement Learning with Bandit Feedback
by: Li, Zitian, et al.
Published: (2026)
by: Li, Zitian, et al.
Published: (2026)
Unified Framework of Distributional Regret in Multi-Armed Bandits and Reinforcement Learning
by: Lee, Harin, et al.
Published: (2026)
by: Lee, Harin, et al.
Published: (2026)
Learning for Bandits under Action Erasures
by: Hanna, Osama, et al.
Published: (2024)
by: Hanna, Osama, et al.
Published: (2024)
On the Hardness of Bandit Learning
by: Brukhim, Nataly, et al.
Published: (2025)
by: Brukhim, Nataly, et al.
Published: (2025)
Symmetry-Preserving Architecture for Multi-NUMA Environments (SPANE): A Deep Reinforcement Learning Approach for Dynamic VM Scheduling
by: Chan, Tin Ping, et al.
Published: (2025)
by: Chan, Tin Ping, et al.
Published: (2025)
Learning to Attack: A Bandit Approach to Adversarial Context Poisoning
by: Telikani, Ray, et al.
Published: (2026)
by: Telikani, Ray, et al.
Published: (2026)
Online Learning to Rank under Corruption: A Robust Cascading Bandits Approach
by: Ghaffari, Fatemeh, et al.
Published: (2025)
by: Ghaffari, Fatemeh, et al.
Published: (2025)
Combinatorial Multivariant Multi-Armed Bandits with Applications to Episodic Reinforcement Learning and Beyond
by: Liu, Xutong, et al.
Published: (2024)
by: Liu, Xutong, et al.
Published: (2024)
Reinforcement Learning for Machine Learning Model Deployment: Evaluating Multi-Armed Bandits in ML Ops Environments
by: McClendon, S. Aaron, et al.
Published: (2025)
by: McClendon, S. Aaron, et al.
Published: (2025)
Provably Efficient Reinforcement Learning for Adversarial Restless Multi-Armed Bandits with Unknown Transitions and Bandit Feedback
by: Xiong, Guojun, et al.
Published: (2024)
by: Xiong, Guojun, et al.
Published: (2024)
Few-Shot Learning for Dynamic Operations of Automated Electric Taxi Fleets under Evolving Charging Infrastructure: A Meta-Deep Reinforcement Learning Approach
by: Li, Xiaozhuang, et al.
Published: (2026)
by: Li, Xiaozhuang, et al.
Published: (2026)
Faster Rates for Private Adversarial Bandits
by: Asi, Hilal, et al.
Published: (2025)
by: Asi, Hilal, et al.
Published: (2025)
Enhancing Courier Scheduling in Crowdsourced Last-Mile Delivery through Dynamic Shift Extensions: A Deep Reinforcement Learning Approach
by: Saleh, Zead, et al.
Published: (2024)
by: Saleh, Zead, et al.
Published: (2024)
Rethinking the Role of Dynamic Sparse Training for Scalable Deep Reinforcement Learning
by: Ma, Guozheng, et al.
Published: (2025)
by: Ma, Guozheng, et al.
Published: (2025)
Deep Proxy Causal Learning and its Application to Confounded Bandit Policy Evaluation
by: Xu, Liyuan, et al.
Published: (2021)
by: Xu, Liyuan, et al.
Published: (2021)
A Conservative Approach for Few-Shot Transfer in Off-Dynamics Reinforcement Learning
by: Daoudi, Paul, et al.
Published: (2023)
by: Daoudi, Paul, et al.
Published: (2023)
Rating-based Reinforcement Learning
by: White, Devin, et al.
Published: (2023)
by: White, Devin, et al.
Published: (2023)
Enhancing Robustness in Deep Reinforcement Learning: A Lyapunov Exponent Approach
by: Young, Rory, et al.
Published: (2024)
by: Young, Rory, et al.
Published: (2024)
Convergence of projected stochastic natural gradient variational inference for various step size and sample or batch size schedules
by: Guilmeau, Thomas, et al.
Published: (2026)
by: Guilmeau, Thomas, et al.
Published: (2026)
Optimal Control of Fluid Restless Multi-armed Bandits: A Machine Learning Approach
by: Bertsimas, Dimitris, et al.
Published: (2025)
by: Bertsimas, Dimitris, et al.
Published: (2025)
Bayesian Experimental Design via Contrastive Diffusions
by: Iollo, Jacopo, et al.
Published: (2024)
by: Iollo, Jacopo, et al.
Published: (2024)
Predictable Reinforcement Learning Dynamics through Entropy Rate Minimization
by: Ornia, Daniel Jarne, et al.
Published: (2023)
by: Ornia, Daniel Jarne, et al.
Published: (2023)
Solving The Dynamic Volatility Fitting Problem: A Deep Reinforcement Learning Approach
by: Gnabeyeu, Emmanuel, et al.
Published: (2024)
by: Gnabeyeu, Emmanuel, et al.
Published: (2024)
The Bandit Whisperer: Communication Learning for Restless Bandits
by: Zhao, Yunfan, et al.
Published: (2024)
by: Zhao, Yunfan, et al.
Published: (2024)
Unsupervised Representation Learning in Deep Reinforcement Learning: A Review
by: Botteghi, Nicolò, et al.
Published: (2022)
by: Botteghi, Nicolò, et al.
Published: (2022)
Advancing Robustness in Deep Reinforcement Learning with an Ensemble Defense Approach
by: Mohan, Adithya, et al.
Published: (2025)
by: Mohan, Adithya, et al.
Published: (2025)
Fast Rates for Inverse Reinforcement Learning
by: Schlaginhaufen, Andreas, et al.
Published: (2026)
by: Schlaginhaufen, Andreas, et al.
Published: (2026)
Similar Items
-
The Pump Scheduling Problem: A Real-World Scenario for Reinforcement Learning
by: Donâncio, Henrique, et al.
Published: (2022) -
Learning Rate Optimization for Deep Neural Networks Using Lipschitz Bandits
by: Priyanka, Padma, et al.
Published: (2024) -
Adaptive Action Duration with Contextual Bandits for Deep Reinforcement Learning in Dynamic Environments
by: Verma, Abhishek, et al.
Published: (2025) -
Uniform Last-Iterate Guarantee for Bandits and Reinforcement Learning
by: Liu, Junyan, et al.
Published: (2024) -
Restless Bandits with Individual Penalty Constraints: Near-Optimal Indices and Deep Reinforcement Learning
by: Zamir, Nida, et al.
Published: (2026)