ADDQ: Adaptive Distributional Double Q-Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Döring, Leif, Wille, Benedikt, Birr, Maximilian, Bîrsan, Mihail, Slowik, Martin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Role of Target Update Frequencies in Q-Learning
by: Weissmann, Simon, et al.
Published: (2026)
by: Weissmann, Simon, et al.
Published: (2026)
Random Function Descent
by: Benning, Felix, et al.
Published: (2023)
by: Benning, Felix, et al.
Published: (2023)
Beyond Stationarity: Convergence Analysis of Stochastic Softmax Policy Gradient Methods
by: Klein, Sara, et al.
Published: (2023)
by: Klein, Sara, et al.
Published: (2023)
Controlling the Flow: Stability and Convergence for Stochastic Gradient Descent with Decaying Regularization
by: Kassing, Sebastian, et al.
Published: (2025)
by: Kassing, Sebastian, et al.
Published: (2025)
Almost sure convergence rates of stochastic gradient methods under gradient domination
by: Weissmann, Simon, et al.
Published: (2024)
by: Weissmann, Simon, et al.
Published: (2024)
An Approximate Ascent Approach To Prove Convergence of PPO
by: Doering, Leif, et al.
Published: (2026)
by: Doering, Leif, et al.
Published: (2026)
Structure Matters: Dynamic Policy Gradient
by: Klein, Sara, et al.
Published: (2024)
by: Klein, Sara, et al.
Published: (2024)
Wasserstein Distributionally Robust Regret Optimization
by: Fiechtner, Lukas-Benedikt, et al.
Published: (2025)
by: Fiechtner, Lukas-Benedikt, et al.
Published: (2025)
Gradient Span Algorithms Make Predictable Progress in High Dimension
by: Benning, Felix, et al.
Published: (2024)
by: Benning, Felix, et al.
Published: (2024)
Distributionally Robust Deep Q-Learning
by: Lu, Chung I, et al.
Published: (2025)
by: Lu, Chung I, et al.
Published: (2025)
Constant Stepsize Q-learning: Distributional Convergence, Bias and Extrapolation
by: Zhang, Yixuan, et al.
Published: (2024)
by: Zhang, Yixuan, et al.
Published: (2024)
Sample Complexity of Variance-reduced Distributionally Robust Q-learning
by: Wang, Shengbo, et al.
Published: (2023)
by: Wang, Shengbo, et al.
Published: (2023)
State Estimation Using Particle Filtering in Adaptive Machine Learning Methods: Integrating Q-Learning and NEAT Algorithms with Noisy Radar Measurements
by: Song, Wonjin, et al.
Published: (2025)
by: Song, Wonjin, et al.
Published: (2025)
Pointer Networks with Q-Learning for Combinatorial Optimization
by: Barro, Alessandro
Published: (2023)
by: Barro, Alessandro
Published: (2023)
Convergence of Distributed Adaptive Optimization with Local Updates
by: Cheng, Ziheng, et al.
Published: (2024)
by: Cheng, Ziheng, et al.
Published: (2024)
Robust Q-Learning under Corrupted Rewards
by: Maity, Sreejeet, et al.
Published: (2024)
by: Maity, Sreejeet, et al.
Published: (2024)
Central Limit Theorems for Asynchronous Averaged Q-Learning
by: Liu, Xingtu
Published: (2025)
by: Liu, Xingtu
Published: (2025)
Convergence and stability of Q-learning in Hierarchical Reinforcement Learning
by: Manenti, Massimiliano, et al.
Published: (2025)
by: Manenti, Massimiliano, et al.
Published: (2025)
The Sample-Communication Complexity Trade-off in Federated Q-Learning
by: Salgia, Sudeep, et al.
Published: (2024)
by: Salgia, Sudeep, et al.
Published: (2024)
Can Temporal-Difference and Q-Learning Learn Representation? A Mean-Field Theory
by: Zhang, Yufeng, et al.
Published: (2020)
by: Zhang, Yufeng, et al.
Published: (2020)
Corruption-Tolerant Asynchronous Q-Learning with Near-Optimal Rates
by: Maity, Sreejeet, et al.
Published: (2025)
by: Maity, Sreejeet, et al.
Published: (2025)
A Minimal-Assumption Analysis of Q-Learning with Time-Varying Policies
by: Nanda, Phalguni, et al.
Published: (2025)
by: Nanda, Phalguni, et al.
Published: (2025)
Q-Measure-Learning for Continuous State RL: Efficient Implementation and Convergence
by: Wang, Shengbo
Published: (2026)
by: Wang, Shengbo
Published: (2026)
Risk-Sensitive Q-Learning in Continuous Time with Application to Dynamic Portfolio Selection
by: Xie, Chuhan
Published: (2025)
by: Xie, Chuhan
Published: (2025)
Locally Adaptive Federated Learning
by: Mukherjee, Sohom, et al.
Published: (2023)
by: Mukherjee, Sohom, et al.
Published: (2023)
On Convergence of Average-Reward Q-Learning in Weakly Communicating Markov Decision Processes
by: Wan, Yi, et al.
Published: (2024)
by: Wan, Yi, et al.
Published: (2024)
Quantizer Design for Finite Model Approximations, Model Learning, and Quantized Q-Learning for MDPs with Unbounded Spaces
by: Bicer, Osman, et al.
Published: (2025)
by: Bicer, Osman, et al.
Published: (2025)
Double Duality: Variational Primal-Dual Policy Optimization for Constrained Reinforcement Learning
by: Li, Zihao, et al.
Published: (2024)
by: Li, Zihao, et al.
Published: (2024)
Communication-Efficient Adaptive Batch Size Strategies for Distributed Local Gradient Methods
by: Lau, Tim Tsz-Kit, et al.
Published: (2024)
by: Lau, Tim Tsz-Kit, et al.
Published: (2024)
Structured Difference-of-Q via Orthogonal Learning
by: Cao, Defu, et al.
Published: (2024)
by: Cao, Defu, et al.
Published: (2024)
Continuous Q-Score Matching: Diffusion Guided Reinforcement Learning for Continuous-Time Control
by: Hua, Chengxiu, et al.
Published: (2025)
by: Hua, Chengxiu, et al.
Published: (2025)
Faster Adaptive Decentralized Learning Algorithms
by: Huang, Feihu, et al.
Published: (2024)
by: Huang, Feihu, et al.
Published: (2024)
Distributionally-Robust Learning to Optimize
by: Ranjan, Vinit, et al.
Published: (2026)
by: Ranjan, Vinit, et al.
Published: (2026)
Adaptive Batch Size Schedules for Distributed Training of Language Models with Data and Model Parallelism
by: Lau, Tim Tsz-Kit, et al.
Published: (2024)
by: Lau, Tim Tsz-Kit, et al.
Published: (2024)
AdaBatchGrad: Combining Adaptive Batch Size and Adaptive Step Size
by: Ostroukhov, Petr, et al.
Published: (2024)
by: Ostroukhov, Petr, et al.
Published: (2024)
Distributionally Robust Learning in Survival Analysis
by: Jin, Yeping, et al.
Published: (2025)
by: Jin, Yeping, et al.
Published: (2025)
Wasserstein Distributionally Robust Online Learning
by: Chen, Guixian, et al.
Published: (2026)
by: Chen, Guixian, et al.
Published: (2026)
Foundations of Multivariate Distributional Reinforcement Learning
by: Wiltzer, Harley, et al.
Published: (2024)
by: Wiltzer, Harley, et al.
Published: (2024)
Communication-Efficient Stochastic Distributed Learning
by: Ren, Xiaoxing, et al.
Published: (2025)
by: Ren, Xiaoxing, et al.
Published: (2025)
On the Foundation of Distributionally Robust Reinforcement Learning
by: Wang, Shengbo, et al.
Published: (2023)
by: Wang, Shengbo, et al.
Published: (2023)
Similar Items
-
The Role of Target Update Frequencies in Q-Learning
by: Weissmann, Simon, et al.
Published: (2026) -
Random Function Descent
by: Benning, Felix, et al.
Published: (2023) -
Beyond Stationarity: Convergence Analysis of Stochastic Softmax Policy Gradient Methods
by: Klein, Sara, et al.
Published: (2023) -
Controlling the Flow: Stability and Convergence for Stochastic Gradient Descent with Decaying Regularization
by: Kassing, Sebastian, et al.
Published: (2025) -
Almost sure convergence rates of stochastic gradient methods under gradient domination
by: Weissmann, Simon, et al.
Published: (2024)