Optimizing Return Distributions with Distributional Dynamic Programming
Fuente:
arXiv
Saved in:
| Main Authors: | Pires, Bernardo Ávila, Rowland, Mark, Borsa, Diana, Guo, Zhaohan Daniel, Khetarpal, Khimya, Barreto, André, Abel, David, Munos, Rémi, Dabney, Will |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Off-policy Distributional Q($λ$): Distributional RL without Importance Sampling
by: Tang, Yunhao, et al.
Published: (2024)
by: Tang, Yunhao, et al.
Published: (2024)
A Unifying Framework for Action-Conditional Self-Predictive Reinforcement Learning
by: Khetarpal, Khimya, et al.
Published: (2024)
by: Khetarpal, Khimya, et al.
Published: (2024)
Representation Learning via Non-Contrastive Mutual Information
by: Guo, Zhaohan Daniel, et al.
Published: (2025)
by: Guo, Zhaohan Daniel, et al.
Published: (2025)
Near-Minimax-Optimal Distributional Reinforcement Learning with a Generative Model
by: Rowland, Mark, et al.
Published: (2024)
by: Rowland, Mark, et al.
Published: (2024)
Generalized Preference Optimization: A Unified Approach to Offline Alignment
by: Tang, Yunhao, et al.
Published: (2024)
by: Tang, Yunhao, et al.
Published: (2024)
Understanding the performance gap between online and offline alignment algorithms
by: Tang, Yunhao, et al.
Published: (2024)
by: Tang, Yunhao, et al.
Published: (2024)
Foundations of Multivariate Distributional Reinforcement Learning
by: Wiltzer, Harley, et al.
Published: (2024)
by: Wiltzer, Harley, et al.
Published: (2024)
Preventing Learning Stagnation in PPO by Scaling to 1 Million Parallel Environments
by: Beukman, Michael, et al.
Published: (2026)
by: Beukman, Michael, et al.
Published: (2026)
A Distributional Analogue to the Successor Representation
by: Wiltzer, Harley, et al.
Published: (2024)
by: Wiltzer, Harley, et al.
Published: (2024)
Robust Intervention Learning from Emergency Stop Interventions
by: Pronovost, Ethan, et al.
Published: (2026)
by: Pronovost, Ethan, et al.
Published: (2026)
Disentangling the Causes of Plasticity Loss in Neural Networks
by: Lyle, Clare, et al.
Published: (2024)
by: Lyle, Clare, et al.
Published: (2024)
Normalization and effective learning rates in reinforcement learning
by: Lyle, Clare, et al.
Published: (2024)
by: Lyle, Clare, et al.
Published: (2024)
Plasticity as the Mirror of Empowerment
by: Abel, David, et al.
Published: (2025)
by: Abel, David, et al.
Published: (2025)
Agency Is Frame-Dependent
by: Abel, David, et al.
Published: (2025)
by: Abel, David, et al.
Published: (2025)
A Fault-Tolerant Distributed Termination Method for Distributed Optimization Algorithms
by: Alkhraijah, Mohannad, et al.
Published: (2024)
by: Alkhraijah, Mohannad, et al.
Published: (2024)
VA-learning as a more efficient alternative to Q-learning
by: Tang, Yunhao, et al.
Published: (2023)
by: Tang, Yunhao, et al.
Published: (2023)
On Degeneracy Issues in Multi-parametric Programming and Critical Region Exploration based Distributed Optimization in Smart Grid Operations
by: Liu, Haitian, et al.
Published: (2023)
by: Liu, Haitian, et al.
Published: (2023)
Dynamics and Optimization in Spatially Distributed Electrical Vehicle Charging
by: Paganini, Fernando, et al.
Published: (2024)
by: Paganini, Fernando, et al.
Published: (2024)
Sensitivity-Based Distributed Programming for Non-Convex Optimization
by: von Esch, Maximilian Pierer, et al.
Published: (2025)
by: von Esch, Maximilian Pierer, et al.
Published: (2025)
Discretized Distributed Optimization over Dynamic Digraphs
by: Doostmohammadian, Mohammadreza, et al.
Published: (2023)
by: Doostmohammadian, Mohammadreza, et al.
Published: (2023)
Decision-Dependent Stochastic Optimization: The Role of Distribution Dynamics
by: He, Zhiyu, et al.
Published: (2025)
by: He, Zhiyu, et al.
Published: (2025)
On the Detection of Shared Data Manipulation in Distributed Optimization
by: Alkhraijah, Mohannad, et al.
Published: (2023)
by: Alkhraijah, Mohannad, et al.
Published: (2023)
Distributed Least-Squares Optimization Solvers with Differential Privacy
by: Liu, Weijia, et al.
Published: (2024)
by: Liu, Weijia, et al.
Published: (2024)
Distributionally Robust Optimization
by: Kuhn, Daniel, et al.
Published: (2024)
by: Kuhn, Daniel, et al.
Published: (2024)
Optimization-Based Control of Distributed Battery Storage in Distribution Networks
by: de Carvalho, Wilhiam, et al.
Published: (2024)
by: de Carvalho, Wilhiam, et al.
Published: (2024)
Delay-Robust Primal-Dual Dynamics for Distributed Optimization
by: Şen, Gökçen Devlet, et al.
Published: (2026)
by: Şen, Gökçen Devlet, et al.
Published: (2026)
Optimization with Multi-sourced Information and Unknown Reliability: A Distributionally Robust Approach
by: Guo, Yanru, et al.
Published: (2025)
by: Guo, Yanru, et al.
Published: (2025)
Dynamic Control of Service Systems with Returns: Application to Design of Post-Discharge Hospital Readmission Prevention Programs
by: Chan, Timothy C. Y., et al.
Published: (2022)
by: Chan, Timothy C. Y., et al.
Published: (2022)
Distributed Optimization Algorithm with Superlinear Convergence Rate
by: Xu, Yeming, et al.
Published: (2024)
by: Xu, Yeming, et al.
Published: (2024)
Self-Predictive Representations for Combinatorial Generalization in Behavioral Cloning
by: Lawson, Daniel, et al.
Published: (2025)
by: Lawson, Daniel, et al.
Published: (2025)
Distributed Optimization Method Based On Optimal Control
by: Guo, Ziyuan, et al.
Published: (2024)
by: Guo, Ziyuan, et al.
Published: (2024)
An Analysis of Quantile Temporal-Difference Learning
by: Rowland, Mark, et al.
Published: (2023)
by: Rowland, Mark, et al.
Published: (2023)
On Tractability, Complexity, and Mixed-Integer Convex Programming Representability of Distributionally Favorable Optimization
by: Jiang, Nan, et al.
Published: (2024)
by: Jiang, Nan, et al.
Published: (2024)
Deep Distributed Optimization for Large-Scale Quadratic Programming
by: Saravanos, Augustinos D., et al.
Published: (2024)
by: Saravanos, Augustinos D., et al.
Published: (2024)
Robustness Measures in Distributionally Robust Optimization
by: Gotoh, Jun-ya, et al.
Published: (2025)
by: Gotoh, Jun-ya, et al.
Published: (2025)
Control-Based Online Distributed Optimization
by: van Weerelt, Wouter J. A., et al.
Published: (2025)
by: van Weerelt, Wouter J. A., et al.
Published: (2025)
Decision-Dependent Distributionally Robust Optimization with Application to Dynamic Pricing
by: Qu, Chengrui, et al.
Published: (2025)
by: Qu, Chengrui, et al.
Published: (2025)
Quadratic Truncated Random Return in Distributional LQR: Positive Definiteness, Density, and Log-Concavity
by: Teng, Ruyi, et al.
Published: (2025)
by: Teng, Ruyi, et al.
Published: (2025)
Affine-coupled Distributed Optimization via Distributed Proximal Jacobian ADMM with Quantized Communication
by: Du, Xu, et al.
Published: (2026)
by: Du, Xu, et al.
Published: (2026)
Logarithmically Quantized Distributed Optimization over Dynamic Multi-Agent Networks
by: Doostmohammadian, Mohammadreza, et al.
Published: (2024)
by: Doostmohammadian, Mohammadreza, et al.
Published: (2024)
Similar Items
-
Off-policy Distributional Q($λ$): Distributional RL without Importance Sampling
by: Tang, Yunhao, et al.
Published: (2024) -
A Unifying Framework for Action-Conditional Self-Predictive Reinforcement Learning
by: Khetarpal, Khimya, et al.
Published: (2024) -
Representation Learning via Non-Contrastive Mutual Information
by: Guo, Zhaohan Daniel, et al.
Published: (2025) -
Near-Minimax-Optimal Distributional Reinforcement Learning with a Generative Model
by: Rowland, Mark, et al.
Published: (2024) -
Generalized Preference Optimization: A Unified Approach to Offline Alignment
by: Tang, Yunhao, et al.
Published: (2024)