Convergence Theorems for Entropy-Regularized and Distributional Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jhaveri, Yash, Wiltzer, Harley, Shafto, Patrick, Bellemare, Marc G., Meger, David |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Action Gaps and Advantages in Continuous-Time Distributional Reinforcement Learning
von: Wiltzer, Harley, et al.
Veröffentlicht: (2024)
von: Wiltzer, Harley, et al.
Veröffentlicht: (2024)
Tractable Representations for Convergent Approximation of Distributional HJB Equations
von: Alhosh, Julie, et al.
Veröffentlicht: (2025)
von: Alhosh, Julie, et al.
Veröffentlicht: (2025)
Policy Optimization in a Noisy Neighborhood: On Return Landscapes in Continuous Control
von: Rahn, Nate, et al.
Veröffentlicht: (2023)
von: Rahn, Nate, et al.
Veröffentlicht: (2023)
Foundations of Multivariate Distributional Reinforcement Learning
von: Wiltzer, Harley, et al.
Veröffentlicht: (2024)
von: Wiltzer, Harley, et al.
Veröffentlicht: (2024)
A Distributional Analogue to the Successor Representation
von: Wiltzer, Harley, et al.
Veröffentlicht: (2024)
von: Wiltzer, Harley, et al.
Veröffentlicht: (2024)
Parseval Regularization for Continual Reinforcement Learning
von: Chung, Wesley, et al.
Veröffentlicht: (2024)
von: Chung, Wesley, et al.
Veröffentlicht: (2024)
Non-Adversarial Inverse Reinforcement Learning via Successor Feature Matching
von: Jain, Arnav Kumar, et al.
Veröffentlicht: (2024)
von: Jain, Arnav Kumar, et al.
Veröffentlicht: (2024)
Multi-Agent Model-Based Reinforcement Learning with Joint State-Action Learned Embeddings
von: Wang, Zhizun, et al.
Veröffentlicht: (2026)
von: Wang, Zhizun, et al.
Veröffentlicht: (2026)
Global Convergence of Wasserstein Policy Gradient for Entropy-Regularized Reinforcement Learning
von: Zhu, Zhaoyu, et al.
Veröffentlicht: (2026)
von: Zhu, Zhaoyu, et al.
Veröffentlicht: (2026)
KerJEPA: Kernel Discrepancies for Euclidean Self-Supervised Learning
von: Zimmermann, Eric, et al.
Veröffentlicht: (2025)
von: Zimmermann, Eric, et al.
Veröffentlicht: (2025)
Multi-scale Predictive Representations for Goal-conditioned Reinforcement Learning
von: Adaikkappan, Valliappan Chidambaram, et al.
Veröffentlicht: (2026)
von: Adaikkappan, Valliappan Chidambaram, et al.
Veröffentlicht: (2026)
State Entropy Regularization for Robust Reinforcement Learning
von: Ashlag, Yonatan, et al.
Veröffentlicht: (2025)
von: Ashlag, Yonatan, et al.
Veröffentlicht: (2025)
Fairness in Reinforcement Learning with Bisimulation Metrics
von: Rezaei-Shoshtari, Sahand, et al.
Veröffentlicht: (2024)
von: Rezaei-Shoshtari, Sahand, et al.
Veröffentlicht: (2024)
Entropy Regularized Task Representation Learning for Offline Meta-Reinforcement Learning
von: Nakhaei, Mohammadreza, et al.
Veröffentlicht: (2024)
von: Nakhaei, Mohammadreza, et al.
Veröffentlicht: (2024)
Efficient Epistemic Uncertainty Estimation in Regression Ensemble Models Using Pairwise-Distance Estimators
von: Berry, Lucas, et al.
Veröffentlicht: (2023)
von: Berry, Lucas, et al.
Veröffentlicht: (2023)
Entropy-Regularized Adjoint Matching for Offline Reinforcement Learning
von: Ghanem, Abdelghani, et al.
Veröffentlicht: (2026)
von: Ghanem, Abdelghani, et al.
Veröffentlicht: (2026)
Distributional Reinforcement Learning with Regularized Wasserstein Loss
von: Sun, Ke, et al.
Veröffentlicht: (2022)
von: Sun, Ke, et al.
Veröffentlicht: (2022)
Structured Evaluation of Synthetic Tabular Data
von: Yang, Scott Cheng-Hsin, et al.
Veröffentlicht: (2024)
von: Yang, Scott Cheng-Hsin, et al.
Veröffentlicht: (2024)
On the geometry and topology of representations: the manifolds of modular addition
von: Moisescu-Pareja, Gabriela, et al.
Veröffentlicht: (2025)
von: Moisescu-Pareja, Gabriela, et al.
Veröffentlicht: (2025)
Predicting The Cop Number Using Machine Learning
von: Mann, Meagan, et al.
Veröffentlicht: (2026)
von: Mann, Meagan, et al.
Veröffentlicht: (2026)
Federated Distributional Reinforcement Learning with Distributional Critic Regularization
von: Millard, David, et al.
Veröffentlicht: (2026)
von: Millard, David, et al.
Veröffentlicht: (2026)
VDFD: Multi-Agent Value Decomposition Framework with Disentangled World Model
von: Wang, Zhizun, et al.
Veröffentlicht: (2023)
von: Wang, Zhizun, et al.
Veröffentlicht: (2023)
DROP: Distributional and Regular Optimism and Pessimism for Reinforcement Learning
von: Kobayashi, Taisuke
Veröffentlicht: (2024)
von: Kobayashi, Taisuke
Veröffentlicht: (2024)
Viability of Future Actions: Robust Safety in Reinforcement Learning via Entropy Regularization
von: Massiani, Pierre-François, et al.
Veröffentlicht: (2025)
von: Massiani, Pierre-François, et al.
Veröffentlicht: (2025)
Provably Safe Reinforcement Learning for Stochastic Reach-Avoid Problems with Entropy Regularization
von: Mazumdar, Abhijit, et al.
Veröffentlicht: (2026)
von: Mazumdar, Abhijit, et al.
Veröffentlicht: (2026)
Addressing Label Shift in Distributed Learning via Entropy Regularization
von: Wu, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Wu, Zhiyuan, et al.
Veröffentlicht: (2025)
Federated Learning Framework via Distributed Mutual Learning
von: Gupta, Yash
Veröffentlicht: (2025)
von: Gupta, Yash
Veröffentlicht: (2025)
Quantile Geometry Regularization for Distributional Reinforcement Learning
von: Zhang, Zhaofan, et al.
Veröffentlicht: (2026)
von: Zhang, Zhaofan, et al.
Veröffentlicht: (2026)
Relative Entropy Regularized Reinforcement Learning for Efficient Encrypted Policy Synthesis
von: Suh, Jihoon, et al.
Veröffentlicht: (2025)
von: Suh, Jihoon, et al.
Veröffentlicht: (2025)
SPINE: Token-Selective Test-Time Reinforcement Learning with Entropy-Band Regularization
von: Wu, Jianghao, et al.
Veröffentlicht: (2025)
von: Wu, Jianghao, et al.
Veröffentlicht: (2025)
Gradient Flow Convergence Guarantee for General Neural Network Architectures
von: Jakhmola, Yash
Veröffentlicht: (2025)
von: Jakhmola, Yash
Veröffentlicht: (2025)
Beyond Exact Gradients: Convergence of Stochastic Soft-Max Policy Gradient Methods with Entropy Regularization
von: Ding, Yuhao, et al.
Veröffentlicht: (2021)
von: Ding, Yuhao, et al.
Veröffentlicht: (2021)
Linear Convergence of Entropy-Regularized Natural Policy Gradient with Linear Function Approximation
von: Cayci, Semih, et al.
Veröffentlicht: (2021)
von: Cayci, Semih, et al.
Veröffentlicht: (2021)
An Analysis of Quantile Temporal-Difference Learning
von: Rowland, Mark, et al.
Veröffentlicht: (2023)
von: Rowland, Mark, et al.
Veröffentlicht: (2023)
Linear Convergence of Independent Natural Policy Gradient in Games with Entropy Regularization
von: Sun, Youbang, et al.
Veröffentlicht: (2024)
von: Sun, Youbang, et al.
Veröffentlicht: (2024)
Beyond Softmax and Entropy: Convergence Rates of Policy Gradients with f-SoftArgmax Parameterization & Coupled Regularization
von: Labbi, Safwan, et al.
Veröffentlicht: (2026)
von: Labbi, Safwan, et al.
Veröffentlicht: (2026)
Imitation Learning from Observation through Optimal Transport
von: Chang, Wei-Di, et al.
Veröffentlicht: (2023)
von: Chang, Wei-Di, et al.
Veröffentlicht: (2023)
Compositional Planning with Jumpy World Models
von: Farebrother, Jesse, et al.
Veröffentlicht: (2026)
von: Farebrother, Jesse, et al.
Veröffentlicht: (2026)
Intrinsic Benefits of Categorical Distributional Loss: Uncertainty-aware Regularized Exploration in Reinforcement Learning
von: Sun, Ke, et al.
Veröffentlicht: (2021)
von: Sun, Ke, et al.
Veröffentlicht: (2021)
Cyclical Entropy Eruption: Entropy Dynamics in Agent Reinforcement Learning
von: Li, Wendi, et al.
Veröffentlicht: (2026)
von: Li, Wendi, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Action Gaps and Advantages in Continuous-Time Distributional Reinforcement Learning
von: Wiltzer, Harley, et al.
Veröffentlicht: (2024) -
Tractable Representations for Convergent Approximation of Distributional HJB Equations
von: Alhosh, Julie, et al.
Veröffentlicht: (2025) -
Policy Optimization in a Noisy Neighborhood: On Return Landscapes in Continuous Control
von: Rahn, Nate, et al.
Veröffentlicht: (2023) -
Foundations of Multivariate Distributional Reinforcement Learning
von: Wiltzer, Harley, et al.
Veröffentlicht: (2024) -
A Distributional Analogue to the Successor Representation
von: Wiltzer, Harley, et al.
Veröffentlicht: (2024)