How Does Return Distribution in Distributional Reinforcement Learning Help Optimization?
Fuente:
arXiv
Saved in:
| Main Authors: | Sun, Ke, Jiang, Bei, Kong, Linglong |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Distributional Reinforcement Learning with Regularized Wasserstein Loss
by: Sun, Ke, et al.
Published: (2022)
by: Sun, Ke, et al.
Published: (2022)
Intrinsic Benefits of Categorical Distributional Loss: Uncertainty-aware Regularized Exploration in Reinforcement Learning
by: Sun, Ke, et al.
Published: (2021)
by: Sun, Ke, et al.
Published: (2021)
When and How Does In-Distribution Label Help Out-of-Distribution Detection?
by: Du, Xuefeng, et al.
Published: (2024)
by: Du, Xuefeng, et al.
Published: (2024)
Deep Fair Learning: A Unified Framework for Fine-tuning Representations with Sufficient Networks
by: Shi, Enze, et al.
Published: (2025)
by: Shi, Enze, et al.
Published: (2025)
How Does Unlabeled Data Provably Help Out-of-Distribution Detection?
by: Du, Xuefeng, et al.
Published: (2024)
by: Du, Xuefeng, et al.
Published: (2024)
A Distance-based Anomaly Detection Framework for Deep Reinforcement Learning
by: Zhang, Hongming, et al.
Published: (2021)
by: Zhang, Hongming, et al.
Published: (2021)
Oblivious subspace embeddings for compressed Tucker decompositions
by: Pietrosanu, Matthew, et al.
Published: (2024)
by: Pietrosanu, Matthew, et al.
Published: (2024)
How Does Distribution Matching Help Domain Generalization: An Information-theoretic Analysis
by: Dong, Yuxin, et al.
Published: (2024)
by: Dong, Yuxin, et al.
Published: (2024)
Non-Asymptotic Analysis of Online Local Private Learning with SGD
by: Shi, Enze, et al.
Published: (2025)
by: Shi, Enze, et al.
Published: (2025)
Principled Fast and Meta Knowledge Learners for Continual Reinforcement Learning
by: Sun, Ke, et al.
Published: (2026)
by: Sun, Ke, et al.
Published: (2026)
Understanding Fairness and Prediction Error through Subspace Decomposition and Influence Analysis
by: Shi, Enze, et al.
Published: (2025)
by: Shi, Enze, et al.
Published: (2025)
A Deep Bayesian Nonparametric Framework for Robust Mutual Information Estimation
by: Fazeliasl, Forough, et al.
Published: (2025)
by: Fazeliasl, Forough, et al.
Published: (2025)
Optimizing Return Distributions with Distributional Dynamic Programming
by: Pires, Bernardo Ávila, et al.
Published: (2025)
by: Pires, Bernardo Ávila, et al.
Published: (2025)
Differentially Private Conformal Prediction
by: Wu, Jiamei, et al.
Published: (2026)
by: Wu, Jiamei, et al.
Published: (2026)
How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models?
by: Cheng, Xiaoyuan, et al.
Published: (2026)
by: Cheng, Xiaoyuan, et al.
Published: (2026)
Imitation Learning as Return Distribution Matching
by: Lazzati, Filippo, et al.
Published: (2025)
by: Lazzati, Filippo, et al.
Published: (2025)
Distributional Reinforcement Learning with Diffusion Bridge Critics
by: Ding, Shutong, et al.
Published: (2026)
by: Ding, Shutong, et al.
Published: (2026)
Moments Matter:Stabilizing Policy Optimization using Return Distributions
by: Jabs, Dennis, et al.
Published: (2026)
by: Jabs, Dennis, et al.
Published: (2026)
Online federated learning framework for classification
by: Guo, Wenxing, et al.
Published: (2025)
by: Guo, Wenxing, et al.
Published: (2025)
In-Dataset Trajectory Return Regularization for Offline Preference-based Reinforcement Learning
by: Tu, Songjun, et al.
Published: (2024)
by: Tu, Songjun, et al.
Published: (2024)
Goal Exploration via Adaptive Skill Distribution for Goal-Conditioned Reinforcement Learning
by: Wu, Lisheng, et al.
Published: (2024)
by: Wu, Lisheng, et al.
Published: (2024)
Distributed Direct Preference Optimization
by: Jiang, Zhanhong
Published: (2026)
by: Jiang, Zhanhong
Published: (2026)
Distributional Inverse Reinforcement Learning
by: Wu, Feiyang, et al.
Published: (2025)
by: Wu, Feiyang, et al.
Published: (2025)
Generalized Distribution Prediction for Asset Returns
by: Pétursson, Ísak, et al.
Published: (2024)
by: Pétursson, Ísak, et al.
Published: (2024)
Reinforcement Learning for Efficient Returns Management
by: Linden, Pascal, et al.
Published: (2025)
by: Linden, Pascal, et al.
Published: (2025)
Flow-based Policy With Distributional Reinforcement Learning in Trajectory Optimization
by: Hao, Ruijie, et al.
Published: (2026)
by: Hao, Ruijie, et al.
Published: (2026)
Prediction-powered Inference by Mixture of Experts
by: Gu, Yanwu, et al.
Published: (2026)
by: Gu, Yanwu, et al.
Published: (2026)
Estimation and Inference in Distributional Reinforcement Learning
by: Zhang, Liangyu, et al.
Published: (2023)
by: Zhang, Liangyu, et al.
Published: (2023)
Fourier Head: Helping Large Language Models Learn Complex Probability Distributions
by: Gillman, Nate, et al.
Published: (2024)
by: Gillman, Nate, et al.
Published: (2024)
More Benefits of Being Distributional: Second-Order Bounds for Reinforcement Learning
by: Wang, Kaiwen, et al.
Published: (2024)
by: Wang, Kaiwen, et al.
Published: (2024)
High-Performance Reinforcement Learning on Spot: Optimizing Simulation Parameters with Distributional Measures
by: Miller, AJ, et al.
Published: (2025)
by: Miller, AJ, et al.
Published: (2025)
Distributional Reinforcement Learning via the Cramér Distance
by: Aziz, Vanya, et al.
Published: (2026)
by: Aziz, Vanya, et al.
Published: (2026)
Assessing Uncertainty in Stock Returns: A Gaussian Mixture Distribution-Based Method
by: Wang, Yanlong, et al.
Published: (2025)
by: Wang, Yanlong, et al.
Published: (2025)
On the Distributed Evaluation of Generative Models
by: Wang, Zixiao, et al.
Published: (2023)
by: Wang, Zixiao, et al.
Published: (2023)
Does Homophily Help in Robust Test-time Node Classification?
by: Jiang, Yan, et al.
Published: (2025)
by: Jiang, Yan, et al.
Published: (2025)
How Does the Pretraining Distribution Shape In-Context Learning? Task Selection, Generalization, and Robustness
by: Azizian, Waïss, et al.
Published: (2025)
by: Azizian, Waïss, et al.
Published: (2025)
Vulnerability Analysis of Safe Reinforcement Learning via Inverse Constrained Reinforcement Learning
by: Fan, Jialiang, et al.
Published: (2026)
by: Fan, Jialiang, et al.
Published: (2026)
Foundations of Multivariate Distributional Reinforcement Learning
by: Wiltzer, Harley, et al.
Published: (2024)
by: Wiltzer, Harley, et al.
Published: (2024)
Distorted Distributional Policy Evaluation for Offline Reinforcement Learning
by: Iwaki, Ryo, et al.
Published: (2026)
by: Iwaki, Ryo, et al.
Published: (2026)
Convergence Theorems for Entropy-Regularized and Distributional Reinforcement Learning
by: Jhaveri, Yash, et al.
Published: (2025)
by: Jhaveri, Yash, et al.
Published: (2025)
Similar Items
-
Distributional Reinforcement Learning with Regularized Wasserstein Loss
by: Sun, Ke, et al.
Published: (2022) -
Intrinsic Benefits of Categorical Distributional Loss: Uncertainty-aware Regularized Exploration in Reinforcement Learning
by: Sun, Ke, et al.
Published: (2021) -
When and How Does In-Distribution Label Help Out-of-Distribution Detection?
by: Du, Xuefeng, et al.
Published: (2024) -
Deep Fair Learning: A Unified Framework for Fine-tuning Representations with Sufficient Networks
by: Shi, Enze, et al.
Published: (2025) -
How Does Unlabeled Data Provably Help Out-of-Distribution Detection?
by: Du, Xuefeng, et al.
Published: (2024)