Intrinsic Benefits of Categorical Distributional Loss: Uncertainty-aware Regularized Exploration in Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Sun, Ke, Zhao, Yingnan, Shi, Enze, Wang, Yafei, Yan, Xiaodong, Jiang, Bei, Kong, Linglong |
|---|---|
| Format: | Preprint |
| Published: |
2021
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Distributional Reinforcement Learning with Regularized Wasserstein Loss
by: Sun, Ke, et al.
Published: (2022)
by: Sun, Ke, et al.
Published: (2022)
How Does Return Distribution in Distributional Reinforcement Learning Help Optimization?
by: Sun, Ke, et al.
Published: (2022)
by: Sun, Ke, et al.
Published: (2022)
Deep Fair Learning: A Unified Framework for Fine-tuning Representations with Sufficient Networks
by: Shi, Enze, et al.
Published: (2025)
by: Shi, Enze, et al.
Published: (2025)
Non-Asymptotic Analysis of Online Local Private Learning with SGD
by: Shi, Enze, et al.
Published: (2025)
by: Shi, Enze, et al.
Published: (2025)
Understanding Fairness and Prediction Error through Subspace Decomposition and Influence Analysis
by: Shi, Enze, et al.
Published: (2025)
by: Shi, Enze, et al.
Published: (2025)
Online differentially private inference in stochastic gradient descent
by: Xie, Jinhan, et al.
Published: (2025)
by: Xie, Jinhan, et al.
Published: (2025)
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon
by: Pan, Bo, et al.
Published: (2025)
by: Pan, Bo, et al.
Published: (2025)
CBMA: Improving conformal prediction through Bayesian model averaging
by: Bhagwat, Pankaj, et al.
Published: (2025)
by: Bhagwat, Pankaj, et al.
Published: (2025)
Oblivious subspace embeddings for compressed Tucker decompositions
by: Pietrosanu, Matthew, et al.
Published: (2024)
by: Pietrosanu, Matthew, et al.
Published: (2024)
Fast Online $L_0$ Elastic Net Subspace Clustering via A Novel Dictionary Update Strategy
by: Qu, Wentao, et al.
Published: (2024)
by: Qu, Wentao, et al.
Published: (2024)
A Distance-based Anomaly Detection Framework for Deep Reinforcement Learning
by: Zhang, Hongming, et al.
Published: (2021)
by: Zhang, Hongming, et al.
Published: (2021)
ARMA-Design: Optimal Treatment Allocation Strategies for A/B Testing in Partially Observable Time Series Experiments
by: Sun, Ke, et al.
Published: (2024)
by: Sun, Ke, et al.
Published: (2024)
Conformal Inference For Missing Data under Multiple Robust Learning
by: Tang, Wenlu, et al.
Published: (2025)
by: Tang, Wenlu, et al.
Published: (2025)
Efficient Reinforcement Learning for Large Language Models with Intrinsic Exploration
by: Sun, Yan, et al.
Published: (2025)
by: Sun, Yan, et al.
Published: (2025)
A Deep Bayesian Nonparametric Framework for Robust Mutual Information Estimation
by: Fazeliasl, Forough, et al.
Published: (2025)
by: Fazeliasl, Forough, et al.
Published: (2025)
Variational Dynamic for Self-Supervised Exploration in Deep Reinforcement Learning
by: Bai, Chenjia, et al.
Published: (2020)
by: Bai, Chenjia, et al.
Published: (2020)
Principled Fast and Meta Knowledge Learners for Continual Reinforcement Learning
by: Sun, Ke, et al.
Published: (2026)
by: Sun, Ke, et al.
Published: (2026)
RN-D: Discretized Categorical Actors with Regularized Networks for On-Policy Reinforcement Learning
by: Bian, Yuexin, et al.
Published: (2026)
by: Bian, Yuexin, et al.
Published: (2026)
The Benefits of Power Regularization in Cooperative Reinforcement Learning
by: Li, Michelle, et al.
Published: (2024)
by: Li, Michelle, et al.
Published: (2024)
Texture-aware Intrinsic Image Decomposition with Model- and Learning-based Priors
by: Wang, Xiaodong, et al.
Published: (2025)
by: Wang, Xiaodong, et al.
Published: (2025)
Differentially Private Conformal Prediction
by: Wu, Jiamei, et al.
Published: (2026)
by: Wu, Jiamei, et al.
Published: (2026)
SuS: Strategy-aware Surprise for Intrinsic Exploration
by: Kashirskiy, Mark, et al.
Published: (2026)
by: Kashirskiy, Mark, et al.
Published: (2026)
Federated Distributional Reinforcement Learning with Distributional Critic Regularization
by: Millard, David, et al.
Published: (2026)
by: Millard, David, et al.
Published: (2026)
Exploring the Relationships Among Teacher Autonomy Support, Intrinsic Motivation, Academic Resilience, Learning Engagement, and Well‐Being: A Mixed Methods Study
by: Yafei Shi, et al.
Published: (2026)
by: Yafei Shi, et al.
Published: (2026)
Goal Exploration via Adaptive Skill Distribution for Goal-Conditioned Reinforcement Learning
by: Wu, Lisheng, et al.
Published: (2024)
by: Wu, Lisheng, et al.
Published: (2024)
Learning Off-policy with Model-based Intrinsic Motivation For Active Online Exploration
by: Wang, Yibo, et al.
Published: (2024)
by: Wang, Yibo, et al.
Published: (2024)
Tape: A Cellular Automata Benchmark for Evaluating Rule-Shift Generalization in Reinforcement Learning
by: Pan, Enze
Published: (2026)
by: Pan, Enze
Published: (2026)
More Benefits of Being Distributional: Second-Order Bounds for Reinforcement Learning
by: Wang, Kaiwen, et al.
Published: (2024)
by: Wang, Kaiwen, et al.
Published: (2024)
Effective Pressure of the FRW Universe
by: Kong, Shi-Bei
Published: (2025)
by: Kong, Shi-Bei
Published: (2025)
Heat Capacity and the Violation of Scaling Laws in Gravitational System
by: Kong, Shi-Bei
Published: (2025)
by: Kong, Shi-Bei
Published: (2025)
Individual Contributions as Intrinsic Exploration Scaffolds for Multi-agent Reinforcement Learning
by: Li, Xinran, et al.
Published: (2024)
by: Li, Xinran, et al.
Published: (2024)
Decoding for Punctured Convolutional and Turbo Codes: A Deep Learning Solution for Protocols Compliance
by: Yan, Yongli, et al.
Published: (2025)
by: Yan, Yongli, et al.
Published: (2025)
T$^2$PO: Uncertainty-Guided Exploration Control for Stable Multi-Turn Agentic Reinforcement Learning
by: Wang, Haixin, et al.
Published: (2026)
by: Wang, Haixin, et al.
Published: (2026)
Reinforcement Learning in Categorical Cybernetics
by: Hedges, Jules, et al.
Published: (2024)
by: Hedges, Jules, et al.
Published: (2024)
In-Dataset Trajectory Return Regularization for Offline Preference-based Reinforcement Learning
by: Tu, Songjun, et al.
Published: (2024)
by: Tu, Songjun, et al.
Published: (2024)
Joint Intrinsic Motivation for Coordinated Exploration in Multi-Agent Deep Reinforcement Learning
by: Toquebiau, Maxime, et al.
Published: (2024)
by: Toquebiau, Maxime, et al.
Published: (2024)
MECap-R1: Emotion-aware Policy with Reinforcement Learning for Multimodal Emotion Captioning
by: Sun, Haoqin, et al.
Published: (2025)
by: Sun, Haoqin, et al.
Published: (2025)
Equation of the Perfect Fluid in the FRW Universe
by: Kong, Shi-Bei, et al.
Published: (2025)
by: Kong, Shi-Bei, et al.
Published: (2025)
Online federated learning framework for classification
by: Guo, Wenxing, et al.
Published: (2025)
by: Guo, Wenxing, et al.
Published: (2025)
Manifold-Aware Exploration for Reinforcement Learning in Video Generation
by: Zheng, Mingzhe, et al.
Published: (2026)
by: Zheng, Mingzhe, et al.
Published: (2026)
Similar Items
-
Distributional Reinforcement Learning with Regularized Wasserstein Loss
by: Sun, Ke, et al.
Published: (2022) -
How Does Return Distribution in Distributional Reinforcement Learning Help Optimization?
by: Sun, Ke, et al.
Published: (2022) -
Deep Fair Learning: A Unified Framework for Fine-tuning Representations with Sufficient Networks
by: Shi, Enze, et al.
Published: (2025) -
Non-Asymptotic Analysis of Online Local Private Learning with SGD
by: Shi, Enze, et al.
Published: (2025) -
Understanding Fairness and Prediction Error through Subspace Decomposition and Influence Analysis
by: Shi, Enze, et al.
Published: (2025)