Generalized Gaussian Temporal Difference Error for Uncertainty-aware Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Kim, Seyeon, Lee, Joonhun, Cho, Namhoon, Han, Sungjun, Hwang, Wooseop |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Feature-aligned N-BEATS with Sinkhorn divergence
by: Lee, Joonhun, et al.
Published: (2023)
by: Lee, Joonhun, et al.
Published: (2023)
Deep Reinforcement Learning and The Tale of Two Temporal Difference Errors
by: Rojas, Juan Sebastian, et al.
Published: (2026)
by: Rojas, Juan Sebastian, et al.
Published: (2026)
Generalisation of Total Uncertainty in AI: A Theoretical Study
by: Shariatmadar, Keivan
Published: (2024)
by: Shariatmadar, Keivan
Published: (2024)
SASSHA: Sharpness-aware Adaptive Second-order Optimization with Stable Hessian Approximation
by: Shin, Dahun, et al.
Published: (2025)
by: Shin, Dahun, et al.
Published: (2025)
Neural Network Parameter-optimization of Gaussian pmDAGs
by: Saremi, Mehrzad
Published: (2023)
by: Saremi, Mehrzad
Published: (2023)
Q-Learning under Finite Model Uncertainty
by: Sester, Julian, et al.
Published: (2024)
by: Sester, Julian, et al.
Published: (2024)
Backstepping Temporal Difference Learning
by: Lim, Han-Dong, et al.
Published: (2023)
by: Lim, Han-Dong, et al.
Published: (2023)
MINR: Implicit Neural Representations with Masked Image Modelling
by: Lee, Sua, et al.
Published: (2025)
by: Lee, Sua, et al.
Published: (2025)
Combining Statistical Depth and Fermat Distance for Uncertainty Quantification
by: Nguyen, Hai-Vy, et al.
Published: (2024)
by: Nguyen, Hai-Vy, et al.
Published: (2024)
A New Error Temporal Difference Algorithm for Deep Reinforcement Learning in Microgrid Optimization
by: Yao, Fulong, et al.
Published: (2025)
by: Yao, Fulong, et al.
Published: (2025)
Multi-hop Upstream Anticipatory Traffic Signal Control with Deep Reinforcement Learning
by: Li, Xiaocan, et al.
Published: (2024)
by: Li, Xiaocan, et al.
Published: (2024)
Strict Subgoal Execution: Reliable Long-Horizon Planning in Hierarchical Reinforcement Learning
by: Hwang, Jaebak, et al.
Published: (2025)
by: Hwang, Jaebak, et al.
Published: (2025)
Bellman Unbiasedness: Toward Provably Efficient Distributional Reinforcement Learning with General Value Function Approximation
by: Cho, Taehyun, et al.
Published: (2024)
by: Cho, Taehyun, et al.
Published: (2024)
Uncertainty Quantification for Transformer Models for Dark-Pattern Detection
by: Muñoz, Javier, et al.
Published: (2024)
by: Muñoz, Javier, et al.
Published: (2024)
Option-aware Temporally Abstracted Value for Offline Goal-Conditioned Reinforcement Learning
by: Ahn, Hongjoon, et al.
Published: (2025)
by: Ahn, Hongjoon, et al.
Published: (2025)
Feature Learning Dynamics in Infinite-Depth Neural Networks
by: Yao, Zihan, et al.
Published: (2025)
by: Yao, Zihan, et al.
Published: (2025)
Hybrid Probabilistic Forecasting of Under-Five Malaria Admissions in Ghana: A Gaussian Process Regression with Holt-Winters Smoothing
by: Ansah-Narh, T., et al.
Published: (2026)
by: Ansah-Narh, T., et al.
Published: (2026)
Advancing Deep Learning through Probability Engineering: A Pragmatic Paradigm for Modern AI
by: Zhang, Jianyi
Published: (2025)
by: Zhang, Jianyi
Published: (2025)
Robust $Q$-learning Algorithm for Markov Decision Processes under Wasserstein Uncertainty
by: Neufeld, Ariel, et al.
Published: (2022)
by: Neufeld, Ariel, et al.
Published: (2022)
Finite-Time Analysis of Temporal Difference Learning with Experience Replay
by: Lim, Han-Dong, et al.
Published: (2023)
by: Lim, Han-Dong, et al.
Published: (2023)
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning
by: Lee, Dongsu, et al.
Published: (2025)
by: Lee, Dongsu, et al.
Published: (2025)
ARCLE: The Abstraction and Reasoning Corpus Learning Environment for Reinforcement Learning
by: Lee, Hosung, et al.
Published: (2024)
by: Lee, Hosung, et al.
Published: (2024)
Stabilizing Temporal Difference Learning via Implicit Stochastic Recursion
by: Kim, Hwanwoo, et al.
Published: (2025)
by: Kim, Hwanwoo, et al.
Published: (2025)
Uncertainty Quantification with Bayesian Higher Order ReLU KANs
by: Giroux, James, et al.
Published: (2024)
by: Giroux, James, et al.
Published: (2024)
Temporal Difference Learning with Compressed Updates: Error-Feedback meets Reinforcement Learning
by: Mitra, Aritra, et al.
Published: (2023)
by: Mitra, Aritra, et al.
Published: (2023)
Spectral-Risk Safe Reinforcement Learning with Convergence Guarantees
by: Kim, Dohyeong, et al.
Published: (2024)
by: Kim, Dohyeong, et al.
Published: (2024)
SAFE: Finding Sparse and Flat Minima to Improve Pruning
by: Lee, Dongyeop, et al.
Published: (2025)
by: Lee, Dongyeop, et al.
Published: (2025)
SPQR: Controlling Q-ensemble Independence with Spiked Random Model for Reinforcement Learning
by: Lee, Dohyeok, et al.
Published: (2024)
by: Lee, Dohyeok, et al.
Published: (2024)
Representative Arm Identification: A fixed confidence approach to identify cluster representatives
by: Gharat, Sarvesh, et al.
Published: (2024)
by: Gharat, Sarvesh, et al.
Published: (2024)
Neural Laplace for learning Stochastic Differential Equations
by: Carrel, Adrien
Published: (2024)
by: Carrel, Adrien
Published: (2024)
Deep Conditional Measure Quantization
by: Turinici, Gabriel
Published: (2023)
by: Turinici, Gabriel
Published: (2023)
$α$-TCAV: A Unified Framework for Testing with Concept Activation Vectors
by: Schnoor, Ekkehard, et al.
Published: (2026)
by: Schnoor, Ekkehard, et al.
Published: (2026)
Causal Effect Identification in Heterogeneous Environments from Higher-Order Moments
by: Kivva, Yaroslav, et al.
Published: (2025)
by: Kivva, Yaroslav, et al.
Published: (2025)
Explicit Density Approximation for Neural Implicit Samplers Using a Bernstein-Based Convex Divergence
by: de Frutos, José Manuel, et al.
Published: (2025)
by: de Frutos, José Manuel, et al.
Published: (2025)
A Unified Theory of $θ$-Expectations
by: Qi, Qian
Published: (2025)
by: Qi, Qian
Published: (2025)
Nonparametric Identification of Latent Concepts
by: Zheng, Yujia, et al.
Published: (2025)
by: Zheng, Yujia, et al.
Published: (2025)
Greedy Selection under Independent Increments: A Toy Model Analysis
by: Yang, Huitao
Published: (2025)
by: Yang, Huitao
Published: (2025)
A Mean-Field Theory of $Θ$-Expectations
by: Qi, Qian
Published: (2025)
by: Qi, Qian
Published: (2025)
Neural Expectation Operators
by: Qi, Qian
Published: (2025)
by: Qi, Qian
Published: (2025)
Note on Martingale Theory and Applications
by: Zou, Xiandong
Published: (2026)
by: Zou, Xiandong
Published: (2026)
Similar Items
-
Feature-aligned N-BEATS with Sinkhorn divergence
by: Lee, Joonhun, et al.
Published: (2023) -
Deep Reinforcement Learning and The Tale of Two Temporal Difference Errors
by: Rojas, Juan Sebastian, et al.
Published: (2026) -
Generalisation of Total Uncertainty in AI: A Theoretical Study
by: Shariatmadar, Keivan
Published: (2024) -
SASSHA: Sharpness-aware Adaptive Second-order Optimization with Stable Hessian Approximation
by: Shin, Dahun, et al.
Published: (2025) -
Neural Network Parameter-optimization of Gaussian pmDAGs
by: Saremi, Mehrzad
Published: (2023)