Backstepping Temporal Difference Learning
Fuente:
arXiv
Salvato in:
| Autori principali: | Lim, Han-Dong, Lee, Donghwan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Finite-Time Analysis of Temporal Difference Learning with Experience Replay
di: Lim, Han-Dong, et al.
Pubblicazione: (2023)
di: Lim, Han-Dong, et al.
Pubblicazione: (2023)
Analysis of Off-Policy $n$-Step TD-Learning with Linear Function Approximation
di: Lim, Han-Dong, et al.
Pubblicazione: (2025)
di: Lim, Han-Dong, et al.
Pubblicazione: (2025)
Understanding the theoretical properties of projected Bellman equation, linear Q-learning, and approximate value iteration
di: Lim, Han-Dong, et al.
Pubblicazione: (2025)
di: Lim, Han-Dong, et al.
Pubblicazione: (2025)
R-GTD: A Geometric Analysis of Gradient Temporal-Difference Learning in Singular Regimes
di: Na, Hyunjun, et al.
Pubblicazione: (2026)
di: Na, Hyunjun, et al.
Pubblicazione: (2026)
Periodic Regularized Q-Learning
di: Yang, Hyukjun, et al.
Pubblicazione: (2026)
di: Yang, Hyukjun, et al.
Pubblicazione: (2026)
Learning the Model While Learning Q: Finite-Time Sample Complexity of Online SyncMBQ
di: Lim, Han-Dong, et al.
Pubblicazione: (2024)
di: Lim, Han-Dong, et al.
Pubblicazione: (2024)
A finite time analysis of distributed Q-learning
di: Lim, Han-Dong, et al.
Pubblicazione: (2024)
di: Lim, Han-Dong, et al.
Pubblicazione: (2024)
Safe-Support Q-Learning: Learning without Unsafe Exploration
di: Lim, Yeeun, et al.
Pubblicazione: (2026)
di: Lim, Yeeun, et al.
Pubblicazione: (2026)
New Versions of Gradient Temporal Difference Learning
di: Lee, Donghwan, et al.
Pubblicazione: (2021)
di: Lee, Donghwan, et al.
Pubblicazione: (2021)
Lyapunov-Certified Direct Switching Theory for Q-Learning
di: Lee, Donghwan
Pubblicazione: (2026)
di: Lee, Donghwan
Pubblicazione: (2026)
Suppressing Overestimation in Q-Learning through Adversarial Behaviors
di: Lee, HyeAnn, et al.
Pubblicazione: (2023)
di: Lee, HyeAnn, et al.
Pubblicazione: (2023)
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms
di: Lee, Donghwan, et al.
Pubblicazione: (2024)
di: Lee, Donghwan, et al.
Pubblicazione: (2024)
Taming the Adversary: Stable Minimax Deep Deterministic Policy Gradient via Fractional Objectives
di: Lee, Taeho, et al.
Pubblicazione: (2026)
di: Lee, Taeho, et al.
Pubblicazione: (2026)
Soft Deterministic Policy Gradient with Gaussian Smoothing
di: Na, Hyunjun, et al.
Pubblicazione: (2026)
di: Na, Hyunjun, et al.
Pubblicazione: (2026)
Adaptive Policy Backbone via Shared Network
di: Park, Bumgeun, et al.
Pubblicazione: (2025)
di: Park, Bumgeun, et al.
Pubblicazione: (2025)
Pretraining a Shared Q-Network for Data-Efficient Offline Reinforcement Learning
di: Park, Jongchan, et al.
Pubblicazione: (2025)
di: Park, Jongchan, et al.
Pubblicazione: (2025)
Generalized Gaussian Temporal Difference Error for Uncertainty-aware Reinforcement Learning
di: Kim, Seyeon, et al.
Pubblicazione: (2024)
di: Kim, Seyeon, et al.
Pubblicazione: (2024)
Deep Reinforcement Learning and The Tale of Two Temporal Difference Errors
di: Rojas, Juan Sebastian, et al.
Pubblicazione: (2026)
di: Rojas, Juan Sebastian, et al.
Pubblicazione: (2026)
A Switching System Theory of Q-Learning with Linear Function Approximation
di: Lee, Donghwan, et al.
Pubblicazione: (2026)
di: Lee, Donghwan, et al.
Pubblicazione: (2026)
Discerning Temporal Difference Learning
di: Ma, Jianfei
Pubblicazione: (2023)
di: Ma, Jianfei
Pubblicazione: (2023)
MahaVar: OOD Detection via Class-wise Mahalanobis Distance Variance under Neural Collapse
di: Kim, Donghwan, et al.
Pubblicazione: (2026)
di: Kim, Donghwan, et al.
Pubblicazione: (2026)
Mitigating the Likelihood Paradox in Flow-based OOD Detection via Entropy Manipulation
di: Kim, Donghwan, et al.
Pubblicazione: (2026)
di: Kim, Donghwan, et al.
Pubblicazione: (2026)
Temporal-Difference Variational Continual Learning
di: Melo, Luckeciano C., et al.
Pubblicazione: (2024)
di: Melo, Luckeciano C., et al.
Pubblicazione: (2024)
Gradient Iterated Temporal-Difference Learning
di: Vincent, Théo, et al.
Pubblicazione: (2026)
di: Vincent, Théo, et al.
Pubblicazione: (2026)
Demystifying the Recency Heuristic in Temporal-Difference Learning
di: Daley, Brett, et al.
Pubblicazione: (2024)
di: Daley, Brett, et al.
Pubblicazione: (2024)
Merge and Bound: Direct Manipulations on Weights for Class Incremental Learning
di: Kim, Taehoon, et al.
Pubblicazione: (2025)
di: Kim, Taehoon, et al.
Pubblicazione: (2025)
Is Temporal Difference Learning the Gold Standard for Stitching in RL?
di: Bortkiewicz, Michał, et al.
Pubblicazione: (2025)
di: Bortkiewicz, Michał, et al.
Pubblicazione: (2025)
A Variance Minimization Approach to Temporal-Difference Learning
di: Chen, Xingguo, et al.
Pubblicazione: (2024)
di: Chen, Xingguo, et al.
Pubblicazione: (2024)
Temporal Difference Flows
di: Farebrother, Jesse, et al.
Pubblicazione: (2025)
di: Farebrother, Jesse, et al.
Pubblicazione: (2025)
Regularized Q-learning
di: Lim, Han-Dong, et al.
Pubblicazione: (2022)
di: Lim, Han-Dong, et al.
Pubblicazione: (2022)
An MRP Formulation for Supervised Learning: Generalized Temporal Difference Learning Models
di: Pan, Yangchen, et al.
Pubblicazione: (2024)
di: Pan, Yangchen, et al.
Pubblicazione: (2024)
Why the Counterintuitive Phenomenon of Likelihood Rarely Appears in Tabular Anomaly Detection with Deep Generative Models?
di: Kim, Donghwan, et al.
Pubblicazione: (2026)
di: Kim, Donghwan, et al.
Pubblicazione: (2026)
Almost Sure Convergence of Linear Temporal Difference Learning with Arbitrary Features
di: Wang, Jiuqi, et al.
Pubblicazione: (2024)
di: Wang, Jiuqi, et al.
Pubblicazione: (2024)
An Analysis of Action-Value Temporal-Difference Methods That Learn State Values
di: Daley, Brett, et al.
Pubblicazione: (2025)
di: Daley, Brett, et al.
Pubblicazione: (2025)
Finite Sample Analysis of Linear Temporal Difference Learning with Arbitrary Features
di: Xie, Zixuan, et al.
Pubblicazione: (2025)
di: Xie, Zixuan, et al.
Pubblicazione: (2025)
Cochain Perspectives on Temporal-Difference Signals for Learning Beyond Markov Dynamics
di: Zhang, Zuyuan, et al.
Pubblicazione: (2026)
di: Zhang, Zuyuan, et al.
Pubblicazione: (2026)
OMG-RL:Offline Model-based Guided Reward Learning for Heparin Treatment
di: Lim, Yooseok, et al.
Pubblicazione: (2024)
di: Lim, Yooseok, et al.
Pubblicazione: (2024)
TANDEM: Temporal Attention-guided Neural Differential Equations for Missingness in Time Series Classification
di: Oh, YongKyung, et al.
Pubblicazione: (2025)
di: Oh, YongKyung, et al.
Pubblicazione: (2025)
A primal-dual perspective for distributed TD-learning
di: Lim, Han-Dong, et al.
Pubblicazione: (2023)
di: Lim, Han-Dong, et al.
Pubblicazione: (2023)
A New Error Temporal Difference Algorithm for Deep Reinforcement Learning in Microgrid Optimization
di: Yao, Fulong, et al.
Pubblicazione: (2025)
di: Yao, Fulong, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Finite-Time Analysis of Temporal Difference Learning with Experience Replay
di: Lim, Han-Dong, et al.
Pubblicazione: (2023) -
Analysis of Off-Policy $n$-Step TD-Learning with Linear Function Approximation
di: Lim, Han-Dong, et al.
Pubblicazione: (2025) -
Understanding the theoretical properties of projected Bellman equation, linear Q-learning, and approximate value iteration
di: Lim, Han-Dong, et al.
Pubblicazione: (2025) -
R-GTD: A Geometric Analysis of Gradient Temporal-Difference Learning in Singular Regimes
di: Na, Hyunjun, et al.
Pubblicazione: (2026) -
Periodic Regularized Q-Learning
di: Yang, Hyukjun, et al.
Pubblicazione: (2026)