Finite time analysis of temporal difference learning with linear function approximation: Tail averaging and regularisation
Fuente:
arXiv
Salvato in:
| Autori principali: | Patil, Gandharv, A., Prashanth L., Nagaraj, Dheeraj, Precup, Doina |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2022
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Capacity-Constrained Continual Learning
di: Wen, Zheng, et al.
Pubblicazione: (2025)
di: Wen, Zheng, et al.
Pubblicazione: (2025)
Fluid-Agent Reinforcement Learning
di: Sharma, Shishir, et al.
Pubblicazione: (2026)
di: Sharma, Shishir, et al.
Pubblicazione: (2026)
Planning to avoid ambiguous states through Gaussian approximations to non-linear sensors in active inference agents
di: Kouw, Wouter M.
Pubblicazione: (2024)
di: Kouw, Wouter M.
Pubblicazione: (2024)
EVT-Based Generative AI for Tail-Aware Channel Estimation
di: Valiahdi, Parmida, et al.
Pubblicazione: (2026)
di: Valiahdi, Parmida, et al.
Pubblicazione: (2026)
A Look at Value-Based Decision-Time vs. Background Planning Methods Across Different Settings
di: Alver, Safa, et al.
Pubblicazione: (2022)
di: Alver, Safa, et al.
Pubblicazione: (2022)
Diversity-Enriched Option-Critic
di: Kamat, Anand, et al.
Pubblicazione: (2020)
di: Kamat, Anand, et al.
Pubblicazione: (2020)
Functional Acceleration for Policy Mirror Descent
di: Chelu, Veronica, et al.
Pubblicazione: (2024)
di: Chelu, Veronica, et al.
Pubblicazione: (2024)
An agent design with goal reaching guarantees for enhancement of learning
di: Osinenko, Pavel, et al.
Pubblicazione: (2024)
di: Osinenko, Pavel, et al.
Pubblicazione: (2024)
Stochastic Approximation with Delayed Updates: Finite-Time Rates under Markovian Sampling
di: Adibi, Arman, et al.
Pubblicazione: (2024)
di: Adibi, Arman, et al.
Pubblicazione: (2024)
How to discretize continuous state-action spaces in Q-learning: A symbolic control approach
di: Alaoui, Sadek Belamfedel, et al.
Pubblicazione: (2024)
di: Alaoui, Sadek Belamfedel, et al.
Pubblicazione: (2024)
Switching-time bioprocess control with pulse-width-modulated optogenetics
di: Espinel-Ríos, Sebastián
Pubblicazione: (2025)
di: Espinel-Ríos, Sebastián
Pubblicazione: (2025)
Real-time system optimal traffic routing under uncertainties -- Can physics models boost reinforcement learning?
di: Ke, Zemian, et al.
Pubblicazione: (2024)
di: Ke, Zemian, et al.
Pubblicazione: (2024)
Q-learning-based Model-free Safety Filter
di: Sue, Guo Ning, et al.
Pubblicazione: (2024)
di: Sue, Guo Ning, et al.
Pubblicazione: (2024)
Achieving Tighter Finite-Time Rates for Heterogeneous Federated Stochastic Approximation under Markovian Sampling
di: Zhu, Feng, et al.
Pubblicazione: (2025)
di: Zhu, Feng, et al.
Pubblicazione: (2025)
Testing predictive automated driving systems: lessons learned and future recommendations
di: Gonzalo, Rubén Izquierdo, et al.
Pubblicazione: (2022)
di: Gonzalo, Rubén Izquierdo, et al.
Pubblicazione: (2022)
Continual uncertainty learning
di: Yonezawa, Heisei, et al.
Pubblicazione: (2026)
di: Yonezawa, Heisei, et al.
Pubblicazione: (2026)
A Digital Twin prototype for traffic sign recognition of a learning-enabled autonomous vehicle
di: AbdElSalam, Mohamed, et al.
Pubblicazione: (2024)
di: AbdElSalam, Mohamed, et al.
Pubblicazione: (2024)
Reinforcement learning meets bioprocess control through behaviour cloning: Real-world deployment in an industrial photobioreactor
di: Gil, Juan D., et al.
Pubblicazione: (2025)
di: Gil, Juan D., et al.
Pubblicazione: (2025)
Contingency-constrained economic dispatch with safe reinforcement learning
di: Eichelbeck, Michael, et al.
Pubblicazione: (2022)
di: Eichelbeck, Michael, et al.
Pubblicazione: (2022)
Learning in Hybrid Active Inference Models
di: Collis, Poppy, et al.
Pubblicazione: (2024)
di: Collis, Poppy, et al.
Pubblicazione: (2024)
Hybrid Recurrent Models Support Emergent Descriptions for Hierarchical Planning and Control
di: Collis, Poppy, et al.
Pubblicazione: (2024)
di: Collis, Poppy, et al.
Pubblicazione: (2024)
Enhanced Transformer architecture for in-context learning of dynamical systems
di: Rufolo, Matteo, et al.
Pubblicazione: (2024)
di: Rufolo, Matteo, et al.
Pubblicazione: (2024)
Distributionally robust minimization in meta-learning for system identification
di: Rufolo, Matteo, et al.
Pubblicazione: (2025)
di: Rufolo, Matteo, et al.
Pubblicazione: (2025)
Shared learning of powertrain control policies for vehicle fleets
di: Kerbel, Lindsey, et al.
Pubblicazione: (2024)
di: Kerbel, Lindsey, et al.
Pubblicazione: (2024)
Explainable Representation of Finite-Memory Policies for POMDPs using Decision Trees
di: Azeem, Muqsit, et al.
Pubblicazione: (2024)
di: Azeem, Muqsit, et al.
Pubblicazione: (2024)
Stabilizing reinforcement learning control: A modular framework for optimizing over all stable behavior
di: Lawrence, Nathan P., et al.
Pubblicazione: (2023)
di: Lawrence, Nathan P., et al.
Pubblicazione: (2023)
A Decision-Language Model (DLM) for Dynamic Restless Multi-Armed Bandit Tasks in Public Health
di: Behari, Nikhil, et al.
Pubblicazione: (2024)
di: Behari, Nikhil, et al.
Pubblicazione: (2024)
Exploring a Physics-Informed Decision Transformer for Distribution System Restoration: Methodology and Performance Analysis
di: Zhao, Hong, et al.
Pubblicazione: (2024)
di: Zhao, Hong, et al.
Pubblicazione: (2024)
Learning a local trading strategy: deep reinforcement learning for grid-scale renewable energy integration
di: Ju, Caleb, et al.
Pubblicazione: (2024)
di: Ju, Caleb, et al.
Pubblicazione: (2024)
Guided Safe Shooting: model based reinforcement learning with safety constraints
di: Paolo, Giuseppe, et al.
Pubblicazione: (2022)
di: Paolo, Giuseppe, et al.
Pubblicazione: (2022)
Intelligent Duty Cycling Management and Wake-up for Energy Harvesting IoT Networks with Correlated Activity
di: Ruíz-Guirola, David E., et al.
Pubblicazione: (2024)
di: Ruíz-Guirola, David E., et al.
Pubblicazione: (2024)
CycLight: learning traffic signal cooperation with a cycle-level strategy
di: Han, Gengyue, et al.
Pubblicazione: (2024)
di: Han, Gengyue, et al.
Pubblicazione: (2024)
Multi-agent deep reinforcement learning with centralized training and decentralized execution for transportation infrastructure management
di: Saifullah, M., et al.
Pubblicazione: (2024)
di: Saifullah, M., et al.
Pubblicazione: (2024)
Real-time Vehicle-to-Vehicle Communication Based Network Cooperative Control System through Distributed Database and Multimodal Perception: Demonstrated in Crossroads
di: Zhu, Xinwen, et al.
Pubblicazione: (2024)
di: Zhu, Xinwen, et al.
Pubblicazione: (2024)
A Novel Feature Learning-based Bio-inspired Neural Network for Real-time Collision-free Rescue of Multi-Robot Systems
di: Li, Junfei, et al.
Pubblicazione: (2024)
di: Li, Junfei, et al.
Pubblicazione: (2024)
Deep reinforcement learning-based spacecraft attitude control with pointing keep-out constraint
di: Yang, Juntang, et al.
Pubblicazione: (2025)
di: Yang, Juntang, et al.
Pubblicazione: (2025)
Adaptive control of reaction-diffusion PDEs via neural operator-approximated gain kernels
di: Bhan, Luke, et al.
Pubblicazione: (2024)
di: Bhan, Luke, et al.
Pubblicazione: (2024)
Understanding the differences in Foundation Models: Attention, State Space Models, and Recurrent Neural Networks
di: Sieber, Jerome, et al.
Pubblicazione: (2024)
di: Sieber, Jerome, et al.
Pubblicazione: (2024)
Optimal Output Feedback Learning Control for Discrete-Time Linear Quadratic Regulation
di: Xie, Kedi, et al.
Pubblicazione: (2025)
di: Xie, Kedi, et al.
Pubblicazione: (2025)
Towards Autonomous Supply Chains: Definition, Characteristics, Conceptual Framework, and Autonomy Levels
di: Xu, Liming, et al.
Pubblicazione: (2023)
di: Xu, Liming, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Capacity-Constrained Continual Learning
di: Wen, Zheng, et al.
Pubblicazione: (2025) -
Fluid-Agent Reinforcement Learning
di: Sharma, Shishir, et al.
Pubblicazione: (2026) -
Planning to avoid ambiguous states through Gaussian approximations to non-linear sensors in active inference agents
di: Kouw, Wouter M.
Pubblicazione: (2024) -
EVT-Based Generative AI for Tail-Aware Channel Estimation
di: Valiahdi, Parmida, et al.
Pubblicazione: (2026) -
A Look at Value-Based Decision-Time vs. Background Planning Methods Across Different Settings
di: Alver, Safa, et al.
Pubblicazione: (2022)