Reinforcement Learning in Non-Markovian Environments
Fuente:
arXiv
Guardado en:
| Autores principales: | Chandak, Siddharth, Shah, Pratik, Borkar, Vivek S, Dodhia, Parth |
|---|---|
| Formato: | Preprint |
| Publicado: |
2022
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
A Concentration Bound for TD(0) with Function Approximation
por: Chandak, Siddharth, et al.
Publicado: (2023)
por: Chandak, Siddharth, et al.
Publicado: (2023)
$O(1/k)$ Finite-Time Bound for Non-Linear Two-Time-Scale Stochastic Approximation
por: Chandak, Siddharth
Publicado: (2025)
por: Chandak, Siddharth
Publicado: (2025)
Finite-Time Bounds for Two-Time-Scale Stochastic Approximation with Arbitrary Norm Contractions and Markovian Noise
por: Chandak, Siddharth, et al.
Publicado: (2025)
por: Chandak, Siddharth, et al.
Publicado: (2025)
High-Probability Bounds for SGD under the Polyak-Lojasiewicz Condition with Markovian Noise
por: Kar, Avik, et al.
Publicado: (2026)
por: Kar, Avik, et al.
Publicado: (2026)
Learning to Control Unknown Strongly Monotone Games
por: Chandak, Siddharth, et al.
Publicado: (2024)
por: Chandak, Siddharth, et al.
Publicado: (2024)
Lagrangian Index Policy for Restless Bandits with Average Reward
por: Avrachenkov, Konstantin, et al.
Publicado: (2024)
por: Avrachenkov, Konstantin, et al.
Publicado: (2024)
Last-Iterate Guarantees for Learning in Co-coercive Games
por: Chandak, Siddharth, et al.
Publicado: (2026)
por: Chandak, Siddharth, et al.
Publicado: (2026)
Heavy-Tailed and Long-Range Dependent Noise in Stochastic Approximation: A Finite-Time Analysis
por: Chandak, Siddharth, et al.
Publicado: (2026)
por: Chandak, Siddharth, et al.
Publicado: (2026)
Choose Your Battles: Distributed Learning Over Multiple Tug of War Games
por: Chandak, Siddharth, et al.
Publicado: (2025)
por: Chandak, Siddharth, et al.
Publicado: (2025)
Non-Expansive Mappings in Two-Time-Scale Stochastic Approximation: Finite-Time Analysis
por: Chandak, Siddharth
Publicado: (2025)
por: Chandak, Siddharth
Publicado: (2025)
End-to-End Learning Framework for Solving Non-Markovian Optimal Control
por: Zhang, Xiaole, et al.
Publicado: (2025)
por: Zhang, Xiaole, et al.
Publicado: (2025)
Regret and Sample Complexity of Online Q-Learning via Concentration of Stochastic Approximation with Time-Inhomogeneous Markov Chains
por: Singh, Rahul, et al.
Publicado: (2026)
por: Singh, Rahul, et al.
Publicado: (2026)
Learning Coverage Paths in Unknown Environments with Deep Reinforcement Learning
por: Jonnarth, Arvi, et al.
Publicado: (2023)
por: Jonnarth, Arvi, et al.
Publicado: (2023)
A dynamic view of some anomalous phenomena in SGD
por: Borkar, Vivek Shripad
Publicado: (2025)
por: Borkar, Vivek Shripad
Publicado: (2025)
BEAVER: Building Environments with Assessable Variation for Evaluating Multi-Objective Reinforcement Learning
por: Liu, Ruohong, et al.
Publicado: (2025)
por: Liu, Ruohong, et al.
Publicado: (2025)
Control in Stochastic Environment with Delays: A Model-based Reinforcement Learning Approach
por: Yao, Zhiyuan, et al.
Publicado: (2024)
por: Yao, Zhiyuan, et al.
Publicado: (2024)
Adaptive Testing Environment Generation for Connected and Automated Vehicles with Dense Reinforcement Learning
por: Yang, Jingxuan, et al.
Publicado: (2024)
por: Yang, Jingxuan, et al.
Publicado: (2024)
Heat Death of Generative Models in Closed-Loop Learning
por: Marchi, Matteo, et al.
Publicado: (2024)
por: Marchi, Matteo, et al.
Publicado: (2024)
Policy Gradient Methods for Non-Markovian Reinforcement Learning
por: Kar, Avik, et al.
Publicado: (2026)
por: Kar, Avik, et al.
Publicado: (2026)
Dyna-Style Reinforcement Learning Modeling and Control of Non-linear Dynamics
por: Abdelsalam, Karim, et al.
Publicado: (2025)
por: Abdelsalam, Karim, et al.
Publicado: (2025)
ExARNN: An Environment-Driven Adaptive RNN for Learning Non-Stationary Power Dynamics
por: Li, Haoran, et al.
Publicado: (2025)
por: Li, Haoran, et al.
Publicado: (2025)
Adversarially-Robust TD Learning with Markovian Data: Finite-Time Rates and Fundamental Limits
por: Maity, Sreejeet, et al.
Publicado: (2025)
por: Maity, Sreejeet, et al.
Publicado: (2025)
Action Mapping for Reinforcement Learning in Continuous Environments with Constraints
por: Theile, Mirco, et al.
Publicado: (2024)
por: Theile, Mirco, et al.
Publicado: (2024)
RL-ADN: A High-Performance Deep Reinforcement Learning Environment for Optimal Energy Storage Systems Dispatch in Active Distribution Networks
por: Hou, Shengren, et al.
Publicado: (2024)
por: Hou, Shengren, et al.
Publicado: (2024)
Stochastic Approximation with Unbounded Markovian Noise: A General-Purpose Theorem
por: Haque, Shaan Ul, et al.
Publicado: (2024)
por: Haque, Shaan Ul, et al.
Publicado: (2024)
Momentum for the Win: Collaborative Federated Reinforcement Learning across Heterogeneous Environments
por: Wang, Han, et al.
Publicado: (2024)
por: Wang, Han, et al.
Publicado: (2024)
Leveling the Playing Field: Carefully Comparing Classical and Learned Controllers for Quadrotor Trajectory Tracking
por: Kunapuli, Pratik, et al.
Publicado: (2025)
por: Kunapuli, Pratik, et al.
Publicado: (2025)
Resilient Constrained Reinforcement Learning
por: Ding, Dongsheng, et al.
Publicado: (2023)
por: Ding, Dongsheng, et al.
Publicado: (2023)
Stone Soup Multi-Target Tracking Feature Extraction For Autonomous Search And Track In Deep Reinforcement Learning Environment
por: Ewers, Jan-Hendrik, et al.
Publicado: (2025)
por: Ewers, Jan-Hendrik, et al.
Publicado: (2025)
On the Foundation of Distributionally Robust Reinforcement Learning
por: Wang, Shengbo, et al.
Publicado: (2023)
por: Wang, Shengbo, et al.
Publicado: (2023)
Learn for Variation: Variationally Guided AAV Trajectory Learning in Differentiable Environments
por: Wang, Xiucheng, et al.
Publicado: (2026)
por: Wang, Xiucheng, et al.
Publicado: (2026)
Tight Finite Time Bounds of Two-Time-Scale Linear Stochastic Approximation with Markovian Noise
por: Haque, Shaan Ul, et al.
Publicado: (2023)
por: Haque, Shaan Ul, et al.
Publicado: (2023)
Learning the Optimal Power Flow: Environment Design Matters
por: Wolgast, Thomas, et al.
Publicado: (2024)
por: Wolgast, Thomas, et al.
Publicado: (2024)
Leveraging Symmetry to Accelerate Learning of Trajectory Tracking Controllers for Free-Flying Robotic Systems
por: Welde, Jake, et al.
Publicado: (2024)
por: Welde, Jake, et al.
Publicado: (2024)
Proximal Reliability Optimization for Reinforcement Learning
por: Patwardhan, Narendra, et al.
Publicado: (2019)
por: Patwardhan, Narendra, et al.
Publicado: (2019)
Offline Reinforcement Learning via Inverse Optimization
por: Dimanidis, Ioannis, et al.
Publicado: (2025)
por: Dimanidis, Ioannis, et al.
Publicado: (2025)
Concurrent Learning of Policy and Unknown Safety Constraints in Reinforcement Learning
por: Yifru, Lunet, et al.
Publicado: (2024)
por: Yifru, Lunet, et al.
Publicado: (2024)
Cross-fitted Proximal Learning for Model-Based Reinforcement Learning
por: Venkatesh, Nishanth, et al.
Publicado: (2026)
por: Venkatesh, Nishanth, et al.
Publicado: (2026)
SINDy-RL: Interpretable and Efficient Model-Based Reinforcement Learning
por: Zolman, Nicholas, et al.
Publicado: (2024)
por: Zolman, Nicholas, et al.
Publicado: (2024)
A Survey on Reinforcement Learning in Aviation Applications
por: Razzaghi, Pouria, et al.
Publicado: (2022)
por: Razzaghi, Pouria, et al.
Publicado: (2022)
Ejemplares similares
-
A Concentration Bound for TD(0) with Function Approximation
por: Chandak, Siddharth, et al.
Publicado: (2023) -
$O(1/k)$ Finite-Time Bound for Non-Linear Two-Time-Scale Stochastic Approximation
por: Chandak, Siddharth
Publicado: (2025) -
Finite-Time Bounds for Two-Time-Scale Stochastic Approximation with Arbitrary Norm Contractions and Markovian Noise
por: Chandak, Siddharth, et al.
Publicado: (2025) -
High-Probability Bounds for SGD under the Polyak-Lojasiewicz Condition with Markovian Noise
por: Kar, Avik, et al.
Publicado: (2026) -
Learning to Control Unknown Strongly Monotone Games
por: Chandak, Siddharth, et al.
Publicado: (2024)