Reinforcement Learning in Non-Markovian Environments
Fuente:
arXiv
Saved in:
| Main Authors: | Chandak, Siddharth, Shah, Pratik, Borkar, Vivek S, Dodhia, Parth |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Concentration Bound for TD(0) with Function Approximation
by: Chandak, Siddharth, et al.
Published: (2023)
by: Chandak, Siddharth, et al.
Published: (2023)
$O(1/k)$ Finite-Time Bound for Non-Linear Two-Time-Scale Stochastic Approximation
by: Chandak, Siddharth
Published: (2025)
by: Chandak, Siddharth
Published: (2025)
Finite-Time Bounds for Two-Time-Scale Stochastic Approximation with Arbitrary Norm Contractions and Markovian Noise
by: Chandak, Siddharth, et al.
Published: (2025)
by: Chandak, Siddharth, et al.
Published: (2025)
High-Probability Bounds for SGD under the Polyak-Lojasiewicz Condition with Markovian Noise
by: Kar, Avik, et al.
Published: (2026)
by: Kar, Avik, et al.
Published: (2026)
Learning to Control Unknown Strongly Monotone Games
by: Chandak, Siddharth, et al.
Published: (2024)
by: Chandak, Siddharth, et al.
Published: (2024)
Lagrangian Index Policy for Restless Bandits with Average Reward
by: Avrachenkov, Konstantin, et al.
Published: (2024)
by: Avrachenkov, Konstantin, et al.
Published: (2024)
Last-Iterate Guarantees for Learning in Co-coercive Games
by: Chandak, Siddharth, et al.
Published: (2026)
by: Chandak, Siddharth, et al.
Published: (2026)
Heavy-Tailed and Long-Range Dependent Noise in Stochastic Approximation: A Finite-Time Analysis
by: Chandak, Siddharth, et al.
Published: (2026)
by: Chandak, Siddharth, et al.
Published: (2026)
Choose Your Battles: Distributed Learning Over Multiple Tug of War Games
by: Chandak, Siddharth, et al.
Published: (2025)
by: Chandak, Siddharth, et al.
Published: (2025)
Non-Expansive Mappings in Two-Time-Scale Stochastic Approximation: Finite-Time Analysis
by: Chandak, Siddharth
Published: (2025)
by: Chandak, Siddharth
Published: (2025)
End-to-End Learning Framework for Solving Non-Markovian Optimal Control
by: Zhang, Xiaole, et al.
Published: (2025)
by: Zhang, Xiaole, et al.
Published: (2025)
Regret and Sample Complexity of Online Q-Learning via Concentration of Stochastic Approximation with Time-Inhomogeneous Markov Chains
by: Singh, Rahul, et al.
Published: (2026)
by: Singh, Rahul, et al.
Published: (2026)
Learning Coverage Paths in Unknown Environments with Deep Reinforcement Learning
by: Jonnarth, Arvi, et al.
Published: (2023)
by: Jonnarth, Arvi, et al.
Published: (2023)
A dynamic view of some anomalous phenomena in SGD
by: Borkar, Vivek Shripad
Published: (2025)
by: Borkar, Vivek Shripad
Published: (2025)
BEAVER: Building Environments with Assessable Variation for Evaluating Multi-Objective Reinforcement Learning
by: Liu, Ruohong, et al.
Published: (2025)
by: Liu, Ruohong, et al.
Published: (2025)
Control in Stochastic Environment with Delays: A Model-based Reinforcement Learning Approach
by: Yao, Zhiyuan, et al.
Published: (2024)
by: Yao, Zhiyuan, et al.
Published: (2024)
Adaptive Testing Environment Generation for Connected and Automated Vehicles with Dense Reinforcement Learning
by: Yang, Jingxuan, et al.
Published: (2024)
by: Yang, Jingxuan, et al.
Published: (2024)
Heat Death of Generative Models in Closed-Loop Learning
by: Marchi, Matteo, et al.
Published: (2024)
by: Marchi, Matteo, et al.
Published: (2024)
Policy Gradient Methods for Non-Markovian Reinforcement Learning
by: Kar, Avik, et al.
Published: (2026)
by: Kar, Avik, et al.
Published: (2026)
Dyna-Style Reinforcement Learning Modeling and Control of Non-linear Dynamics
by: Abdelsalam, Karim, et al.
Published: (2025)
by: Abdelsalam, Karim, et al.
Published: (2025)
ExARNN: An Environment-Driven Adaptive RNN for Learning Non-Stationary Power Dynamics
by: Li, Haoran, et al.
Published: (2025)
by: Li, Haoran, et al.
Published: (2025)
Adversarially-Robust TD Learning with Markovian Data: Finite-Time Rates and Fundamental Limits
by: Maity, Sreejeet, et al.
Published: (2025)
by: Maity, Sreejeet, et al.
Published: (2025)
Action Mapping for Reinforcement Learning in Continuous Environments with Constraints
by: Theile, Mirco, et al.
Published: (2024)
by: Theile, Mirco, et al.
Published: (2024)
RL-ADN: A High-Performance Deep Reinforcement Learning Environment for Optimal Energy Storage Systems Dispatch in Active Distribution Networks
by: Hou, Shengren, et al.
Published: (2024)
by: Hou, Shengren, et al.
Published: (2024)
Stochastic Approximation with Unbounded Markovian Noise: A General-Purpose Theorem
by: Haque, Shaan Ul, et al.
Published: (2024)
by: Haque, Shaan Ul, et al.
Published: (2024)
Momentum for the Win: Collaborative Federated Reinforcement Learning across Heterogeneous Environments
by: Wang, Han, et al.
Published: (2024)
by: Wang, Han, et al.
Published: (2024)
Leveling the Playing Field: Carefully Comparing Classical and Learned Controllers for Quadrotor Trajectory Tracking
by: Kunapuli, Pratik, et al.
Published: (2025)
by: Kunapuli, Pratik, et al.
Published: (2025)
Resilient Constrained Reinforcement Learning
by: Ding, Dongsheng, et al.
Published: (2023)
by: Ding, Dongsheng, et al.
Published: (2023)
Stone Soup Multi-Target Tracking Feature Extraction For Autonomous Search And Track In Deep Reinforcement Learning Environment
by: Ewers, Jan-Hendrik, et al.
Published: (2025)
by: Ewers, Jan-Hendrik, et al.
Published: (2025)
On the Foundation of Distributionally Robust Reinforcement Learning
by: Wang, Shengbo, et al.
Published: (2023)
by: Wang, Shengbo, et al.
Published: (2023)
Learn for Variation: Variationally Guided AAV Trajectory Learning in Differentiable Environments
by: Wang, Xiucheng, et al.
Published: (2026)
by: Wang, Xiucheng, et al.
Published: (2026)
Tight Finite Time Bounds of Two-Time-Scale Linear Stochastic Approximation with Markovian Noise
by: Haque, Shaan Ul, et al.
Published: (2023)
by: Haque, Shaan Ul, et al.
Published: (2023)
Learning the Optimal Power Flow: Environment Design Matters
by: Wolgast, Thomas, et al.
Published: (2024)
by: Wolgast, Thomas, et al.
Published: (2024)
Leveraging Symmetry to Accelerate Learning of Trajectory Tracking Controllers for Free-Flying Robotic Systems
by: Welde, Jake, et al.
Published: (2024)
by: Welde, Jake, et al.
Published: (2024)
Proximal Reliability Optimization for Reinforcement Learning
by: Patwardhan, Narendra, et al.
Published: (2019)
by: Patwardhan, Narendra, et al.
Published: (2019)
Offline Reinforcement Learning via Inverse Optimization
by: Dimanidis, Ioannis, et al.
Published: (2025)
by: Dimanidis, Ioannis, et al.
Published: (2025)
Concurrent Learning of Policy and Unknown Safety Constraints in Reinforcement Learning
by: Yifru, Lunet, et al.
Published: (2024)
by: Yifru, Lunet, et al.
Published: (2024)
Cross-fitted Proximal Learning for Model-Based Reinforcement Learning
by: Venkatesh, Nishanth, et al.
Published: (2026)
by: Venkatesh, Nishanth, et al.
Published: (2026)
SINDy-RL: Interpretable and Efficient Model-Based Reinforcement Learning
by: Zolman, Nicholas, et al.
Published: (2024)
by: Zolman, Nicholas, et al.
Published: (2024)
A Survey on Reinforcement Learning in Aviation Applications
by: Razzaghi, Pouria, et al.
Published: (2022)
by: Razzaghi, Pouria, et al.
Published: (2022)
Similar Items
-
A Concentration Bound for TD(0) with Function Approximation
by: Chandak, Siddharth, et al.
Published: (2023) -
$O(1/k)$ Finite-Time Bound for Non-Linear Two-Time-Scale Stochastic Approximation
by: Chandak, Siddharth
Published: (2025) -
Finite-Time Bounds for Two-Time-Scale Stochastic Approximation with Arbitrary Norm Contractions and Markovian Noise
by: Chandak, Siddharth, et al.
Published: (2025) -
High-Probability Bounds for SGD under the Polyak-Lojasiewicz Condition with Markovian Noise
by: Kar, Avik, et al.
Published: (2026) -
Learning to Control Unknown Strongly Monotone Games
by: Chandak, Siddharth, et al.
Published: (2024)