Q-Learning for Stochastic Control under General Information Structures and Non-Markovian Environments
Fuente:
arXiv
Saved in:
| Main Authors: | Kara, Ali Devran, Yuksel, Serdar |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Partially Observed Optimal Stochastic Control: Regularity, Optimality, Approximations, and Learning
by: Kara, Ali Devran, et al.
Published: (2024)
by: Kara, Ali Devran, et al.
Published: (2024)
Near Optimal Approximations and Finite Memory Policies for POMPDs with Continuous Spaces
by: Kara, Ali Devran, et al.
Published: (2024)
by: Kara, Ali Devran, et al.
Published: (2024)
Sensitivity of Filter Kernels and Robustness Bounds to Transition and Measurement Kernel Perturbations in Partially Observable Stochastic Control
by: Demirci, Yunus Emre, et al.
Published: (2025)
by: Demirci, Yunus Emre, et al.
Published: (2025)
Learning POMDPs with Linear Function Approximation and Finite Memory
by: Kara, Ali Devran
Published: (2025)
by: Kara, Ali Devran
Published: (2025)
Refined Bounds on Near Optimality Finite Window Policies in POMDPs and Their Reinforcement Learning
by: Demirci, Yunus Emre, et al.
Published: (2024)
by: Demirci, Yunus Emre, et al.
Published: (2024)
Stochastic Approximation with Delayed Updates: Finite-Time Rates under Markovian Sampling
by: Adibi, Arman, et al.
Published: (2024)
by: Adibi, Arman, et al.
Published: (2024)
Reinforcement Learning with Function Approximation for Non-Markov Processes
by: Kara, Ali Devran
Published: (2026)
by: Kara, Ali Devran
Published: (2026)
Achieving Tighter Finite-Time Rates for Heterogeneous Federated Stochastic Approximation under Markovian Sampling
by: Zhu, Feng, et al.
Published: (2025)
by: Zhu, Feng, et al.
Published: (2025)
Kernel Mean Embedding Topology: Weak and Strong Forms for Stochastic Kernels and Implications for Model Learning
by: Saldi, Naci, et al.
Published: (2025)
by: Saldi, Naci, et al.
Published: (2025)
Robustness to Model Approximation, Model Learning From Data, and Sample Complexity in Wasserstein Regular MDPs
by: Zhou, Yichen, et al.
Published: (2024)
by: Zhou, Yichen, et al.
Published: (2024)
Safe Control and Learning Using Generalized Action Governor
by: Fang, Peiyuan, et al.
Published: (2022)
by: Fang, Peiyuan, et al.
Published: (2022)
Approximate Information States for Worst-Case Control and Learning in Uncertain Systems
by: Dave, Aditya, et al.
Published: (2023)
by: Dave, Aditya, et al.
Published: (2023)
Quantizer Design for Finite Model Approximations, Model Learning, and Quantized Q-Learning for MDPs with Unbounded Spaces
by: Bicer, Osman, et al.
Published: (2025)
by: Bicer, Osman, et al.
Published: (2025)
Average Cost Optimality of Partially Observed MDPS: Contraction of Non-linear Filters, Optimal Solutions and Approximations
by: Demirci, Yunus Emre, et al.
Published: (2023)
by: Demirci, Yunus Emre, et al.
Published: (2023)
Another Look at Partially Observed Optimal Stochastic Control: Existence, Ergodicity, and Approximations without Belief-Reduction
by: Yüksel, Serdar
Published: (2023)
by: Yüksel, Serdar
Published: (2023)
Reinforcement Learning for Discounted and Ergodic Control of Diffusion Processes
by: Bayraktar, Erhan, et al.
Published: (2026)
by: Bayraktar, Erhan, et al.
Published: (2026)
Quantum Markov Decision Processes: General Theory, Approximations, and Classes of Policies
by: Saldi, Naci, et al.
Published: (2024)
by: Saldi, Naci, et al.
Published: (2024)
Stochastic Learning of Computational Resource Usage as Graph Structured Multimarginal Schrödinger Bridge
by: Bondar, Georgiy A., et al.
Published: (2024)
by: Bondar, Georgiy A., et al.
Published: (2024)
Reinforcement Learning for Jointly Optimal Coding and Control Policies for a Controlled Markovian System over a Communication Channel
by: Hubbard, Evelyn, et al.
Published: (2024)
by: Hubbard, Evelyn, et al.
Published: (2024)
Decentralized Learning for Optimality in Stochastic Dynamic Teams and Games with Local Control and Global State Information
by: Yongacoglu, Bora, et al.
Published: (2019)
by: Yongacoglu, Bora, et al.
Published: (2019)
Nonlinear Control Allocation: A Learning Based Approach
by: Khan, Hafiz Zeeshan Iqbal, et al.
Published: (2022)
by: Khan, Hafiz Zeeshan Iqbal, et al.
Published: (2022)
Distributionally Robust Control with End-to-End Statistically Guaranteed Metric Learning
by: Wu, Jingyi, et al.
Published: (2025)
by: Wu, Jingyi, et al.
Published: (2025)
Learning with Linear Function Approximations in Mean-Field Control
by: Bayraktar, Erhan, et al.
Published: (2024)
by: Bayraktar, Erhan, et al.
Published: (2024)
Approximate Model Predictive Control for Microgrid Energy Management via Imitation Learning
by: Liu, Changrui, et al.
Published: (2025)
by: Liu, Changrui, et al.
Published: (2025)
Non-Sequential Decentralized Stochastic Control Revisited: Causality and Static Reducibility
by: Mrani-Zentar, Omar, et al.
Published: (2022)
by: Mrani-Zentar, Omar, et al.
Published: (2022)
Model Predictive Control and Reinforcement Learning: A Unified Framework Based on Dynamic Programming
by: Bertsekas, Dimitri P.
Published: (2024)
by: Bertsekas, Dimitri P.
Published: (2024)
Multi-Robot Multi-Queue Control via Exhaustive Assignment Actor-Critic Learning
by: Merati, Mohammad, et al.
Published: (2026)
by: Merati, Mohammad, et al.
Published: (2026)
Benchmarking Reinforcement Learning via Stochastic Converse Optimality: Generating Systems with Known Optimal Policies
by: Ibrahim, Sinan, et al.
Published: (2026)
by: Ibrahim, Sinan, et al.
Published: (2026)
Optimality of Symmetric Independent Policies under Decentralized Mean-Field Information Sharing for Stochastic Teams and Equivalence with McKean-Vlasov Control of a Representative Agent
by: Sanjari, Sina, et al.
Published: (2024)
by: Sanjari, Sina, et al.
Published: (2024)
On Borkar and Young Relaxed Control Topologies and Continuous Dependence of Invariant Measures on Control Policy
by: Yüksel, Serdar
Published: (2023)
by: Yüksel, Serdar
Published: (2023)
Unifying Controller Design for Stabilizing Nonlinear Systems with Norm-Bounded Control Inputs
by: Li, Ming, et al.
Published: (2024)
by: Li, Ming, et al.
Published: (2024)
Quantum Markov Decision Processes: Dynamic and Semi-Definite Programs for Optimal Solutions
by: Saldi, Naci, et al.
Published: (2024)
by: Saldi, Naci, et al.
Published: (2024)
Finite Approximations for Mean Field Type Multi-Agent Control and Their Near Optimality
by: Bayraktar, Erhan, et al.
Published: (2022)
by: Bayraktar, Erhan, et al.
Published: (2022)
Deep Reinforcement Learning for Multi-Objective Optimization: Enhancing Wind Turbine Energy Generation while Mitigating Noise Emissions
by: de Frutos, Martín, et al.
Published: (2024)
by: de Frutos, Martín, et al.
Published: (2024)
Optimal Output Feedback Learning Control for Discrete-Time Linear Quadratic Regulation
by: Xie, Kedi, et al.
Published: (2025)
by: Xie, Kedi, et al.
Published: (2025)
Policy Optimization for PDE Control with a Warm Start
by: Zhang, Xiangyuan, et al.
Published: (2024)
by: Zhang, Xiangyuan, et al.
Published: (2024)
Hierarchical Reinforcement Learning with Low-Level MPC for Multi-Agent Control
by: Studt, Max, et al.
Published: (2025)
by: Studt, Max, et al.
Published: (2025)
Graphon Particle Systems, Part II: Dynamics of Distributed Stochastic Continuum Optimization
by: Chen, Yan, et al.
Published: (2024)
by: Chen, Yan, et al.
Published: (2024)
Using Laplace Transform To Optimize the Hallucination of Generation Models
by: Kang, Cheng, et al.
Published: (2026)
by: Kang, Cheng, et al.
Published: (2026)
Multi-Year Maintenance Planning for Large-Scale Infrastructure Systems: A Novel Network Deep Q-Learning Approach
by: Fard, Amir, et al.
Published: (2025)
by: Fard, Amir, et al.
Published: (2025)
Similar Items
-
Partially Observed Optimal Stochastic Control: Regularity, Optimality, Approximations, and Learning
by: Kara, Ali Devran, et al.
Published: (2024) -
Near Optimal Approximations and Finite Memory Policies for POMPDs with Continuous Spaces
by: Kara, Ali Devran, et al.
Published: (2024) -
Sensitivity of Filter Kernels and Robustness Bounds to Transition and Measurement Kernel Perturbations in Partially Observable Stochastic Control
by: Demirci, Yunus Emre, et al.
Published: (2025) -
Learning POMDPs with Linear Function Approximation and Finite Memory
by: Kara, Ali Devran
Published: (2025) -
Refined Bounds on Near Optimality Finite Window Policies in POMDPs and Their Reinforcement Learning
by: Demirci, Yunus Emre, et al.
Published: (2024)