Posterior Sampling-based Online Learning for Episodic POMDPs
Fuente:
arXiv
Saved in:
| Main Authors: | Tang, Dengwang, Ye, Dongze, Jain, Rahul, Nayyar, Ashutosh, Nuzzo, Pierluigi |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Efficient Online Learning with Offline Datasets for Infinite Horizon MDPs: A Bayesian Approach
by: Tang, Dengwang, et al.
Published: (2023)
by: Tang, Dengwang, et al.
Published: (2023)
Multi-agent assignment via state augmented reinforcement learning
by: Agorio, Leopoldo, et al.
Published: (2024)
by: Agorio, Leopoldo, et al.
Published: (2024)
Object Tracking Incorporating Transfer Learning into Unscented and Cubature Kalman Filters
by: Alotaibi, Omar, et al.
Published: (2024)
by: Alotaibi, Omar, et al.
Published: (2024)
Nonparametric Sparse Online Learning of the Koopman Operator
by: Hou, Boya, et al.
Published: (2024)
by: Hou, Boya, et al.
Published: (2024)
Optimal Control of Logically Constrained Partially Observable and Multi-Agent Markov Decision Processes
by: Kalagarla, Krishna C., et al.
Published: (2023)
by: Kalagarla, Krishna C., et al.
Published: (2023)
Automatic Link Selection in Multi-Channel Multiple Access with Link Failures
by: Wijewardena, Mevan, et al.
Published: (2025)
by: Wijewardena, Mevan, et al.
Published: (2025)
Robust Probability Hypothesis Density Filtering: Theory and Algorithms
by: Lei, Ming, et al.
Published: (2025)
by: Lei, Ming, et al.
Published: (2025)
Controllability and Vector Potential
by: Shankar, Shiva
Published: (2019)
by: Shankar, Shiva
Published: (2019)
On Robustness of Double Linear Policy with Time-Varying Weights
by: Wang, Xin-Yu, et al.
Published: (2023)
by: Wang, Xin-Yu, et al.
Published: (2023)
Sample Complexity of Policy Gradient for Log-Growth Control
by: Pan, Qiuhua, et al.
Published: (2026)
by: Pan, Qiuhua, et al.
Published: (2026)
Reinforcement Learning Method for Zero-Sum Linear-Quadratic Stochastic Differential Games in Infinite Horizons
by: Wang, Yiyuan
Published: (2026)
by: Wang, Yiyuan
Published: (2026)
Compositional Planning for Logically Constrained Multi-Agent Markov Decision Processes
by: Kalagarla, Krishna C., et al.
Published: (2024)
by: Kalagarla, Krishna C., et al.
Published: (2024)
On the continuity and smoothness of the value function in reinforcement learning and optimal control
by: Harder, Hans, et al.
Published: (2024)
by: Harder, Hans, et al.
Published: (2024)
Continuous-time Risk-sensitive Reinforcement Learning via Quadratic Variation Penalty
by: Jia, Yanwei
Published: (2024)
by: Jia, Yanwei
Published: (2024)
Pure Exploration for Constrained Best Mixed Arm Identification with a Fixed Budget
by: Tang, Dengwang, et al.
Published: (2024)
by: Tang, Dengwang, et al.
Published: (2024)
Optimal control of SDEs with merely measurable drift: an HJB approach
by: Du, Kai, et al.
Published: (2025)
by: Du, Kai, et al.
Published: (2025)
Finite-Time Analysis of Projected Two-Time-Scale Stochastic Approximation
by: Bai, Yitao, et al.
Published: (2026)
by: Bai, Yitao, et al.
Published: (2026)
Max-Entropy Moment Filtering for Stochastic Hybrid Systems
by: Iwasaki, Kaito, et al.
Published: (2026)
by: Iwasaki, Kaito, et al.
Published: (2026)
Continuous time Stochastic optimal control under discrete time partial observations
by: Bayer, Christian, et al.
Published: (2024)
by: Bayer, Christian, et al.
Published: (2024)
Mean--Variance Portfolio Selection by Continuous-Time Reinforcement Learning: Algorithms, Regret Analysis, and Empirical Study
by: Huang, Yilie, et al.
Published: (2024)
by: Huang, Yilie, et al.
Published: (2024)
Pointwise-Sparse Actuator Scheduling for Linear Systems with Controllability Guarantee
by: Ballotta, Luca, et al.
Published: (2024)
by: Ballotta, Luca, et al.
Published: (2024)
A Moreau Envelope Approach for LQR Meta-Policy Estimation
by: Aravind, Ashwin, et al.
Published: (2024)
by: Aravind, Ashwin, et al.
Published: (2024)
Convergence Analysis for Entropy-Regularized Control Problems: A Probabilistic Approach
by: Ma, Jin, et al.
Published: (2024)
by: Ma, Jin, et al.
Published: (2024)
Continuous Policy and Value Iteration for Stochastic Control Problems and Its Convergence
by: Feng, Qi, et al.
Published: (2025)
by: Feng, Qi, et al.
Published: (2025)
Markov Decision Processes of the Third Kind: Learning Distributions by Policy Gradient Descent
by: Bäuerle, Nicole, et al.
Published: (2026)
by: Bäuerle, Nicole, et al.
Published: (2026)
Optimistic Training and Convergence of Q-Learning -- Extended Version
by: Mehta, Prashant, et al.
Published: (2026)
by: Mehta, Prashant, et al.
Published: (2026)
Stability and Sensitivity Analysis of Relative Temporal-Difference Learning: Extended Version
by: Sakha, Masoud S., et al.
Published: (2026)
by: Sakha, Masoud S., et al.
Published: (2026)
Formalising the intentional stance 2: a coinductive approach
by: McGregor, Simon, et al.
Published: (2025)
by: McGregor, Simon, et al.
Published: (2025)
Formalising the intentional stance 1: attributing goals and beliefs to stochastic processes
by: McGregor, Simon, et al.
Published: (2024)
by: McGregor, Simon, et al.
Published: (2024)
The Quantum Advantage in Binary Teams and the Coordination Dilemma: Supplementary
by: Deshpande, Shashank A., et al.
Published: (2023)
by: Deshpande, Shashank A., et al.
Published: (2023)
Intermittent Encryption Strategies for Anti-Eavesdropping Estimation
by: Hu, Zhongyao, et al.
Published: (2024)
by: Hu, Zhongyao, et al.
Published: (2024)
Quantum advantage in decentralized control of POMDPs: A control-theoretic view of the Mermin-Peres square
by: Anantharam, Venkat
Published: (2025)
by: Anantharam, Venkat
Published: (2025)
Data-driven Control of T-Product-based Dynamical Systems
by: He, Ziqin, et al.
Published: (2025)
by: He, Ziqin, et al.
Published: (2025)
Fractional Backward Stochastic Partial Differential Equations with Applications to Stochastic Optimal Control of Partially Observed Systems driven by Lévy Processes
by: Ye, Yuyang, et al.
Published: (2024)
by: Ye, Yuyang, et al.
Published: (2024)
The Dynamic Search for the Minimal Dynamic Extension
by: D'Souza, Rollen S.
Published: (2026)
by: D'Souza, Rollen S.
Published: (2026)
Importance sampling for rare event tracking within the ensemble Kalman filtering framework
by: Rached, Nadhir Ben, et al.
Published: (2024)
by: Rached, Nadhir Ben, et al.
Published: (2024)
Policy Gradient for Continuous-Time Mean-Field Control
by: Bayraktar, Erhan, et al.
Published: (2026)
by: Bayraktar, Erhan, et al.
Published: (2026)
Robust Recurrence of Discrete-Time Infinite-Horizon Stochastic Optimal Control with Discounted Cost
by: Moldenhauer, Robert H., et al.
Published: (2025)
by: Moldenhauer, Robert H., et al.
Published: (2025)
Dynamic Weight Optimization for Double Linear Policy: A Stochastic Model Predictive Control Approach
by: Hong, Tan Chin, et al.
Published: (2026)
by: Hong, Tan Chin, et al.
Published: (2026)
Stochastic Model Predictive Control for Sub-Gaussian Noise
by: Ao, Yunke, et al.
Published: (2025)
by: Ao, Yunke, et al.
Published: (2025)
Similar Items
-
Efficient Online Learning with Offline Datasets for Infinite Horizon MDPs: A Bayesian Approach
by: Tang, Dengwang, et al.
Published: (2023) -
Multi-agent assignment via state augmented reinforcement learning
by: Agorio, Leopoldo, et al.
Published: (2024) -
Object Tracking Incorporating Transfer Learning into Unscented and Cubature Kalman Filters
by: Alotaibi, Omar, et al.
Published: (2024) -
Nonparametric Sparse Online Learning of the Koopman Operator
by: Hou, Boya, et al.
Published: (2024) -
Optimal Control of Logically Constrained Partially Observable and Multi-Agent Markov Decision Processes
by: Kalagarla, Krishna C., et al.
Published: (2023)