Accelerated Online Risk-Averse Policy Evaluation in POMDPs with Theoretical Guarantees and Novel CVaR Bounds
Fuente:
arXiv
Saved in:
| Main Authors: | Pariente, Yaacov, Indelman, Vadim |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Online Risk-Averse Planning in POMDPs Using Iterated CVaR Value Function
by: Pariente, Yaacov, et al.
Published: (2026)
by: Pariente, Yaacov, et al.
Published: (2026)
Simplification of Risk Averse POMDPs with Performance Guarantees
by: Pariente, Yaacov, et al.
Published: (2024)
by: Pariente, Yaacov, et al.
Published: (2024)
Bounding Conditional Value-at-Risk via Auxiliary Distributions with Bounded Discrepancies
by: Pariente, Yaacov, et al.
Published: (2025)
by: Pariente, Yaacov, et al.
Published: (2025)
POMDPPlanners: Open-Source Package for POMDP Planning
by: Pariente, Yaacov, et al.
Published: (2026)
by: Pariente, Yaacov, et al.
Published: (2026)
Towards Optimal Performance and Action Consistency Guarantees in Dec-POMDPs with Inconsistent Beliefs and Limited Communication
by: Shimron, Moshe Rafaeli, et al.
Published: (2025)
by: Shimron, Moshe Rafaeli, et al.
Published: (2025)
Online POMDP Planning with Anytime Deterministic Optimality Guarantees
by: Barenboim, Moran, et al.
Published: (2023)
by: Barenboim, Moran, et al.
Published: (2023)
Return Capping: Sample-Efficient CVaR Policy Gradient Optimisation
by: Mead, Harry, et al.
Published: (2025)
by: Mead, Harry, et al.
Published: (2025)
A Martingale approach to continuous Portfolio Optimization under CVaR like constraints
by: Lelong, Jérôme, et al.
Published: (2025)
by: Lelong, Jérôme, et al.
Published: (2025)
Measurement Simplification in ρ-POMDP with Performance Guarantees
by: Yotam, Tom, et al.
Published: (2023)
by: Yotam, Tom, et al.
Published: (2023)
No Compromise in Solution Quality: Speeding Up Belief-dependent Continuous POMDPs via Adaptive Multilevel Simplification
by: Zhitnikov, Andrey, et al.
Published: (2023)
by: Zhitnikov, Andrey, et al.
Published: (2023)
Can Generative Artificial Intelligence Survive Data Contamination? Theoretical Guarantees under Contaminated Recursive Training
by: Wang, Kevin, et al.
Published: (2026)
by: Wang, Kevin, et al.
Published: (2026)
Conditional Performance Guarantee for Large Reasoning Models
by: Huang, Jianguo, et al.
Published: (2026)
by: Huang, Jianguo, et al.
Published: (2026)
A Computational Theory for Efficient Mini Agent Evaluation with Causal Guarantees
by: Yan, Hedong
Published: (2025)
by: Yan, Hedong
Published: (2025)
Anytime Probabilistically Constrained Provably Convergent Online Belief Space Planning
by: Zhitnikov, Andrey, et al.
Published: (2024)
by: Zhitnikov, Andrey, et al.
Published: (2024)
A Measure-Theoretic Axiomatisation of Causality
by: Park, Junhyung, et al.
Published: (2023)
by: Park, Junhyung, et al.
Published: (2023)
On the Provable Performance Guarantee of Efficient Reasoning Models
by: Zeng, Hao, et al.
Published: (2025)
by: Zeng, Hao, et al.
Published: (2025)
Simplifying Complex Observation Models in Continuous POMDP Planning with Probabilistic Guarantees and Practice
by: Lev-Yehudi, Idan, et al.
Published: (2023)
by: Lev-Yehudi, Idan, et al.
Published: (2023)
Reward Redistribution for CVaR MDPs using a Bellman Operator on L-infinity
by: Muni, Aneri, et al.
Published: (2026)
by: Muni, Aneri, et al.
Published: (2026)
Adaptive Insurance Reserving with CVaR-Constrained Reinforcement Learning under Macroeconomic Regimes
by: Dong, Stella C.
Published: (2025)
by: Dong, Stella C.
Published: (2025)
The Geometry of Knowing: From Possibilistic Ignorance to Probabilistic Certainty -- A Measure-Theoretic Framework for Epistemic Convergence
by: Jah, Moriba Kemessia
Published: (2026)
by: Jah, Moriba Kemessia
Published: (2026)
Guaranteed Recovery of Unambiguous Clusters
by: Mazooji, Kayvon, et al.
Published: (2025)
by: Mazooji, Kayvon, et al.
Published: (2025)
Online Learning with Unknown Constraints
by: Sridharan, Karthik, et al.
Published: (2024)
by: Sridharan, Karthik, et al.
Published: (2024)
Conformal Policy Control
by: Prinster, Drew, et al.
Published: (2026)
by: Prinster, Drew, et al.
Published: (2026)
Distributionally Robust Safety Verification of Neural Networks via Worst-Case CVaR
by: Kishida, Masako
Published: (2025)
by: Kishida, Masako
Published: (2025)
Finite-Time Analysis of MCTS in Continuous POMDP Planning
by: Kong, Da, et al.
Published: (2026)
by: Kong, Da, et al.
Published: (2026)
Minimax Optimal Variance-Aware Regret Bounds for Multinomial Logistic MDPs
by: Boudart, Pierre, et al.
Published: (2026)
by: Boudart, Pierre, et al.
Published: (2026)
Generalization Bounds: Perspectives from Information Theory and PAC-Bayes
by: Hellström, Fredrik, et al.
Published: (2023)
by: Hellström, Fredrik, et al.
Published: (2023)
Neural Networks Learn Generic Multi-Index Models Near Information-Theoretic Limit
by: Zhang, Bohan, et al.
Published: (2025)
by: Zhang, Bohan, et al.
Published: (2025)
Outcome-Based Online Reinforcement Learning: Algorithms and Fundamental Limits
by: Chen, Fan, et al.
Published: (2025)
by: Chen, Fan, et al.
Published: (2025)
From Thomas Bayes to Big Data: On the feasibility of being a subjective Bayesian
by: Ritov, Ya'acov
Published: (2025)
by: Ritov, Ya'acov
Published: (2025)
No need for an oracle: the nonparametric maximum likelihood decision in the compound decision problem is minimax
by: Ritov, Ya'acov
Published: (2023)
by: Ritov, Ya'acov
Published: (2023)
A mixture of a normal distribution with random mean and variance -- Examples of inconsistency of maximum likelihood estimates
by: Ritov, Ya'acov
Published: (2024)
by: Ritov, Ya'acov
Published: (2024)
Tail-Aware Information-Theoretic Generalization for RLHF and SGLD
by: Zhang, Huiming, et al.
Published: (2026)
by: Zhang, Huiming, et al.
Published: (2026)
A Diffusion Analysis of Policy Gradient for Stochastic Bandits
by: Lattimore, Tor
Published: (2026)
by: Lattimore, Tor
Published: (2026)
The Good, the Bad, and the Sampled: a No-Regret Approach to Safe Online Classification
by: Baharav, Tavor Z., et al.
Published: (2025)
by: Baharav, Tavor Z., et al.
Published: (2025)
Risk Analysis and Design Against Adversarial Actions
by: Campi, Marco C., et al.
Published: (2025)
by: Campi, Marco C., et al.
Published: (2025)
Le Cam Distortion: A Decision-Theoretic Framework for Robust Transfer Learning
by: Akdemir, Deniz
Published: (2025)
by: Akdemir, Deniz
Published: (2025)
RACER: Risk-Aware Calibrated Efficient Routing for Large Language Models
by: Hao, Sai, et al.
Published: (2026)
by: Hao, Sai, et al.
Published: (2026)
Generalizability of Neural Networks Minimizing Empirical Risk Based on Expressive Ability
by: Yu, Lijia, et al.
Published: (2025)
by: Yu, Lijia, et al.
Published: (2025)
Conformal Risk Control
by: Angelopoulos, Anastasios N., et al.
Published: (2022)
by: Angelopoulos, Anastasios N., et al.
Published: (2022)
Similar Items
-
Online Risk-Averse Planning in POMDPs Using Iterated CVaR Value Function
by: Pariente, Yaacov, et al.
Published: (2026) -
Simplification of Risk Averse POMDPs with Performance Guarantees
by: Pariente, Yaacov, et al.
Published: (2024) -
Bounding Conditional Value-at-Risk via Auxiliary Distributions with Bounded Discrepancies
by: Pariente, Yaacov, et al.
Published: (2025) -
POMDPPlanners: Open-Source Package for POMDP Planning
by: Pariente, Yaacov, et al.
Published: (2026) -
Towards Optimal Performance and Action Consistency Guarantees in Dec-POMDPs with Inconsistent Beliefs and Limited Communication
by: Shimron, Moshe Rafaeli, et al.
Published: (2025)