Conformal Off-Policy Evaluation in Markov Decision Processes
Fuente:
arXiv
Saved in:
| Main Authors: | Foffano, Daniele, Russo, Alessio, Proutiere, Alexandre |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Adversarial Diffusion for Robust Reinforcement Learning
by: Foffano, Daniele, et al.
Published: (2025)
by: Foffano, Daniele, et al.
Published: (2025)
Receding-Horizon Control via Drifting Models
by: Foffano, Daniele, et al.
Published: (2026)
by: Foffano, Daniele, et al.
Published: (2026)
An Offline Risk-aware Policy Selection Method for Bayesian Markov Decision Processes
by: Angelotti, Giorgio, et al.
Published: (2021)
by: Angelotti, Giorgio, et al.
Published: (2021)
On the Convergence of Modified Policy Iteration in Risk Sensitive Exponential Cost Markov Decision Processes
by: Murthy, Yashaswini, et al.
Published: (2023)
by: Murthy, Yashaswini, et al.
Published: (2023)
OCMDP: Observation-Constrained Markov Decision Process
by: Wang, Taiyi, et al.
Published: (2024)
by: Wang, Taiyi, et al.
Published: (2024)
Success Conditioning as Policy Improvement: The Optimization Problem Solved by Imitating Success
by: Russo, Daniel
Published: (2026)
by: Russo, Daniel
Published: (2026)
RL in Latent MDPs is Tractable: Online Guarantees via Off-Policy Evaluation
by: Kwon, Jeongyeol, et al.
Published: (2024)
by: Kwon, Jeongyeol, et al.
Published: (2024)
1-2-3-Go! Policy Synthesis for Parameterized Markov Decision Processes via Decision-Tree Learning and Generalization
by: Azeem, Muqsit, et al.
Published: (2024)
by: Azeem, Muqsit, et al.
Published: (2024)
Offline Reinforcement Learning and Sequence Modeling for Downlink Link Adaptation
by: Peri, Samuele, et al.
Published: (2024)
by: Peri, Samuele, et al.
Published: (2024)
Advantage-Guided Diffusion for Model-Based Reinforcement Learning
by: Foffano, Daniele, et al.
Published: (2026)
by: Foffano, Daniele, et al.
Published: (2026)
Improved Monte Carlo Planning via Causal Disentanglement for Structurally-Decomposed Markov Decision Processes
by: Liu, Larkin, et al.
Published: (2024)
by: Liu, Larkin, et al.
Published: (2024)
Optimal Centered Active Excitation in Linear System Identification
by: Ito, Kaito, et al.
Published: (2026)
by: Ito, Kaito, et al.
Published: (2026)
GenSafe: A Generalizable Safety Enhancer for Safe Reinforcement Learning Algorithms Based on Reduced Order Markov Decision Process Model
by: Zhou, Zhehua, et al.
Published: (2024)
by: Zhou, Zhehua, et al.
Published: (2024)
Efficient and Sharp Off-Policy Evaluation in Robust Markov Decision Processes
by: Bennett, Andrew, et al.
Published: (2024)
by: Bennett, Andrew, et al.
Published: (2024)
Worst-Case Regret Bounds for Exploration via Randomized Value Functions
by: Russo, Daniel
Published: (2019)
by: Russo, Daniel
Published: (2019)
Frugal Actor-Critic: Sample Efficient Off-Policy Deep Reinforcement Learning Using Unique Experiences
by: Singh, Nikhil Kumar, et al.
Published: (2024)
by: Singh, Nikhil Kumar, et al.
Published: (2024)
Explainable Representation of Finite-Memory Policies for POMDPs using Decision Trees
by: Azeem, Muqsit, et al.
Published: (2024)
by: Azeem, Muqsit, et al.
Published: (2024)
Conformal Safety Shielding for Imperfect-Perception Agents
by: Scarbro, William, et al.
Published: (2025)
by: Scarbro, William, et al.
Published: (2025)
On the Global Convergence of Policy Gradient in Average Reward Markov Decision Processes
by: Kumar, Navdeep, et al.
Published: (2024)
by: Kumar, Navdeep, et al.
Published: (2024)
Intermittently Observable Markov Decision Processes
by: Chen, Gongpu, et al.
Published: (2023)
by: Chen, Gongpu, et al.
Published: (2023)
Physics-Informed State Space Models for Reliable Solar Irradiance Forecasting in Off-Grid Systems
by: Abdullah, Mohammed Ezzaldin Babiker
Published: (2026)
by: Abdullah, Mohammed Ezzaldin Babiker
Published: (2026)
Spatiotemporal Decision Transformer for Traffic Coordination
by: Su, Haoran, et al.
Published: (2026)
by: Su, Haoran, et al.
Published: (2026)
Zono-Conformal Prediction: Zonotope-Based Uncertainty Quantification for Regression and Classification Tasks
by: Lützow, Laura, et al.
Published: (2025)
by: Lützow, Laura, et al.
Published: (2025)
Thermodynamic Liquid Manifold Networks: Physics-Bounded Deep Learning for Solar Forecasting in Autonomous Off-Grid Microgrids
by: Abdullah, Mohammed Ezzaldin Babiker
Published: (2026)
by: Abdullah, Mohammed Ezzaldin Babiker
Published: (2026)
Causal Temporal Reasoning for Markov Decision Processes
by: Kazemi, Milad, et al.
Published: (2022)
by: Kazemi, Milad, et al.
Published: (2022)
Benchmarking Domain Adaptation for Chemical Processes on the Tennessee Eastman Process
by: Montesuma, Eduardo Fernandes, et al.
Published: (2023)
by: Montesuma, Eduardo Fernandes, et al.
Published: (2023)
Bayesian Learning of Optimal Policies in Markov Decision Processes with Countably Infinite State-Space
by: Adler, Saghar, et al.
Published: (2023)
by: Adler, Saghar, et al.
Published: (2023)
Zeroth-Order Actor-Critic: An Evolutionary Framework for Sequential Decision Problems
by: Lei, Yuheng, et al.
Published: (2022)
by: Lei, Yuheng, et al.
Published: (2022)
An Optimal Policy for Learning Controllable Dynamics by Exploration
by: Loxley, Peter N.
Published: (2025)
by: Loxley, Peter N.
Published: (2025)
Policy Optimization Algorithms in a Unified Framework
by: Wu, Shuang
Published: (2025)
by: Wu, Shuang
Published: (2025)
Certifiably Robust Policies for Uncertain Parametric Environments
by: Schnitzer, Yannik, et al.
Published: (2024)
by: Schnitzer, Yannik, et al.
Published: (2024)
Multi-Objective Reinforcement Learning for Tactical Decision Making for Trucks in Highway Traffic
by: Pathare, Deepthi, et al.
Published: (2026)
by: Pathare, Deepthi, et al.
Published: (2026)
Interval Markov Decision Processes with Continuous Action-Spaces
by: Delimpaltadakis, Giannis, et al.
Published: (2022)
by: Delimpaltadakis, Giannis, et al.
Published: (2022)
Policy Testing in Markov Decision Processes
by: Ariu, Kaito, et al.
Published: (2025)
by: Ariu, Kaito, et al.
Published: (2025)
InfraLib: Enabling Reinforcement Learning and Decision-Making for Large-Scale Infrastructure Management
by: Thangeda, Pranay, et al.
Published: (2024)
by: Thangeda, Pranay, et al.
Published: (2024)
Stabilizing Policy Gradient Methods via Reward Profiling
by: Ahmed, Shihab, et al.
Published: (2025)
by: Ahmed, Shihab, et al.
Published: (2025)
Deep Learning for Sequential Decision Making under Uncertainty: Foundations, Frameworks, and Frontiers
by: Buyuktahtakin, I. Esra
Published: (2026)
by: Buyuktahtakin, I. Esra
Published: (2026)
Verification of Neural Reachable Tubes via Scenario Optimization and Conformal Prediction
by: Lin, Albert, et al.
Published: (2023)
by: Lin, Albert, et al.
Published: (2023)
Bayesian Optimization of Process Parameters of a Sensor-Based Sorting System using Gaussian Processes as Surrogate Models
by: Kronenwett, Felix, et al.
Published: (2025)
by: Kronenwett, Felix, et al.
Published: (2025)
Optimizing Inventory Routing: A Decision-Focused Learning Approach using Neural Networks
by: Islam, MD Shafikul, et al.
Published: (2023)
by: Islam, MD Shafikul, et al.
Published: (2023)
Similar Items
-
Adversarial Diffusion for Robust Reinforcement Learning
by: Foffano, Daniele, et al.
Published: (2025) -
Receding-Horizon Control via Drifting Models
by: Foffano, Daniele, et al.
Published: (2026) -
An Offline Risk-aware Policy Selection Method for Bayesian Markov Decision Processes
by: Angelotti, Giorgio, et al.
Published: (2021) -
On the Convergence of Modified Policy Iteration in Risk Sensitive Exponential Cost Markov Decision Processes
by: Murthy, Yashaswini, et al.
Published: (2023) -
OCMDP: Observation-Constrained Markov Decision Process
by: Wang, Taiyi, et al.
Published: (2024)