Certifiably Robust Policies for Uncertain Parametric Environments
Fuente:
arXiv
Saved in:
| Main Authors: | Schnitzer, Yannik, Abate, Alessandro, Parker, David |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Efficient Solution and Learning of Robust Factored MDPs
by: Schnitzer, Yannik, et al.
Published: (2025)
by: Schnitzer, Yannik, et al.
Published: (2025)
Robust Parameter Learning for Uncertain MDPs
by: Schnitzer, Yannik, et al.
Published: (2026)
by: Schnitzer, Yannik, et al.
Published: (2026)
Probabilistic Performance Guarantees for Multi-Task Reinforcement Learning
by: Schnitzer, Yannik, et al.
Published: (2026)
by: Schnitzer, Yannik, et al.
Published: (2026)
Certified Approximate Reachability (CARe): Formal Error Bounds on Deep Learning of Reachable Sets
by: Solanki, Prashant, et al.
Published: (2025)
by: Solanki, Prashant, et al.
Published: (2025)
Neural Proofs for Sound Verification and Control of Complex Systems
by: Abate, Alessandro
Published: (2025)
by: Abate, Alessandro
Published: (2025)
Networked Communication for Decentralised Agents in Mean-Field Games
by: Benjamin, Patrick, et al.
Published: (2023)
by: Benjamin, Patrick, et al.
Published: (2023)
Formal Synthesis of Certifiably Robust Neural Lyapunov-Barrier Certificates
by: Wang, Chengxiao, et al.
Published: (2026)
by: Wang, Chengxiao, et al.
Published: (2026)
Networked Communication for Mean-Field Games with Function Approximation and Empirical Mean-Field Estimation
by: Benjamin, Patrick, et al.
Published: (2024)
by: Benjamin, Patrick, et al.
Published: (2024)
Learning Robust Policies for Uncertain Parametric Markov Decision Processes
by: Rickard, Luke, et al.
Published: (2023)
by: Rickard, Luke, et al.
Published: (2023)
Runtime-Certified Bounded-Error Quantized Attention
by: Calver, Dean
Published: (2026)
by: Calver, Dean
Published: (2026)
Lyapunov-Certified Direct Switching Theory for Q-Learning
by: Lee, Donghwan
Published: (2026)
by: Lee, Donghwan
Published: (2026)
Distributed Risk-Sensitive Safety Filters for Uncertain Discrete-Time Systems
by: Lederer, Armin, et al.
Published: (2025)
by: Lederer, Armin, et al.
Published: (2025)
Certified Neural Approximations of Nonlinear Dynamics
by: Mathiesen, Frederik Baymler, et al.
Published: (2025)
by: Mathiesen, Frederik Baymler, et al.
Published: (2025)
Certifiable Safe RLHF: Fixed-Penalty Constraint Optimization for Safer Language Models
by: Pandit, Kartik, et al.
Published: (2025)
by: Pandit, Kartik, et al.
Published: (2025)
Reinforcement Learning with Partial Parametric Model Knowledge
by: Wang, Shuyuan, et al.
Published: (2023)
by: Wang, Shuyuan, et al.
Published: (2023)
Certified Training with Branch-and-Bound for Lyapunov-stable Neural Control
by: Shi, Zhouxing, et al.
Published: (2024)
by: Shi, Zhouxing, et al.
Published: (2024)
ReACT: Reinforcement Learning for Controller Parametrization using B-Spline Geometries
by: Rudolf, Thomas, et al.
Published: (2024)
by: Rudolf, Thomas, et al.
Published: (2024)
Parallel Differentiable Reachability for Learning and Planning with Certified Neural Dynamics and Controllers
by: Shen, Keyi, et al.
Published: (2026)
by: Shen, Keyi, et al.
Published: (2026)
Efficient Policy Optimization in Robust Constrained MDPs with Iteration Complexity Guarantees
by: Ganguly, Sourav, et al.
Published: (2025)
by: Ganguly, Sourav, et al.
Published: (2025)
One Filter to Deploy Them All: Robust Safety for Quadrupedal Navigation in Unknown Environments
by: Lin, Albert, et al.
Published: (2024)
by: Lin, Albert, et al.
Published: (2024)
GUIDEd Agents: Enhancing Navigation Policies through Task-Specific Uncertainty Abstraction in Localization-Limited Environments
by: Puthumanaillam, Gokul, et al.
Published: (2024)
by: Puthumanaillam, Gokul, et al.
Published: (2024)
Action Mapping for Reinforcement Learning in Continuous Environments with Constraints
by: Theile, Mirco, et al.
Published: (2024)
by: Theile, Mirco, et al.
Published: (2024)
An Optimal Policy for Learning Controllable Dynamics by Exploration
by: Loxley, Peter N.
Published: (2025)
by: Loxley, Peter N.
Published: (2025)
Policy Optimization Algorithms in a Unified Framework
by: Wu, Shuang
Published: (2025)
by: Wu, Shuang
Published: (2025)
TIP-Search: Time-Predictable Inference Scheduling for Market Prediction under Uncertain Load
by: Wang, Xibai
Published: (2025)
by: Wang, Xibai
Published: (2025)
Stabilizing Policy Gradient Methods via Reward Profiling
by: Ahmed, Shihab, et al.
Published: (2025)
by: Ahmed, Shihab, et al.
Published: (2025)
Conformal Off-Policy Evaluation in Markov Decision Processes
by: Foffano, Daniele, et al.
Published: (2023)
by: Foffano, Daniele, et al.
Published: (2023)
Mining--Gym: A Configurable RL Benchmarking Environment for Truck Dispatch Scheduling
by: Banerjee, Chayan, et al.
Published: (2025)
by: Banerjee, Chayan, et al.
Published: (2025)
A General Approach of Automated Environment Design for Learning the Optimal Power Flow
by: Wolgast, Thomas, et al.
Published: (2025)
by: Wolgast, Thomas, et al.
Published: (2025)
Success Conditioning as Policy Improvement: The Optimization Problem Solved by Imitating Success
by: Russo, Daniel
Published: (2026)
by: Russo, Daniel
Published: (2026)
RL-Based Method for Benchmarking the Adversarial Resilience and Robustness of Deep Reinforcement Learning Policies
by: Behzadan, Vahid, et al.
Published: (2019)
by: Behzadan, Vahid, et al.
Published: (2019)
Differentiable Distributionally Robust Optimization Layers
by: Ma, Xutao, et al.
Published: (2024)
by: Ma, Xutao, et al.
Published: (2024)
Control Policy Correction Framework for Reinforcement Learning-based Energy Arbitrage Strategies
by: Madahi, Seyed Soroush Karimi, et al.
Published: (2024)
by: Madahi, Seyed Soroush Karimi, et al.
Published: (2024)
RL in Latent MDPs is Tractable: Online Guarantees via Off-Policy Evaluation
by: Kwon, Jeongyeol, et al.
Published: (2024)
by: Kwon, Jeongyeol, et al.
Published: (2024)
Analyzing Generalization in Policy Networks: A Case Study with the Double-Integrator System
by: Zhang, Ruining, et al.
Published: (2023)
by: Zhang, Ruining, et al.
Published: (2023)
Joint Optimization of Multi-Objective Reinforcement Learning with Policy Gradient Based Algorithm
by: Bai, Qinbo, et al.
Published: (2021)
by: Bai, Qinbo, et al.
Published: (2021)
An Offline Risk-aware Policy Selection Method for Bayesian Markov Decision Processes
by: Angelotti, Giorgio, et al.
Published: (2021)
by: Angelotti, Giorgio, et al.
Published: (2021)
From Explainability to Interpretability: Interpretable Policies in Reinforcement Learning Via Model Explanation
by: Li, Peilang, et al.
Published: (2025)
by: Li, Peilang, et al.
Published: (2025)
CORL: Reinforcement Learning of MILP Policies Solved via Branch and Bound
by: Anand, Akhil S, et al.
Published: (2025)
by: Anand, Akhil S, et al.
Published: (2025)
Revisiting LQR Control from the Perspective of Receding-Horizon Policy Gradient
by: Zhang, Xiangyuan, et al.
Published: (2023)
by: Zhang, Xiangyuan, et al.
Published: (2023)
Similar Items
-
Efficient Solution and Learning of Robust Factored MDPs
by: Schnitzer, Yannik, et al.
Published: (2025) -
Robust Parameter Learning for Uncertain MDPs
by: Schnitzer, Yannik, et al.
Published: (2026) -
Probabilistic Performance Guarantees for Multi-Task Reinforcement Learning
by: Schnitzer, Yannik, et al.
Published: (2026) -
Certified Approximate Reachability (CARe): Formal Error Bounds on Deep Learning of Reachable Sets
by: Solanki, Prashant, et al.
Published: (2025) -
Neural Proofs for Sound Verification and Control of Complex Systems
by: Abate, Alessandro
Published: (2025)