Certifiably Robust Policies for Uncertain Parametric Environments
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Schnitzer, Yannik, Abate, Alessandro, Parker, David |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Efficient Solution and Learning of Robust Factored MDPs
von: Schnitzer, Yannik, et al.
Veröffentlicht: (2025)
von: Schnitzer, Yannik, et al.
Veröffentlicht: (2025)
Robust Parameter Learning for Uncertain MDPs
von: Schnitzer, Yannik, et al.
Veröffentlicht: (2026)
von: Schnitzer, Yannik, et al.
Veröffentlicht: (2026)
Probabilistic Performance Guarantees for Multi-Task Reinforcement Learning
von: Schnitzer, Yannik, et al.
Veröffentlicht: (2026)
von: Schnitzer, Yannik, et al.
Veröffentlicht: (2026)
Certified Approximate Reachability (CARe): Formal Error Bounds on Deep Learning of Reachable Sets
von: Solanki, Prashant, et al.
Veröffentlicht: (2025)
von: Solanki, Prashant, et al.
Veröffentlicht: (2025)
Neural Proofs for Sound Verification and Control of Complex Systems
von: Abate, Alessandro
Veröffentlicht: (2025)
von: Abate, Alessandro
Veröffentlicht: (2025)
Networked Communication for Decentralised Agents in Mean-Field Games
von: Benjamin, Patrick, et al.
Veröffentlicht: (2023)
von: Benjamin, Patrick, et al.
Veröffentlicht: (2023)
Formal Synthesis of Certifiably Robust Neural Lyapunov-Barrier Certificates
von: Wang, Chengxiao, et al.
Veröffentlicht: (2026)
von: Wang, Chengxiao, et al.
Veröffentlicht: (2026)
Networked Communication for Mean-Field Games with Function Approximation and Empirical Mean-Field Estimation
von: Benjamin, Patrick, et al.
Veröffentlicht: (2024)
von: Benjamin, Patrick, et al.
Veröffentlicht: (2024)
Learning Robust Policies for Uncertain Parametric Markov Decision Processes
von: Rickard, Luke, et al.
Veröffentlicht: (2023)
von: Rickard, Luke, et al.
Veröffentlicht: (2023)
Runtime-Certified Bounded-Error Quantized Attention
von: Calver, Dean
Veröffentlicht: (2026)
von: Calver, Dean
Veröffentlicht: (2026)
Lyapunov-Certified Direct Switching Theory for Q-Learning
von: Lee, Donghwan
Veröffentlicht: (2026)
von: Lee, Donghwan
Veröffentlicht: (2026)
Distributed Risk-Sensitive Safety Filters for Uncertain Discrete-Time Systems
von: Lederer, Armin, et al.
Veröffentlicht: (2025)
von: Lederer, Armin, et al.
Veröffentlicht: (2025)
Certified Neural Approximations of Nonlinear Dynamics
von: Mathiesen, Frederik Baymler, et al.
Veröffentlicht: (2025)
von: Mathiesen, Frederik Baymler, et al.
Veröffentlicht: (2025)
Certifiable Safe RLHF: Fixed-Penalty Constraint Optimization for Safer Language Models
von: Pandit, Kartik, et al.
Veröffentlicht: (2025)
von: Pandit, Kartik, et al.
Veröffentlicht: (2025)
Reinforcement Learning with Partial Parametric Model Knowledge
von: Wang, Shuyuan, et al.
Veröffentlicht: (2023)
von: Wang, Shuyuan, et al.
Veröffentlicht: (2023)
Certified Training with Branch-and-Bound for Lyapunov-stable Neural Control
von: Shi, Zhouxing, et al.
Veröffentlicht: (2024)
von: Shi, Zhouxing, et al.
Veröffentlicht: (2024)
ReACT: Reinforcement Learning for Controller Parametrization using B-Spline Geometries
von: Rudolf, Thomas, et al.
Veröffentlicht: (2024)
von: Rudolf, Thomas, et al.
Veröffentlicht: (2024)
Parallel Differentiable Reachability for Learning and Planning with Certified Neural Dynamics and Controllers
von: Shen, Keyi, et al.
Veröffentlicht: (2026)
von: Shen, Keyi, et al.
Veröffentlicht: (2026)
Efficient Policy Optimization in Robust Constrained MDPs with Iteration Complexity Guarantees
von: Ganguly, Sourav, et al.
Veröffentlicht: (2025)
von: Ganguly, Sourav, et al.
Veröffentlicht: (2025)
One Filter to Deploy Them All: Robust Safety for Quadrupedal Navigation in Unknown Environments
von: Lin, Albert, et al.
Veröffentlicht: (2024)
von: Lin, Albert, et al.
Veröffentlicht: (2024)
GUIDEd Agents: Enhancing Navigation Policies through Task-Specific Uncertainty Abstraction in Localization-Limited Environments
von: Puthumanaillam, Gokul, et al.
Veröffentlicht: (2024)
von: Puthumanaillam, Gokul, et al.
Veröffentlicht: (2024)
Action Mapping for Reinforcement Learning in Continuous Environments with Constraints
von: Theile, Mirco, et al.
Veröffentlicht: (2024)
von: Theile, Mirco, et al.
Veröffentlicht: (2024)
An Optimal Policy for Learning Controllable Dynamics by Exploration
von: Loxley, Peter N.
Veröffentlicht: (2025)
von: Loxley, Peter N.
Veröffentlicht: (2025)
Policy Optimization Algorithms in a Unified Framework
von: Wu, Shuang
Veröffentlicht: (2025)
von: Wu, Shuang
Veröffentlicht: (2025)
TIP-Search: Time-Predictable Inference Scheduling for Market Prediction under Uncertain Load
von: Wang, Xibai
Veröffentlicht: (2025)
von: Wang, Xibai
Veröffentlicht: (2025)
Stabilizing Policy Gradient Methods via Reward Profiling
von: Ahmed, Shihab, et al.
Veröffentlicht: (2025)
von: Ahmed, Shihab, et al.
Veröffentlicht: (2025)
Conformal Off-Policy Evaluation in Markov Decision Processes
von: Foffano, Daniele, et al.
Veröffentlicht: (2023)
von: Foffano, Daniele, et al.
Veröffentlicht: (2023)
Mining--Gym: A Configurable RL Benchmarking Environment for Truck Dispatch Scheduling
von: Banerjee, Chayan, et al.
Veröffentlicht: (2025)
von: Banerjee, Chayan, et al.
Veröffentlicht: (2025)
A General Approach of Automated Environment Design for Learning the Optimal Power Flow
von: Wolgast, Thomas, et al.
Veröffentlicht: (2025)
von: Wolgast, Thomas, et al.
Veröffentlicht: (2025)
Success Conditioning as Policy Improvement: The Optimization Problem Solved by Imitating Success
von: Russo, Daniel
Veröffentlicht: (2026)
von: Russo, Daniel
Veröffentlicht: (2026)
RL-Based Method for Benchmarking the Adversarial Resilience and Robustness of Deep Reinforcement Learning Policies
von: Behzadan, Vahid, et al.
Veröffentlicht: (2019)
von: Behzadan, Vahid, et al.
Veröffentlicht: (2019)
Differentiable Distributionally Robust Optimization Layers
von: Ma, Xutao, et al.
Veröffentlicht: (2024)
von: Ma, Xutao, et al.
Veröffentlicht: (2024)
Control Policy Correction Framework for Reinforcement Learning-based Energy Arbitrage Strategies
von: Madahi, Seyed Soroush Karimi, et al.
Veröffentlicht: (2024)
von: Madahi, Seyed Soroush Karimi, et al.
Veröffentlicht: (2024)
RL in Latent MDPs is Tractable: Online Guarantees via Off-Policy Evaluation
von: Kwon, Jeongyeol, et al.
Veröffentlicht: (2024)
von: Kwon, Jeongyeol, et al.
Veröffentlicht: (2024)
Analyzing Generalization in Policy Networks: A Case Study with the Double-Integrator System
von: Zhang, Ruining, et al.
Veröffentlicht: (2023)
von: Zhang, Ruining, et al.
Veröffentlicht: (2023)
Joint Optimization of Multi-Objective Reinforcement Learning with Policy Gradient Based Algorithm
von: Bai, Qinbo, et al.
Veröffentlicht: (2021)
von: Bai, Qinbo, et al.
Veröffentlicht: (2021)
An Offline Risk-aware Policy Selection Method for Bayesian Markov Decision Processes
von: Angelotti, Giorgio, et al.
Veröffentlicht: (2021)
von: Angelotti, Giorgio, et al.
Veröffentlicht: (2021)
From Explainability to Interpretability: Interpretable Policies in Reinforcement Learning Via Model Explanation
von: Li, Peilang, et al.
Veröffentlicht: (2025)
von: Li, Peilang, et al.
Veröffentlicht: (2025)
CORL: Reinforcement Learning of MILP Policies Solved via Branch and Bound
von: Anand, Akhil S, et al.
Veröffentlicht: (2025)
von: Anand, Akhil S, et al.
Veröffentlicht: (2025)
Revisiting LQR Control from the Perspective of Receding-Horizon Policy Gradient
von: Zhang, Xiangyuan, et al.
Veröffentlicht: (2023)
von: Zhang, Xiangyuan, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Efficient Solution and Learning of Robust Factored MDPs
von: Schnitzer, Yannik, et al.
Veröffentlicht: (2025) -
Robust Parameter Learning for Uncertain MDPs
von: Schnitzer, Yannik, et al.
Veröffentlicht: (2026) -
Probabilistic Performance Guarantees for Multi-Task Reinforcement Learning
von: Schnitzer, Yannik, et al.
Veröffentlicht: (2026) -
Certified Approximate Reachability (CARe): Formal Error Bounds on Deep Learning of Reachable Sets
von: Solanki, Prashant, et al.
Veröffentlicht: (2025) -
Neural Proofs for Sound Verification and Control of Complex Systems
von: Abate, Alessandro
Veröffentlicht: (2025)