Adaptive Shielding for Safe Reinforcement Learning under Hidden-Parameter Dynamics Shifts
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kwon, Minjae, Ingebrand, Tyler, Topcu, Ufuk, Feng, Lu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Zero-Shot Reinforcement Learning via Function Encoders
von: Ingebrand, Tyler, et al.
Veröffentlicht: (2024)
von: Ingebrand, Tyler, et al.
Veröffentlicht: (2024)
Function Encoders: A Principled Approach to Transfer Learning in Hilbert Spaces
von: Ingebrand, Tyler, et al.
Veröffentlicht: (2025)
von: Ingebrand, Tyler, et al.
Veröffentlicht: (2025)
Safe Reinforcement Learning in Black-Box Environments via Adaptive Shielding
von: Bethell, Daniel, et al.
Veröffentlicht: (2024)
von: Bethell, Daniel, et al.
Veröffentlicht: (2024)
Dynamic Model Predictive Shielding for Provably Safe Reinforcement Learning
von: Banerjee, Arko, et al.
Veröffentlicht: (2024)
von: Banerjee, Arko, et al.
Veröffentlicht: (2024)
Robust Multi-Agent Reinforcement Learning for Small UAS Separation Assurance under GPS Degradation and Spoofing
von: Zongo, Alex, et al.
Veröffentlicht: (2026)
von: Zongo, Alex, et al.
Veröffentlicht: (2026)
Probabilistic Shielding for Safe Reinforcement Learning
von: Court, Edwin Hamel-De le, et al.
Veröffentlicht: (2025)
von: Court, Edwin Hamel-De le, et al.
Veröffentlicht: (2025)
Adaptive Reward Design for Reinforcement Learning
von: Kwon, Minjae, et al.
Veröffentlicht: (2024)
von: Kwon, Minjae, et al.
Veröffentlicht: (2024)
Robust Probabilistic Shielding for Safe Offline Reinforcement Learning
von: Galesloot, Maris F. L., et al.
Veröffentlicht: (2026)
von: Galesloot, Maris F. L., et al.
Veröffentlicht: (2026)
Safe Reinforcement Learning via Recovery-based Shielding with Gaussian Process Dynamics Models
von: Goodall, Alexander W., et al.
Veröffentlicht: (2026)
von: Goodall, Alexander W., et al.
Veröffentlicht: (2026)
Basis-to-Basis Operator Learning Using Function Encoders
von: Ingebrand, Tyler, et al.
Veröffentlicht: (2024)
von: Ingebrand, Tyler, et al.
Veröffentlicht: (2024)
Neural Port-Hamiltonian Differential Algebraic Equations for Compositional Learning of Electrical Networks
von: Neary, Cyrus, et al.
Veröffentlicht: (2024)
von: Neary, Cyrus, et al.
Veröffentlicht: (2024)
Robust Shielding for Safe Reinforcement Learning
von: Court, Edwin Hamel-De le, et al.
Veröffentlicht: (2026)
von: Court, Edwin Hamel-De le, et al.
Veröffentlicht: (2026)
Approximate Model-Based Shielding for Safe Reinforcement Learning
von: Goodall, Alexander W., et al.
Veröffentlicht: (2023)
von: Goodall, Alexander W., et al.
Veröffentlicht: (2023)
Reduce, Reuse, Recycle: Categories for Compositional Reinforcement Learning
von: Bakirtzis, Georgios, et al.
Veröffentlicht: (2024)
von: Bakirtzis, Georgios, et al.
Veröffentlicht: (2024)
Using Large Language Models to Automate and Expedite Reinforcement Learning with Reward Machine
von: Alsadat, Shayan Meshkat, et al.
Veröffentlicht: (2024)
von: Alsadat, Shayan Meshkat, et al.
Veröffentlicht: (2024)
Sparsity-based Safety Conservatism for Constrained Offline Reinforcement Learning
von: Cho, Minjae, et al.
Veröffentlicht: (2024)
von: Cho, Minjae, et al.
Veröffentlicht: (2024)
Hierarchical Meta-Reinforcement Learning via Automated Macro-Action Discovery
von: Cho, Minjae, et al.
Veröffentlicht: (2024)
von: Cho, Minjae, et al.
Veröffentlicht: (2024)
Decoupled Guidance Diffusion for Adaptive Offline Safe Reinforcement Learning
von: Chen, Rufeng, et al.
Veröffentlicht: (2026)
von: Chen, Rufeng, et al.
Veröffentlicht: (2026)
Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning
von: Chemingui, Yassine, et al.
Veröffentlicht: (2024)
von: Chemingui, Yassine, et al.
Veröffentlicht: (2024)
Contraction Actor-Critic: Contraction Metric-Guided Reinforcement Learning for Robust Path Tracking
von: Cho, Minjae, et al.
Veröffentlicht: (2025)
von: Cho, Minjae, et al.
Veröffentlicht: (2025)
Verified Safe Reinforcement Learning for Neural Network Dynamic Models
von: Wu, Junlin, et al.
Veröffentlicht: (2024)
von: Wu, Junlin, et al.
Veröffentlicht: (2024)
Do LLMs Strategically Reveal, Conceal, and Infer Information? A Theoretical and Empirical Analysis in The Chameleon Game
von: Karabag, Mustafa O., et al.
Veröffentlicht: (2025)
von: Karabag, Mustafa O., et al.
Veröffentlicht: (2025)
Composite Flow Matching for Reinforcement Learning with Shifted-Dynamics Data
von: Kong, Lingkai, et al.
Veröffentlicht: (2025)
von: Kong, Lingkai, et al.
Veröffentlicht: (2025)
Adaptive Primal-Dual Method for Safe Reinforcement Learning
von: Chen, Weiqin, et al.
Veröffentlicht: (2024)
von: Chen, Weiqin, et al.
Veröffentlicht: (2024)
Enhancing Instruction Following of LLMs via Activation Steering with Dynamic Rejection
von: Kang, Minjae, et al.
Veröffentlicht: (2026)
von: Kang, Minjae, et al.
Veröffentlicht: (2026)
Spectral Invariant Learning for Dynamic Graphs under Distribution Shifts
von: Zhang, Zeyang, et al.
Veröffentlicht: (2024)
von: Zhang, Zeyang, et al.
Veröffentlicht: (2024)
Learning Hidden Subgoals under Temporal Ordering Constraints in Reinforcement Learning
von: Xu, Duo, et al.
Veröffentlicht: (2024)
von: Xu, Duo, et al.
Veröffentlicht: (2024)
A Flow Matching Algorithm for Many-Shot Adaptation to Unseen Distributions
von: Ingebrand, Tyler, et al.
Veröffentlicht: (2026)
von: Ingebrand, Tyler, et al.
Veröffentlicht: (2026)
JaxWildfire: A GPU-Accelerated Wildfire Simulator for Reinforcement Learning
von: Çakır, Ufuk, et al.
Veröffentlicht: (2025)
von: Çakır, Ufuk, et al.
Veröffentlicht: (2025)
Oracle-Guided Soft Shielding for Safe Move Prediction in Chess
von: Rajendran, Prajit T, et al.
Veröffentlicht: (2026)
von: Rajendran, Prajit T, et al.
Veröffentlicht: (2026)
Safe In-Context Reinforcement Learning
von: Moeini, Amir, et al.
Veröffentlicht: (2025)
von: Moeini, Amir, et al.
Veröffentlicht: (2025)
ProSh: Probabilistic Shielding for Model-free Reinforcement Learning
von: Court, Edwin Hamel-De le, et al.
Veröffentlicht: (2025)
von: Court, Edwin Hamel-De le, et al.
Veröffentlicht: (2025)
Reinforcement Learning by Guided Safe Exploration
von: Yang, Qisong, et al.
Veröffentlicht: (2023)
von: Yang, Qisong, et al.
Veröffentlicht: (2023)
Online Foundation Model Selection in Robotics
von: Li, Po-han, et al.
Veröffentlicht: (2024)
von: Li, Po-han, et al.
Veröffentlicht: (2024)
SafeAdapt: Provably Safe Policy Updates in Deep Reinforcement Learning
von: Anisimov, Maksim, et al.
Veröffentlicht: (2026)
von: Anisimov, Maksim, et al.
Veröffentlicht: (2026)
Safe and Balanced: A Framework for Constrained Multi-Objective Reinforcement Learning
von: Gu, Shangding, et al.
Veröffentlicht: (2024)
von: Gu, Shangding, et al.
Veröffentlicht: (2024)
Hierarchical Reinforcement Learning with Runtime Safety Shielding for Power Grid Operation
von: Malik, Gitesh
Veröffentlicht: (2026)
von: Malik, Gitesh
Veröffentlicht: (2026)
Online Optimization for Offline Safe Reinforcement Learning
von: Chemingui, Yassine, et al.
Veröffentlicht: (2025)
von: Chemingui, Yassine, et al.
Veröffentlicht: (2025)
Safe Flow Q-Learning: Offline Safe Reinforcement Learning with Reachability-Based Flow Policies
von: Tayal, Mumuksh, et al.
Veröffentlicht: (2026)
von: Tayal, Mumuksh, et al.
Veröffentlicht: (2026)
Implicit Safe Set Algorithm for Provably Safe Reinforcement Learning
von: Zhao, Weiye, et al.
Veröffentlicht: (2024)
von: Zhao, Weiye, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Zero-Shot Reinforcement Learning via Function Encoders
von: Ingebrand, Tyler, et al.
Veröffentlicht: (2024) -
Function Encoders: A Principled Approach to Transfer Learning in Hilbert Spaces
von: Ingebrand, Tyler, et al.
Veröffentlicht: (2025) -
Safe Reinforcement Learning in Black-Box Environments via Adaptive Shielding
von: Bethell, Daniel, et al.
Veröffentlicht: (2024) -
Dynamic Model Predictive Shielding for Provably Safe Reinforcement Learning
von: Banerjee, Arko, et al.
Veröffentlicht: (2024) -
Robust Multi-Agent Reinforcement Learning for Small UAS Separation Assurance under GPS Degradation and Spoofing
von: Zongo, Alex, et al.
Veröffentlicht: (2026)