Adaptive GR(1) Specification Repair for Liveness-Preserving Shielding in Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Georgescu, Tiberiu-Andrei, Goodall, Alexander W., Alrajeh, Dalal, Belardinelli, Francesco, Uchitel, Sebastian |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Approximate Model-Based Shielding for Safe Reinforcement Learning
von: Goodall, Alexander W., et al.
Veröffentlicht: (2023)
von: Goodall, Alexander W., et al.
Veröffentlicht: (2023)
Safe Reinforcement Learning via Recovery-based Shielding with Gaussian Process Dynamics Models
von: Goodall, Alexander W., et al.
Veröffentlicht: (2026)
von: Goodall, Alexander W., et al.
Veröffentlicht: (2026)
Probabilistic Shielding for Safe Reinforcement Learning
von: Court, Edwin Hamel-De le, et al.
Veröffentlicht: (2025)
von: Court, Edwin Hamel-De le, et al.
Veröffentlicht: (2025)
Leveraging Approximate Model-based Shielding for Probabilistic Safety Guarantees in Continuous Environments
von: Goodall, Alexander W., et al.
Veröffentlicht: (2024)
von: Goodall, Alexander W., et al.
Veröffentlicht: (2024)
Behaviour Policy Optimization: Provably Lower Variance Return Estimates for Off-Policy Reinforcement Learning
von: Goodall, Alexander W., et al.
Veröffentlicht: (2025)
von: Goodall, Alexander W., et al.
Veröffentlicht: (2025)
ProSh: Probabilistic Shielding for Model-free Reinforcement Learning
von: Court, Edwin Hamel-De le, et al.
Veröffentlicht: (2025)
von: Court, Edwin Hamel-De le, et al.
Veröffentlicht: (2025)
Robust Shielding for Safe Reinforcement Learning
von: Court, Edwin Hamel-De le, et al.
Veröffentlicht: (2026)
von: Court, Edwin Hamel-De le, et al.
Veröffentlicht: (2026)
Expressive Temporal Specifications for Reward Monitoring
von: Adalat, Omar, et al.
Veröffentlicht: (2025)
von: Adalat, Omar, et al.
Veröffentlicht: (2025)
SafeAdapt: Provably Safe Policy Updates in Deep Reinforcement Learning
von: Anisimov, Maksim, et al.
Veröffentlicht: (2026)
von: Anisimov, Maksim, et al.
Veröffentlicht: (2026)
Scaling GR(1) Synthesis via a Compositional Framework for LTL Discrete Event Control
von: Gagliardi, Hernan, et al.
Veröffentlicht: (2025)
von: Gagliardi, Hernan, et al.
Veröffentlicht: (2025)
Safe Reinforcement Learning in Black-Box Environments via Adaptive Shielding
von: Bethell, Daniel, et al.
Veröffentlicht: (2024)
von: Bethell, Daniel, et al.
Veröffentlicht: (2024)
A Logic of General Attention Using Edge-Conditioned Event Models (Extended Version)
von: Belardinelli, Gaia, et al.
Veröffentlicht: (2025)
von: Belardinelli, Gaia, et al.
Veröffentlicht: (2025)
Adaptive Shielding for Safe Reinforcement Learning under Hidden-Parameter Dynamics Shifts
von: Kwon, Minjae, et al.
Veröffentlicht: (2025)
von: Kwon, Minjae, et al.
Veröffentlicht: (2025)
Gradient Boosting Reinforcement Learning
von: Fuhrer, Benjamin, et al.
Veröffentlicht: (2024)
von: Fuhrer, Benjamin, et al.
Veröffentlicht: (2024)
Explainable Reinforcement Learning for Formula One Race Strategy
von: Thomas, Devin, et al.
Veröffentlicht: (2025)
von: Thomas, Devin, et al.
Veröffentlicht: (2025)
Gaze-based intention estimation: principles, methodologies, and applications in HRI
von: Belardinelli, Anna
Veröffentlicht: (2023)
von: Belardinelli, Anna
Veröffentlicht: (2023)
The Geometry of Grokking: Norm Minimization on the Zero-Loss Manifold
von: Musat, Tiberiu
Veröffentlicht: (2025)
von: Musat, Tiberiu
Veröffentlicht: (2025)
Robust Probabilistic Shielding for Safe Offline Reinforcement Learning
von: Galesloot, Maris F. L., et al.
Veröffentlicht: (2026)
von: Galesloot, Maris F. L., et al.
Veröffentlicht: (2026)
Inventory problems and the parametric measure $m_λ$
von: Georgescu, Irina
Veröffentlicht: (2024)
von: Georgescu, Irina
Veröffentlicht: (2024)
Expressive Reward Synthesis with the Runtime Monitoring Language
von: Donnelly, Daniel, et al.
Veröffentlicht: (2025)
von: Donnelly, Daniel, et al.
Veröffentlicht: (2025)
GR-Agent: Adaptive Graph Reasoning Agent under Incomplete Knowledge
von: Zhou, Dongzhuoran, et al.
Veröffentlicht: (2025)
von: Zhou, Dongzhuoran, et al.
Veröffentlicht: (2025)
Dynamic Model Predictive Shielding for Provably Safe Reinforcement Learning
von: Banerjee, Arko, et al.
Veröffentlicht: (2024)
von: Banerjee, Arko, et al.
Veröffentlicht: (2024)
Silent Neuron Theory and Plasticity Preservation for Deep Reinforcement Learning in Adaptive Video Streaming
von: He, Zhiqiang, et al.
Veröffentlicht: (2025)
von: He, Zhiqiang, et al.
Veröffentlicht: (2025)
Entropy-Preserving Reinforcement Learning
von: Petrenko, Aleksei, et al.
Veröffentlicht: (2026)
von: Petrenko, Aleksei, et al.
Veröffentlicht: (2026)
Measuring Goal-Directedness
von: MacDermott, Matt, et al.
Veröffentlicht: (2024)
von: MacDermott, Matt, et al.
Veröffentlicht: (2024)
Planning Task Shielding: Detecting and Repairing Flaws in Planning Tasks through Turning them Unsolvable
von: Pozanco, Alberto, et al.
Veröffentlicht: (2026)
von: Pozanco, Alberto, et al.
Veröffentlicht: (2026)
Hierarchical Reinforcement Learning with Runtime Safety Shielding for Power Grid Operation
von: Malik, Gitesh
Veröffentlicht: (2026)
von: Malik, Gitesh
Veröffentlicht: (2026)
Step-wise Adaptive Integration of Supervised Fine-tuning and Reinforcement Learning for Task-Specific LLMs
von: Chen, Jack, et al.
Veröffentlicht: (2025)
von: Chen, Jack, et al.
Veröffentlicht: (2025)
The Reasons that Agents Act: Intention and Instrumental Goals
von: Ward, Francis Rhys, et al.
Veröffentlicht: (2024)
von: Ward, Francis Rhys, et al.
Veröffentlicht: (2024)
Predictive Safety Shield for Dyna-Q Reinforcement Learning
von: Pin, Jin, et al.
Veröffentlicht: (2025)
von: Pin, Jin, et al.
Veröffentlicht: (2025)
Compositional Shielding and Reinforcement Learning for Multi-Agent Systems
von: Brorholt, Asger Horn, et al.
Veröffentlicht: (2024)
von: Brorholt, Asger Horn, et al.
Veröffentlicht: (2024)
Constructive Symbolic Reinforcement Learning via Intuitionistic Logic and Goal-Chaining Inference
von: Patrascu, Andrei T.
Veröffentlicht: (2025)
von: Patrascu, Andrei T.
Veröffentlicht: (2025)
Synthesis of Safety Specifications for Probabilistic Systems
von: Ohlmann, Gaspard, et al.
Veröffentlicht: (2025)
von: Ohlmann, Gaspard, et al.
Veröffentlicht: (2025)
Automating the Refinement of Reinforcement Learning Specifications
von: Ambadkar, Tanmay, et al.
Veröffentlicht: (2025)
von: Ambadkar, Tanmay, et al.
Veröffentlicht: (2025)
Regret-Free Reinforcement Learning for LTL Specifications
von: Majumdar, Rupak, et al.
Veröffentlicht: (2024)
von: Majumdar, Rupak, et al.
Veröffentlicht: (2024)
Adaptive Preference Scaling for Reinforcement Learning with Human Feedback
von: Hong, Ilgee, et al.
Veröffentlicht: (2024)
von: Hong, Ilgee, et al.
Veröffentlicht: (2024)
LLM Access Shield: Domain-Specific LLM Framework for Privacy Policy Compliance
von: Wang, Yu, et al.
Veröffentlicht: (2025)
von: Wang, Yu, et al.
Veröffentlicht: (2025)
Interpretable by Design: Query-Specific Neural Modules for Explainable Reinforcement Learning
von: Zakershahrak, Mehrdad
Veröffentlicht: (2025)
von: Zakershahrak, Mehrdad
Veröffentlicht: (2025)
Reinforcement Learning with LTL and $ω$-Regular Objectives via Optimality-Preserving Translation to Average Rewards
von: Le, Xuan-Bach, et al.
Veröffentlicht: (2024)
von: Le, Xuan-Bach, et al.
Veröffentlicht: (2024)
Think Smart, Act SMARL! Analyzing Probabilistic Logic Shields for Multi-Agent Reinforcement Learning
von: Chatterji, Satchit, et al.
Veröffentlicht: (2024)
von: Chatterji, Satchit, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Approximate Model-Based Shielding for Safe Reinforcement Learning
von: Goodall, Alexander W., et al.
Veröffentlicht: (2023) -
Safe Reinforcement Learning via Recovery-based Shielding with Gaussian Process Dynamics Models
von: Goodall, Alexander W., et al.
Veröffentlicht: (2026) -
Probabilistic Shielding for Safe Reinforcement Learning
von: Court, Edwin Hamel-De le, et al.
Veröffentlicht: (2025) -
Leveraging Approximate Model-based Shielding for Probabilistic Safety Guarantees in Continuous Environments
von: Goodall, Alexander W., et al.
Veröffentlicht: (2024) -
Behaviour Policy Optimization: Provably Lower Variance Return Estimates for Off-Policy Reinforcement Learning
von: Goodall, Alexander W., et al.
Veröffentlicht: (2025)