Counterexample-Guided Repair of Reinforcement Learning Systems Using Safety Critics
Fuente:
arXiv
Guardado en:
| Autores principales: | Boetius, David, Leue, Stefan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Verified SHAP: Provable Bounds for Exact Shapley Values of Neural Networks
por: Boetius, David, et al.
Publicado: (2026)
por: Boetius, David, et al.
Publicado: (2026)
Counterexample-Guided Interval Weakening
por: Andrew, Ben M., et al.
Publicado: (2026)
por: Andrew, Ben M., et al.
Publicado: (2026)
Progressive Safeguards for Safe and Model-Agnostic Reinforcement Learning
por: Omi, Nabil, et al.
Publicado: (2024)
por: Omi, Nabil, et al.
Publicado: (2024)
Skill Machines: Temporal Logic Skill Composition in Reinforcement Learning
por: Tasse, Geraud Nangue, et al.
Publicado: (2022)
por: Tasse, Geraud Nangue, et al.
Publicado: (2022)
Compositional Shielding and Reinforcement Learning for Multi-Agent Systems
por: Brorholt, Asger Horn, et al.
Publicado: (2024)
por: Brorholt, Asger Horn, et al.
Publicado: (2024)
A-IC3: Learning-Guided Adaptive Inductive Generalization for Hardware Model Checking
por: Zhou, Xiaofeng, et al.
Publicado: (2026)
por: Zhou, Xiaofeng, et al.
Publicado: (2026)
Conformal Signal Temporal Logic for Robust Reinforcement Learning Control: A Case Study
por: Beirami, Hani, et al.
Publicado: (2026)
por: Beirami, Hani, et al.
Publicado: (2026)
Logical GANs: Adversarial Learning through Ehrenfeucht Fraisse Games
por: Mannucci, Mirco A.
Publicado: (2025)
por: Mannucci, Mirco A.
Publicado: (2025)
Counterexample-Guided Abstraction Refinement for Generalized Graph Transformation Systems (Full Version)
por: König, Barbara, et al.
Publicado: (2025)
por: König, Barbara, et al.
Publicado: (2025)
Model Checking for Reinforcement Learning in Autonomous Driving: One Can Do More Than You Think!
por: Gu, Rong
Publicado: (2024)
por: Gu, Rong
Publicado: (2024)
Feasibility-Guided Fair Adaptive Offline Reinforcement Learning for Medicaid Care Management
por: Basu, Sanjay, et al.
Publicado: (2025)
por: Basu, Sanjay, et al.
Publicado: (2025)
Guiding LLM Temporal Logic Generation with Explicit Separation of Data and Control
por: Murphy, William, et al.
Publicado: (2024)
por: Murphy, William, et al.
Publicado: (2024)
Programmatic Reinforcement Learning: Navigating Gridworlds
por: Shabadi, Guruprerana, et al.
Publicado: (2024)
por: Shabadi, Guruprerana, et al.
Publicado: (2024)
Robust Shielding for Safe Reinforcement Learning
por: Court, Edwin Hamel-De le, et al.
Publicado: (2026)
por: Court, Edwin Hamel-De le, et al.
Publicado: (2026)
Inductive Generalization in Reinforcement Learning from Specifications
por: Subramanian, Vignesh, et al.
Publicado: (2024)
por: Subramanian, Vignesh, et al.
Publicado: (2024)
How (and when) can you fit examples to logic-based hypothesis classes over infinite structures?
por: Benedikt, Michael, et al.
Publicado: (2026)
por: Benedikt, Michael, et al.
Publicado: (2026)
From learnable objects to learnable random objects
por: Anderson, Aaron, et al.
Publicado: (2025)
por: Anderson, Aaron, et al.
Publicado: (2025)
Programs as Singularities
por: Murfet, Daniel, et al.
Publicado: (2025)
por: Murfet, Daniel, et al.
Publicado: (2025)
Bisimulation Learning
por: Abate, Alessandro, et al.
Publicado: (2024)
por: Abate, Alessandro, et al.
Publicado: (2024)
A Reinforcement Learning based Reset Policy for CDCL SAT Solvers
por: Li, Chunxiao, et al.
Publicado: (2024)
por: Li, Chunxiao, et al.
Publicado: (2024)
Deep Learning with Parametric Lenses
por: Cruttwell, Geoffrey S. H., et al.
Publicado: (2024)
por: Cruttwell, Geoffrey S. H., et al.
Publicado: (2024)
Value Function Initialization for Knowledge Transfer and Jump-start in Deep Reinforcement Learning
por: Mehimeh, Soumia
Publicado: (2025)
por: Mehimeh, Soumia
Publicado: (2025)
Output-decomposed Learning of Mealy Machines
por: Koenders, Rick, et al.
Publicado: (2024)
por: Koenders, Rick, et al.
Publicado: (2024)
Learning to Estimate System Specifications in Linear Temporal Logic using Transformers and Mamba
por: Işık, İlker, et al.
Publicado: (2024)
por: Işık, İlker, et al.
Publicado: (2024)
Scalable Interconnect Learning in Boolean Networks
por: Kresse, Fabian, et al.
Publicado: (2025)
por: Kresse, Fabian, et al.
Publicado: (2025)
Null Measurability at the Symmetrization Interface in VC Learning
por: Gupta, Dhruv
Publicado: (2026)
por: Gupta, Dhruv
Publicado: (2026)
On Improving Deep Active Learning with Formal Verification
por: Spiegelman, Jonathan, et al.
Publicado: (2025)
por: Spiegelman, Jonathan, et al.
Publicado: (2025)
Error-awareness Accelerates Active Automata Learning
por: Kruger, Loes, et al.
Publicado: (2026)
por: Kruger, Loes, et al.
Publicado: (2026)
Verifying the Generalization of Deep Learning to Out-of-Distribution Domains
por: Amir, Guy, et al.
Publicado: (2024)
por: Amir, Guy, et al.
Publicado: (2024)
Learning logic programs by finding minimal unsatisfiable subprograms
por: Cropper, Andrew, et al.
Publicado: (2024)
por: Cropper, Andrew, et al.
Publicado: (2024)
A General Framework for Property-Driven Machine Learning
por: Flinkow, Thomas, et al.
Publicado: (2025)
por: Flinkow, Thomas, et al.
Publicado: (2025)
Learning Representations Through Contrastive Neural Model Checking
por: Krsmanovic, Vladimir, et al.
Publicado: (2025)
por: Krsmanovic, Vladimir, et al.
Publicado: (2025)
Quantitative Linear Logic for Neuro-Symbolic Learning and Verification
por: Flinkow, Thomas, et al.
Publicado: (2026)
por: Flinkow, Thomas, et al.
Publicado: (2026)
A Hybrid Real-Time Framework for Efficient Fussell-Vesely Importance Evaluation Using Virtual Fault Trees and Graph Neural Networks
por: Xiao, Xingyu, et al.
Publicado: (2024)
por: Xiao, Xingyu, et al.
Publicado: (2024)
Do It for HER: First-Order Temporal Logic Reward Specification in Reinforcement Learning (Extended Version)
por: Olivieri, Pierriccardo, et al.
Publicado: (2026)
por: Olivieri, Pierriccardo, et al.
Publicado: (2026)
Learning Better Representations From Less Data For Propositional Satisfiability
por: Ghanem, Mohamed, et al.
Publicado: (2024)
por: Ghanem, Mohamed, et al.
Publicado: (2024)
State Matching and Multiple References in Adaptive Active Automata Learning
por: Kruger, Loes, et al.
Publicado: (2024)
por: Kruger, Loes, et al.
Publicado: (2024)
Infectious Disease Forecasting in India using LLM's and Deep Learning
por: Shah, Chaitya, et al.
Publicado: (2024)
por: Shah, Chaitya, et al.
Publicado: (2024)
PICID: Proof-Driven Clause Learning in Neural Network Verification
por: Isac, Omri, et al.
Publicado: (2025)
por: Isac, Omri, et al.
Publicado: (2025)
A PAC Learning Algorithm for LTL and Omega-regular Objectives in MDPs
por: Perez, Mateo, et al.
Publicado: (2023)
por: Perez, Mateo, et al.
Publicado: (2023)
Ejemplares similares
-
Verified SHAP: Provable Bounds for Exact Shapley Values of Neural Networks
por: Boetius, David, et al.
Publicado: (2026) -
Counterexample-Guided Interval Weakening
por: Andrew, Ben M., et al.
Publicado: (2026) -
Progressive Safeguards for Safe and Model-Agnostic Reinforcement Learning
por: Omi, Nabil, et al.
Publicado: (2024) -
Skill Machines: Temporal Logic Skill Composition in Reinforcement Learning
por: Tasse, Geraud Nangue, et al.
Publicado: (2022) -
Compositional Shielding and Reinforcement Learning for Multi-Agent Systems
por: Brorholt, Asger Horn, et al.
Publicado: (2024)