Learning Explainable and Better Performing Representations of POMDP Strategies
Fuente:
arXiv
Guardado en:
| Autores principales: | Bork, Alexander, Chakraborty, Debraj, Grover, Kush, Kretinsky, Jan, Mohr, Stefanie |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Resilient Strategies for Stochastic Systems: How Much Does It Take to Break a Winning Strategy?
por: Grover, Kush, et al.
Publicado: (2026)
por: Grover, Kush, et al.
Publicado: (2026)
1-2-3-Go! Policy Synthesis for Parameterized Markov Decision Processes via Decision-Tree Learning and Generalization
por: Azeem, Muqsit, et al.
Publicado: (2024)
por: Azeem, Muqsit, et al.
Publicado: (2024)
Logic of Fuzzy Paths
por: Grover, Kush, et al.
Publicado: (2026)
por: Grover, Kush, et al.
Publicado: (2026)
Synthesizing POMDP Policies: Sampling Meets Model-checking via Learning
por: Chakraborty, Debraj, et al.
Publicado: (2026)
por: Chakraborty, Debraj, et al.
Publicado: (2026)
Explainable Representation of Finite-Memory Policies for POMDPs using Decision Trees
por: Azeem, Muqsit, et al.
Publicado: (2024)
por: Azeem, Muqsit, et al.
Publicado: (2024)
Locally Pareto-Optimal Interpretations for Black-Box Machine Learning Models
por: Joshi, Aniruddha, et al.
Publicado: (2025)
por: Joshi, Aniruddha, et al.
Publicado: (2025)
Comparing Neural Network Encodings for Logic-based Explainability
por: Carvalho, Levi Cordeiro, et al.
Publicado: (2025)
por: Carvalho, Levi Cordeiro, et al.
Publicado: (2025)
MULTIGAIN 2.0: MDP controller synthesis for multiple mean-payoff, LTL and steady-state constraints
por: Bals, Severin, et al.
Publicado: (2023)
por: Bals, Severin, et al.
Publicado: (2023)
Transfer Learning from Foundational Optimization Embeddings to Unsupervised SAT Representations
por: Pal, Koyena, et al.
Publicado: (2026)
por: Pal, Koyena, et al.
Publicado: (2026)
SDSC:A Structure-Aware Metric for Semantic Signal Representation Learning
por: Lee, Jeyoung, et al.
Publicado: (2025)
por: Lee, Jeyoung, et al.
Publicado: (2025)
Hierarchical Attention Generates Better Proofs
por: Chen, Jianlong, et al.
Publicado: (2025)
por: Chen, Jianlong, et al.
Publicado: (2025)
Machine Learning for Quantifier Selection in cvc5
por: Jakubův, Jan, et al.
Publicado: (2024)
por: Jakubův, Jan, et al.
Publicado: (2024)
Solving Hard Mizar Problems with Instantiation and Strategy Invention
por: Jakubův, Jan, et al.
Publicado: (2024)
por: Jakubův, Jan, et al.
Publicado: (2024)
Stopping Criteria for Value Iteration on Concurrent Stochastic Reachability and Safety Games
por: Grobelna, Marta, et al.
Publicado: (2025)
por: Grobelna, Marta, et al.
Publicado: (2025)
Regional, Lattice and Logical Representations of Neural Networks
por: Preto, Sandro, et al.
Publicado: (2025)
por: Preto, Sandro, et al.
Publicado: (2025)
Circuit Representations of Random Forests with Applications to XAI
por: Ji, Chunxi, et al.
Publicado: (2026)
por: Ji, Chunxi, et al.
Publicado: (2026)
Learning Optimal Strategies for Temporal Tasks in Stochastic Games
por: Bozkurt, Alper Kamil, et al.
Publicado: (2021)
por: Bozkurt, Alper Kamil, et al.
Publicado: (2021)
From Symbolic Tasks to Code Generation: Diversification Yields Better Task Performers
por: Zhang, Dylan, et al.
Publicado: (2024)
por: Zhang, Dylan, et al.
Publicado: (2024)
RocqSmith: Can Automatic Optimization Forge Better Proof Agents?
por: Kozyrev, Andrei, et al.
Publicado: (2026)
por: Kozyrev, Andrei, et al.
Publicado: (2026)
Halting Recurrent GNNs and the Graded $μ$-Calculus
por: Bollen, Jeroen, et al.
Publicado: (2025)
por: Bollen, Jeroen, et al.
Publicado: (2025)
Learning to Solve and Optimize by Evolving Code
por: Semmelrock, Veronika, et al.
Publicado: (2026)
por: Semmelrock, Veronika, et al.
Publicado: (2026)
Robust Shielding for Safe Reinforcement Learning
por: Court, Edwin Hamel-De le, et al.
Publicado: (2026)
por: Court, Edwin Hamel-De le, et al.
Publicado: (2026)
Automated Design of Linear Bounding Functions for Sigmoidal Nonlinearities in Neural Networks
por: König, Matthias, et al.
Publicado: (2024)
por: König, Matthias, et al.
Publicado: (2024)
Inductive Generalization in Reinforcement Learning from Specifications
por: Subramanian, Vignesh, et al.
Publicado: (2024)
por: Subramanian, Vignesh, et al.
Publicado: (2024)
SATformer: Transformer-Based UNSAT Core Learning
por: Shi, Zhengyuan, et al.
Publicado: (2022)
por: Shi, Zhengyuan, et al.
Publicado: (2022)
The Geometry of Reasoning: Flowing Logics in Representation Space
por: Zhou, Yufa, et al.
Publicado: (2025)
por: Zhou, Yufa, et al.
Publicado: (2025)
SemML 2.0: Synthesizing Controllers for LTL
por: Křetínský, Jan, et al.
Publicado: (2026)
por: Křetínský, Jan, et al.
Publicado: (2026)
Learning big logical rules by joining small rules
por: Hocquette, Céline, et al.
Publicado: (2024)
por: Hocquette, Céline, et al.
Publicado: (2024)
Inference of Abstraction for a Unified Account of Reasoning and Learning
por: Kido, Hiroyuki
Publicado: (2024)
por: Kido, Hiroyuki
Publicado: (2024)
Compositional Shielding and Reinforcement Learning for Multi-Agent Systems
por: Brorholt, Asger Horn, et al.
Publicado: (2024)
por: Brorholt, Asger Horn, et al.
Publicado: (2024)
LeanAgent: Lifelong Learning for Formal Theorem Proving
por: Kumarappan, Adarsh, et al.
Publicado: (2024)
por: Kumarappan, Adarsh, et al.
Publicado: (2024)
Can Transformers Learn to Verify During Backtracking Search?
por: Phua, Yin Jun, et al.
Publicado: (2026)
por: Phua, Yin Jun, et al.
Publicado: (2026)
Learning Concepts Definable in First-Order Logic with Counting
por: van Bergerem, Steffen
Publicado: (2019)
por: van Bergerem, Steffen
Publicado: (2019)
Learning Brave Assumption-Based Argumentation Frameworks via ASP
por: De Angelis, Emanuele, et al.
Publicado: (2024)
por: De Angelis, Emanuele, et al.
Publicado: (2024)
Learning Temporal Logic Predicates from Data with Statistical Guarantees
por: Soroka, Emi, et al.
Publicado: (2024)
por: Soroka, Emi, et al.
Publicado: (2024)
PiShield: A PyTorch Package for Learning with Requirements
por: Stoian, Mihaela Cătălina, et al.
Publicado: (2024)
por: Stoian, Mihaela Cătălina, et al.
Publicado: (2024)
Reduced Implication-bias Logic Loss for Neuro-Symbolic Learning
por: He, Haoyuan, et al.
Publicado: (2022)
por: He, Haoyuan, et al.
Publicado: (2022)
Compact Rule-Based Classifier Learning via Gradient Descent
por: Fumanal-Idocin, Javier, et al.
Publicado: (2025)
por: Fumanal-Idocin, Javier, et al.
Publicado: (2025)
Multitask Kernel-based Learning with First-Order Logic Constraints
por: Diligenti, Michelangelo, et al.
Publicado: (2023)
por: Diligenti, Michelangelo, et al.
Publicado: (2023)
A Reinforcement Learning based Reset Policy for CDCL SAT Solvers
por: Li, Chunxiao, et al.
Publicado: (2024)
por: Li, Chunxiao, et al.
Publicado: (2024)
Ejemplares similares
-
Resilient Strategies for Stochastic Systems: How Much Does It Take to Break a Winning Strategy?
por: Grover, Kush, et al.
Publicado: (2026) -
1-2-3-Go! Policy Synthesis for Parameterized Markov Decision Processes via Decision-Tree Learning and Generalization
por: Azeem, Muqsit, et al.
Publicado: (2024) -
Logic of Fuzzy Paths
por: Grover, Kush, et al.
Publicado: (2026) -
Synthesizing POMDP Policies: Sampling Meets Model-checking via Learning
por: Chakraborty, Debraj, et al.
Publicado: (2026) -
Explainable Representation of Finite-Memory Policies for POMDPs using Decision Trees
por: Azeem, Muqsit, et al.
Publicado: (2024)