Inductive Generalization in Reinforcement Learning from Specifications
Fuente:
arXiv
Guardado en:
| Autores principales: | Subramanian, Vignesh, Kushwah, Rohit, Roy, Subhajit, Bansal, Suguman |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Decoupled Behavioral Cloning for Scalable Inductive Generalization in RL from Specifications
por: Subramanian, Vignesh, et al.
Publicado: (2026)
por: Subramanian, Vignesh, et al.
Publicado: (2026)
Learning Logic Specifications for Policy Guidance in POMDPs: an Inductive Logic Programming Approach
por: Meli, Daniele, et al.
Publicado: (2024)
por: Meli, Daniele, et al.
Publicado: (2024)
Certificate-Guided Evaluation of Reinforcement Learning Generalization
por: Subramanian, Vignesh, et al.
Publicado: (2026)
por: Subramanian, Vignesh, et al.
Publicado: (2026)
Do It for HER: First-Order Temporal Logic Reward Specification in Reinforcement Learning (Extended Version)
por: Olivieri, Pierriccardo, et al.
Publicado: (2026)
por: Olivieri, Pierriccardo, et al.
Publicado: (2026)
Temporal Inductive Logic Reasoning over Hypergraphs
por: Yang, Yuan, et al.
Publicado: (2022)
por: Yang, Yuan, et al.
Publicado: (2022)
GLIDR: Graph-Like Inductive Logic Programming with Differentiable Reasoning
por: Johnson, Blair, et al.
Publicado: (2025)
por: Johnson, Blair, et al.
Publicado: (2025)
From Circuit Evidence to Mechanistic Theory: An Inductive Logic Approach
por: Aljaafari, Nura, et al.
Publicado: (2026)
por: Aljaafari, Nura, et al.
Publicado: (2026)
Structured Abductive-Deductive-Inductive Reasoning for LLMs via Algebraic Invariants
por: Gilda, Sankalp, et al.
Publicado: (2026)
por: Gilda, Sankalp, et al.
Publicado: (2026)
ANDRE: An Attention-based Neuro-symbolic Differentiable Rule Extractor for Inductive Logic Programming
por: Sharifi, Iman, et al.
Publicado: (2026)
por: Sharifi, Iman, et al.
Publicado: (2026)
Robust Shielding for Safe Reinforcement Learning
por: Court, Edwin Hamel-De le, et al.
Publicado: (2026)
por: Court, Edwin Hamel-De le, et al.
Publicado: (2026)
Compositional Shielding and Reinforcement Learning for Multi-Agent Systems
por: Brorholt, Asger Horn, et al.
Publicado: (2024)
por: Brorholt, Asger Horn, et al.
Publicado: (2024)
Expressive Temporal Specifications for Reward Monitoring
por: Adalat, Omar, et al.
Publicado: (2025)
por: Adalat, Omar, et al.
Publicado: (2025)
A Reinforcement Learning based Reset Policy for CDCL SAT Solvers
por: Li, Chunxiao, et al.
Publicado: (2024)
por: Li, Chunxiao, et al.
Publicado: (2024)
Chronosymbolic Learning: Efficient CHC Solving with Symbolic Reasoning and Inductive Learning
por: Luo, Ziyan, et al.
Publicado: (2023)
por: Luo, Ziyan, et al.
Publicado: (2023)
Value Function Initialization for Knowledge Transfer and Jump-start in Deep Reinforcement Learning
por: Mehimeh, Soumia
Publicado: (2025)
por: Mehimeh, Soumia
Publicado: (2025)
k-Inductive Neural Barrier Certificates for Unknown Nonlinear Dynamics
por: Wooding, Ben, et al.
Publicado: (2026)
por: Wooding, Ben, et al.
Publicado: (2026)
Compiling High-Level Neural Network Specifications into VNN-LIB Queries
por: Daggitt, Matthew L., et al.
Publicado: (2024)
por: Daggitt, Matthew L., et al.
Publicado: (2024)
Transformers Can Learn Connectivity in Some Graphs but Not Others
por: Roy, Amit, et al.
Publicado: (2025)
por: Roy, Amit, et al.
Publicado: (2025)
Integrating LTL Constraints into PPO for Safe Reinforcement Learning
por: Zhang, Maifang, et al.
Publicado: (2026)
por: Zhang, Maifang, et al.
Publicado: (2026)
Feasibility-Guided Fair Adaptive Offline Reinforcement Learning for Medicaid Care Management
por: Basu, Sanjay, et al.
Publicado: (2025)
por: Basu, Sanjay, et al.
Publicado: (2025)
SATQuest: A Verifier for Logical Reasoning Evaluation and Reinforcement Fine-Tuning of LLMs
por: Zhao, Yanxiao, et al.
Publicado: (2025)
por: Zhao, Yanxiao, et al.
Publicado: (2025)
Hyperproperty-Constrained Secure Reinforcement Learning
por: Bonnah, Ernest, et al.
Publicado: (2025)
por: Bonnah, Ernest, et al.
Publicado: (2025)
Learning Temporal Logic Predicates from Data with Statistical Guarantees
por: Soroka, Emi, et al.
Publicado: (2024)
por: Soroka, Emi, et al.
Publicado: (2024)
Transfer Learning from Foundational Optimization Embeddings to Unsupervised SAT Representations
por: Pal, Koyena, et al.
Publicado: (2026)
por: Pal, Koyena, et al.
Publicado: (2026)
Munkres' General Topology Autoformalized in Isabelle/HOL
por: Bryant, Dustin, et al.
Publicado: (2026)
por: Bryant, Dustin, et al.
Publicado: (2026)
Most General Explanations of Tree Ensembles (Extended Version)
por: Izza, Yacine, et al.
Publicado: (2025)
por: Izza, Yacine, et al.
Publicado: (2025)
Logic Tensor Network-Enhanced Generative Adversarial Network
por: Upreti, Nijesh, et al.
Publicado: (2026)
por: Upreti, Nijesh, et al.
Publicado: (2026)
Generating $SROI^-$ Ontologies via Knowledge Graph Query Embedding Learning
por: He, Yunjie, et al.
Publicado: (2024)
por: He, Yunjie, et al.
Publicado: (2024)
Procedural Adherence and Interpretability Through Neuro-Symbolic Generative Agents
por: Rothkopf, Raven, et al.
Publicado: (2024)
por: Rothkopf, Raven, et al.
Publicado: (2024)
MathlibLemma: Folklore Lemma Generation and Benchmark for Formal Mathematics
por: Liu, Xinyu, et al.
Publicado: (2026)
por: Liu, Xinyu, et al.
Publicado: (2026)
Learning to Solve and Optimize by Evolving Code
por: Semmelrock, Veronika, et al.
Publicado: (2026)
por: Semmelrock, Veronika, et al.
Publicado: (2026)
Geospatial Trajectory Generation via Efficient Abduction: Deployment for Independent Testing
por: Bavikadi, Divyagna, et al.
Publicado: (2024)
por: Bavikadi, Divyagna, et al.
Publicado: (2024)
MathConstraint: Automated Generation of Verified Combinatorial Reasoning Instances for LLMs
por: Pati, Viresh, et al.
Publicado: (2026)
por: Pati, Viresh, et al.
Publicado: (2026)
MC3G: Model Agnostic Causally Constrained Counterfactual Generation
por: Dasgupta, Sopam, et al.
Publicado: (2025)
por: Dasgupta, Sopam, et al.
Publicado: (2025)
Machine Learning for Quantifier Selection in cvc5
por: Jakubův, Jan, et al.
Publicado: (2024)
por: Jakubův, Jan, et al.
Publicado: (2024)
SATformer: Transformer-Based UNSAT Core Learning
por: Shi, Zhengyuan, et al.
Publicado: (2022)
por: Shi, Zhengyuan, et al.
Publicado: (2022)
DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search
por: Xin, Huajian, et al.
Publicado: (2024)
por: Xin, Huajian, et al.
Publicado: (2024)
Learning big logical rules by joining small rules
por: Hocquette, Céline, et al.
Publicado: (2024)
por: Hocquette, Céline, et al.
Publicado: (2024)
Learning Explainable and Better Performing Representations of POMDP Strategies
por: Bork, Alexander, et al.
Publicado: (2024)
por: Bork, Alexander, et al.
Publicado: (2024)
Inference of Abstraction for a Unified Account of Reasoning and Learning
por: Kido, Hiroyuki
Publicado: (2024)
por: Kido, Hiroyuki
Publicado: (2024)
Ejemplares similares
-
Decoupled Behavioral Cloning for Scalable Inductive Generalization in RL from Specifications
por: Subramanian, Vignesh, et al.
Publicado: (2026) -
Learning Logic Specifications for Policy Guidance in POMDPs: an Inductive Logic Programming Approach
por: Meli, Daniele, et al.
Publicado: (2024) -
Certificate-Guided Evaluation of Reinforcement Learning Generalization
por: Subramanian, Vignesh, et al.
Publicado: (2026) -
Do It for HER: First-Order Temporal Logic Reward Specification in Reinforcement Learning (Extended Version)
por: Olivieri, Pierriccardo, et al.
Publicado: (2026) -
Temporal Inductive Logic Reasoning over Hypergraphs
por: Yang, Yuan, et al.
Publicado: (2022)