Do It for HER: First-Order Temporal Logic Reward Specification in Reinforcement Learning (Extended Version)
Fuente:
arXiv
Guardado en:
| Autores principales: | Olivieri, Pierriccardo, Lasca, Fausto, Gianola, Alessandro, Papini, Matteo |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
A General Automata Model for First-Order Temporal Logics (Extended Version)
por: Geatti, Luca, et al.
Publicado: (2024)
por: Geatti, Luca, et al.
Publicado: (2024)
Object-Centric Conformance Alignments with Synchronization (Extended Version)
por: Gianola, Alessandro, et al.
Publicado: (2023)
por: Gianola, Alessandro, et al.
Publicado: (2023)
Learning Logic Specifications for Policy Guidance in POMDPs: an Inductive Logic Programming Approach
por: Meli, Daniele, et al.
Publicado: (2024)
por: Meli, Daniele, et al.
Publicado: (2024)
Expressive Temporal Specifications for Reward Monitoring
por: Adalat, Omar, et al.
Publicado: (2025)
por: Adalat, Omar, et al.
Publicado: (2025)
Learning Concepts Definable in First-Order Logic with Counting
por: van Bergerem, Steffen
Publicado: (2019)
por: van Bergerem, Steffen
Publicado: (2019)
Multitask Kernel-based Learning with First-Order Logic Constraints
por: Diligenti, Michelangelo, et al.
Publicado: (2023)
por: Diligenti, Michelangelo, et al.
Publicado: (2023)
Inferring Causal Graph Temporal Logic Formulas to Expedite Reinforcement Learning in Temporally Extended Tasks
por: Aria, Hadi Partovi, et al.
Publicado: (2026)
por: Aria, Hadi Partovi, et al.
Publicado: (2026)
First Order Logic with Fuzzy Semantics for Describing and Recognizing Nerves in Medical Images
por: Bloch, Isabelle, et al.
Publicado: (2025)
por: Bloch, Isabelle, et al.
Publicado: (2025)
The Boolean Solution Problem from the Perspective of Predicate Logic -- Extended Version
por: Wernhard, Christoph
Publicado: (2017)
por: Wernhard, Christoph
Publicado: (2017)
Learning Temporal Logic Predicates from Data with Statistical Guarantees
por: Soroka, Emi, et al.
Publicado: (2024)
por: Soroka, Emi, et al.
Publicado: (2024)
Inductive Generalization in Reinforcement Learning from Specifications
por: Subramanian, Vignesh, et al.
Publicado: (2024)
por: Subramanian, Vignesh, et al.
Publicado: (2024)
Most General Explanations of Tree Ensembles (Extended Version)
por: Izza, Yacine, et al.
Publicado: (2025)
por: Izza, Yacine, et al.
Publicado: (2025)
Autoformalizing Natural Language to First-Order Logic: A Case Study in Logical Fallacy Detection
por: Lalwani, Abhinav, et al.
Publicado: (2024)
por: Lalwani, Abhinav, et al.
Publicado: (2024)
The Logical Expressiveness of Temporal GNNs via Two-Dimensional Product Logics
por: Sälzer, Marco, et al.
Publicado: (2025)
por: Sälzer, Marco, et al.
Publicado: (2025)
Machine Learning Model Integration with Open World Temporal Logic for Process Automation
por: Aditya, Dyuman, et al.
Publicado: (2025)
por: Aditya, Dyuman, et al.
Publicado: (2025)
Temporal Inductive Logic Reasoning over Hypergraphs
por: Yang, Yuan, et al.
Publicado: (2022)
por: Yang, Yuan, et al.
Publicado: (2022)
Lifted Inference beyond First-Order Logic
por: Malhotra, Sagar, et al.
Publicado: (2023)
por: Malhotra, Sagar, et al.
Publicado: (2023)
Guiding Multi-agent Multi-task Reinforcement Learning by a Hierarchical Framework with Logical Reward Shaping
por: Liu, Chanjuan, et al.
Publicado: (2024)
por: Liu, Chanjuan, et al.
Publicado: (2024)
FORM: Learning Expressive and Transferable First-Order Logic Reward Machines
por: Ardon, Leo, et al.
Publicado: (2024)
por: Ardon, Leo, et al.
Publicado: (2024)
Logic of Hypotheses: from Zero to Full Knowledge in Neurosymbolic Integration
por: Bizzaro, Davide, et al.
Publicado: (2025)
por: Bizzaro, Davide, et al.
Publicado: (2025)
Formally Verified Neurosymbolic Trajectory Learning via Tensor-based Linear Temporal Logic on Finite Traces
por: Chevallier, Mark, et al.
Publicado: (2025)
por: Chevallier, Mark, et al.
Publicado: (2025)
Adding Circumscription to Decidable Fragments of First-Order Logic: A Complexity Rollercoaster
por: Lutz, Carsten, et al.
Publicado: (2024)
por: Lutz, Carsten, et al.
Publicado: (2024)
TGPO: Temporal Grounded Policy Optimization for Signal Temporal Logic Tasks
por: Meng, Yue, et al.
Publicado: (2025)
por: Meng, Yue, et al.
Publicado: (2025)
Robust Shielding for Safe Reinforcement Learning
por: Court, Edwin Hamel-De le, et al.
Publicado: (2026)
por: Court, Edwin Hamel-De le, et al.
Publicado: (2026)
Discussion Graph Semantics of First-Order Logic with Equality for Reasoning about Discussion and Argumentation
por: Arisaka, Ryuta
Publicado: (2024)
por: Arisaka, Ryuta
Publicado: (2024)
Diverse Controllable Diffusion Policy with Signal Temporal Logic
por: Meng, Yue, et al.
Publicado: (2025)
por: Meng, Yue, et al.
Publicado: (2025)
Model Predictive Robustness of Signal Temporal Logic Predicates
por: Lin, Yuanfei, et al.
Publicado: (2022)
por: Lin, Yuanfei, et al.
Publicado: (2022)
Learning Probabilistic Temporal Logic Specifications for Stochastic Systems
por: Roy, Rajarshi, et al.
Publicado: (2025)
por: Roy, Rajarshi, et al.
Publicado: (2025)
Extending Defeasibility for Propositional Standpoint Logics
por: Leisegang, Nicholas, et al.
Publicado: (2025)
por: Leisegang, Nicholas, et al.
Publicado: (2025)
SATQuest: A Verifier for Logical Reasoning Evaluation and Reinforcement Fine-Tuning of LLMs
por: Zhao, Yanxiao, et al.
Publicado: (2025)
por: Zhao, Yanxiao, et al.
Publicado: (2025)
A First-Order Logic-Based Alternative to Reward Models in RLHF
por: Jian, Chunjin, et al.
Publicado: (2025)
por: Jian, Chunjin, et al.
Publicado: (2025)
Applications of Intuitionistic Temporal Logic to Temporal Answer Set Programming
por: Cabalar, Pedro, et al.
Publicado: (2026)
por: Cabalar, Pedro, et al.
Publicado: (2026)
Technical Report -- A Context-Sensitive Multi-Level Similarity Framework for First-Order Logic Arguments: An Axiomatic Study
por: David, Victor, et al.
Publicado: (2026)
por: David, Victor, et al.
Publicado: (2026)
Reduced Implication-bias Logic Loss for Neuro-Symbolic Learning
por: He, Haoyuan, et al.
Publicado: (2022)
por: He, Haoyuan, et al.
Publicado: (2022)
A Variable Occurrence-Centric Framework for Inconsistency Handling (Extended Version)
por: Salhi, Yakoub
Publicado: (2024)
por: Salhi, Yakoub
Publicado: (2024)
Automated planning with ontologies under coherence update semantics (Extended Version)
por: Borgwardt, Stefan, et al.
Publicado: (2025)
por: Borgwardt, Stefan, et al.
Publicado: (2025)
The Shape of a Benedictine Monastery: The SaintGall Ontology (Extended Version)
por: Cantale, Claudia, et al.
Publicado: (2017)
por: Cantale, Claudia, et al.
Publicado: (2017)
Splitting Answer Set Programs with respect to Intensionality Statements (Extended Version)
por: Fandinno, Jorge, et al.
Publicado: (2025)
por: Fandinno, Jorge, et al.
Publicado: (2025)
Lattice Annotated Temporal (LAT) Logic for Non-Markovian Reasoning
por: Mukherji, Kaustuv, et al.
Publicado: (2025)
por: Mukherji, Kaustuv, et al.
Publicado: (2025)
Non-Rigid Designators in Modal and Temporal Free Description Logics (Extended Version)
por: Artale, Alessandro, et al.
Publicado: (2024)
por: Artale, Alessandro, et al.
Publicado: (2024)
Ejemplares similares
-
A General Automata Model for First-Order Temporal Logics (Extended Version)
por: Geatti, Luca, et al.
Publicado: (2024) -
Object-Centric Conformance Alignments with Synchronization (Extended Version)
por: Gianola, Alessandro, et al.
Publicado: (2023) -
Learning Logic Specifications for Policy Guidance in POMDPs: an Inductive Logic Programming Approach
por: Meli, Daniele, et al.
Publicado: (2024) -
Expressive Temporal Specifications for Reward Monitoring
por: Adalat, Omar, et al.
Publicado: (2025) -
Learning Concepts Definable in First-Order Logic with Counting
por: van Bergerem, Steffen
Publicado: (2019)