Learning Optimal Strategies for Temporal Tasks in Stochastic Games
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Bozkurt, Alper Kamil, Wang, Yu, Zavlanos, Michael M., Pajic, Miroslav |
|---|---|
| Format: | Preprint |
| Publié: |
2021
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Model-Free Reinforcement Learning for Stochastic Games with Linear Temporal Logic Objectives
par: Bozkurt, Alper Kamil, et autres
Publié: (2020)
par: Bozkurt, Alper Kamil, et autres
Publié: (2020)
Control Synthesis from Linear Temporal Logic Specifications using Model-Free Reinforcement Learning
par: Bozkurt, Alper Kamil, et autres
Publié: (2019)
par: Bozkurt, Alper Kamil, et autres
Publié: (2019)
Model-Free Learning of Safe yet Effective Controllers
par: Bozkurt, Alper Kamil, et autres
Publié: (2021)
par: Bozkurt, Alper Kamil, et autres
Publié: (2021)
On the Uniqueness of Solution for the Bellman Equation of LTL Objectives
par: Xuan, Zetong, et autres
Publié: (2024)
par: Xuan, Zetong, et autres
Publié: (2024)
TGPO: Temporal Grounded Policy Optimization for Signal Temporal Logic Tasks
par: Meng, Yue, et autres
Publié: (2025)
par: Meng, Yue, et autres
Publié: (2025)
Secure Planning Against Stealthy Attacks via Model-Free Reinforcement Learning
par: Bozkurt, Alper Kamil, et autres
Publié: (2020)
par: Bozkurt, Alper Kamil, et autres
Publié: (2020)
Value Functions for Temporal Logic: Optimal Policies and Safety Filters
par: So, Oswin, et autres
Publié: (2026)
par: So, Oswin, et autres
Publié: (2026)
Diverse Controllable Diffusion Policy with Signal Temporal Logic
par: Meng, Yue, et autres
Publié: (2025)
par: Meng, Yue, et autres
Publié: (2025)
Model Predictive Robustness of Signal Temporal Logic Predicates
par: Lin, Yuanfei, et autres
Publié: (2022)
par: Lin, Yuanfei, et autres
Publié: (2022)
Adaptive Policy Selection and Fine-Tuning under Interaction Budgets for Offline-to-Online Reinforcement Learning
par: Bozkurt, Alper Kamil, et autres
Publié: (2026)
par: Bozkurt, Alper Kamil, et autres
Publié: (2026)
Integrating LTL Constraints into PPO for Safe Reinforcement Learning
par: Zhang, Maifang, et autres
Publié: (2026)
par: Zhang, Maifang, et autres
Publié: (2026)
Recover: A Neuro-Symbolic Framework for Failure Detection and Recovery
par: Cornelio, Cristina, et autres
Publié: (2024)
par: Cornelio, Cristina, et autres
Publié: (2024)
Shield Synthesis for LTL Modulo Theories
par: Rodriguez, Andoni, et autres
Publié: (2024)
par: Rodriguez, Andoni, et autres
Publié: (2024)
Towards Safe Autonomous Driving Policies using a Neuro-Symbolic Deep Reinforcement Learning Approach
par: Sharifi, Iman, et autres
Publié: (2023)
par: Sharifi, Iman, et autres
Publié: (2023)
Nl2Hltl2Plan: Scaling Up Natural Language Understanding for Multi-Robots Through Hierarchical Temporal Logic Task Representation
par: Xu, Shaojun, et autres
Publié: (2024)
par: Xu, Shaojun, et autres
Publié: (2024)
Inductive Learning of Robot Task Knowledge from Raw Data and Online Expert Feedback
par: Meli, Daniele, et autres
Publié: (2025)
par: Meli, Daniele, et autres
Publié: (2025)
Accelerated Learning with Linear Temporal Logic using Differentiable Simulation
par: Bozkurt, Alper Kamil, et autres
Publié: (2025)
par: Bozkurt, Alper Kamil, et autres
Publié: (2025)
Efficient Dynamic Shielding for Parametric Safety Specifications
par: Corsi, Davide, et autres
Publié: (2025)
par: Corsi, Davide, et autres
Publié: (2025)
Learning Temporal Logic Predicates from Data with Statistical Guarantees
par: Soroka, Emi, et autres
Publié: (2024)
par: Soroka, Emi, et autres
Publié: (2024)
Learning Explainable and Better Performing Representations of POMDP Strategies
par: Bork, Alexander, et autres
Publié: (2024)
par: Bork, Alexander, et autres
Publié: (2024)
Machine Learning Model Integration with Open World Temporal Logic for Process Automation
par: Aditya, Dyuman, et autres
Publié: (2025)
par: Aditya, Dyuman, et autres
Publié: (2025)
Locally Pareto-Optimal Interpretations for Black-Box Machine Learning Models
par: Joshi, Aniruddha, et autres
Publié: (2025)
par: Joshi, Aniruddha, et autres
Publié: (2025)
Do It for HER: First-Order Temporal Logic Reward Specification in Reinforcement Learning (Extended Version)
par: Olivieri, Pierriccardo, et autres
Publié: (2026)
par: Olivieri, Pierriccardo, et autres
Publié: (2026)
Formally Verified Neurosymbolic Trajectory Learning via Tensor-based Linear Temporal Logic on Finite Traces
par: Chevallier, Mark, et autres
Publié: (2025)
par: Chevallier, Mark, et autres
Publié: (2025)
Expressive Temporal Specifications for Reward Monitoring
par: Adalat, Omar, et autres
Publié: (2025)
par: Adalat, Omar, et autres
Publié: (2025)
Temporal Inductive Logic Reasoning over Hypergraphs
par: Yang, Yuan, et autres
Publié: (2022)
par: Yang, Yuan, et autres
Publié: (2022)
Approximating Pareto Frontiers in Stochastic Multi-Objective Optimization via Hashing and Randomization
par: Li, Jinzhao, et autres
Publié: (2026)
par: Li, Jinzhao, et autres
Publié: (2026)
The Logical Expressiveness of Temporal GNNs via Two-Dimensional Product Logics
par: Sälzer, Marco, et autres
Publié: (2025)
par: Sälzer, Marco, et autres
Publié: (2025)
Constraint-aware Learning of Probabilistic Sequential Models for Multi-Label Classification
par: Buleshnyi, Mykhailo, et autres
Publié: (2025)
par: Buleshnyi, Mykhailo, et autres
Publié: (2025)
Model-Free Reinforcement Learning for Symbolic Automata-encoded Objectives
par: Balakrishnan, Anand, et autres
Publié: (2022)
par: Balakrishnan, Anand, et autres
Publié: (2022)
The Optimal Choice of Hypothesis Is the Weakest, Not the Shortest
par: Bennett, Michael Timothy
Publié: (2023)
par: Bennett, Michael Timothy
Publié: (2023)
Dualformer: Controllable Fast and Slow Thinking by Learning with Randomized Reasoning Traces
par: Su, DiJia, et autres
Publié: (2024)
par: Su, DiJia, et autres
Publié: (2024)
Preference-Based Planning in Stochastic Environments: From Partially-Ordered Temporal Goals to Most Preferred Policies
par: Rahmani, Hazhar, et autres
Publié: (2024)
par: Rahmani, Hazhar, et autres
Publié: (2024)
Learning to Solve and Optimize by Evolving Code
par: Semmelrock, Veronika, et autres
Publié: (2026)
par: Semmelrock, Veronika, et autres
Publié: (2026)
Robust Shielding for Safe Reinforcement Learning
par: Court, Edwin Hamel-De le, et autres
Publié: (2026)
par: Court, Edwin Hamel-De le, et autres
Publié: (2026)
Machine Learning for Quantifier Selection in cvc5
par: Jakubův, Jan, et autres
Publié: (2024)
par: Jakubův, Jan, et autres
Publié: (2024)
Inductive Generalization in Reinforcement Learning from Specifications
par: Subramanian, Vignesh, et autres
Publié: (2024)
par: Subramanian, Vignesh, et autres
Publié: (2024)
SATformer: Transformer-Based UNSAT Core Learning
par: Shi, Zhengyuan, et autres
Publié: (2022)
par: Shi, Zhengyuan, et autres
Publié: (2022)
Learning big logical rules by joining small rules
par: Hocquette, Céline, et autres
Publié: (2024)
par: Hocquette, Céline, et autres
Publié: (2024)
Can Transformers Learn to Verify During Backtracking Search?
par: Phua, Yin Jun, et autres
Publié: (2026)
par: Phua, Yin Jun, et autres
Publié: (2026)
Documents similaires
-
Model-Free Reinforcement Learning for Stochastic Games with Linear Temporal Logic Objectives
par: Bozkurt, Alper Kamil, et autres
Publié: (2020) -
Control Synthesis from Linear Temporal Logic Specifications using Model-Free Reinforcement Learning
par: Bozkurt, Alper Kamil, et autres
Publié: (2019) -
Model-Free Learning of Safe yet Effective Controllers
par: Bozkurt, Alper Kamil, et autres
Publié: (2021) -
On the Uniqueness of Solution for the Bellman Equation of LTL Objectives
par: Xuan, Zetong, et autres
Publié: (2024) -
TGPO: Temporal Grounded Policy Optimization for Signal Temporal Logic Tasks
par: Meng, Yue, et autres
Publié: (2025)