Pushdown Reward Machines for Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Varricchione, Giovanni, Klassen, Toryn Q., Alechina, Natasha, Dastani, Mehdi, Logan, Brian, McIlraith, Sheila A. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Maximally Permissive Reward Machines
by: Varricchione, Giovanni, et al.
Published: (2024)
by: Varricchione, Giovanni, et al.
Published: (2024)
Pluralistic Alignment Over Time
by: Klassen, Toryn Q., et al.
Published: (2024)
by: Klassen, Toryn Q., et al.
Published: (2024)
Formal Methods Meet LLMs: Auditing, Monitoring, and Intervention for Compliance of Advanced AI Systems
by: Alamdari, Parand A., et al.
Published: (2026)
by: Alamdari, Parand A., et al.
Published: (2026)
Remembering to Be Fair: Non-Markovian Fairness in Sequential Decision Making
by: Alamdari, Parand A., et al.
Published: (2023)
by: Alamdari, Parand A., et al.
Published: (2023)
Being Considerate as a Pathway Towards Pluralistic Alignment for Agentic AI
by: Alamdari, Parand A., et al.
Published: (2024)
by: Alamdari, Parand A., et al.
Published: (2024)
Learning Bilevel Policies over Symbolic World Models for Long-Horizon Planning
by: Chen, Dillon Z., et al.
Published: (2026)
by: Chen, Dillon Z., et al.
Published: (2026)
Satisficing and Optimal Generalised Planning via Goal Regression (Extended Version)
by: Chen, Dillon Z., et al.
Published: (2025)
by: Chen, Dillon Z., et al.
Published: (2025)
Ground-Compose-Reinforce: Grounding Language in Agentic Behaviours using Limited Data
by: Li, Andrew C., et al.
Published: (2025)
by: Li, Andrew C., et al.
Published: (2025)
Reward Machines for Deep RL in Noisy and Uncertain Environments
by: Li, Andrew C., et al.
Published: (2024)
by: Li, Andrew C., et al.
Published: (2024)
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning
by: Zhu, Changxi, et al.
Published: (2025)
by: Zhu, Changxi, et al.
Published: (2025)
Causes and Strategies in Multiagent Systems
by: Kerkhove, Sylvia S., et al.
Published: (2025)
by: Kerkhove, Sylvia S., et al.
Published: (2025)
Temporal Causal Reasoning with (Non-Recursive) Structural Equation Models
by: Gladyshev, Maksim, et al.
Published: (2025)
by: Gladyshev, Maksim, et al.
Published: (2025)
Learning Communication Skills in Multi-task Multi-agent Deep Reinforcement Learning
by: Zhu, Changxi, et al.
Published: (2025)
by: Zhu, Changxi, et al.
Published: (2025)
From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning
by: Park, Junseok, et al.
Published: (2025)
by: Park, Junseok, et al.
Published: (2025)
Deep Reinforcement Learning Xiangqi Player with Monte Carlo Tree Search
by: Yilmaz, Berk, et al.
Published: (2025)
by: Yilmaz, Berk, et al.
Published: (2025)
BatteryML:An Open-source platform for Machine Learning on Battery Degradation
by: Zhang, Han, et al.
Published: (2023)
by: Zhang, Han, et al.
Published: (2023)
Multi-State TD Target for Model-Free Reinforcement Learning
by: Wang, Wuhao, et al.
Published: (2024)
by: Wang, Wuhao, et al.
Published: (2024)
STEVE-1: A Generative Model for Text-to-Behavior in Minecraft
by: Lifshitz, Shalev, et al.
Published: (2023)
by: Lifshitz, Shalev, et al.
Published: (2023)
Topological Foundations of Reinforcement Learning
by: Kadurha, David Krame
Published: (2024)
by: Kadurha, David Krame
Published: (2024)
Multi-Agent Verification: Scaling Test-Time Compute with Multiple Verifiers
by: Lifshitz, Shalev, et al.
Published: (2025)
by: Lifshitz, Shalev, et al.
Published: (2025)
Unsupervised Ensemble Learning Through Deep Energy-based Models
by: Maymon, Ariel, et al.
Published: (2026)
by: Maymon, Ariel, et al.
Published: (2026)
Reciprocal Learning
by: Rodemann, Julian, et al.
Published: (2024)
by: Rodemann, Julian, et al.
Published: (2024)
Grouped Sequential Optimization Strategy -- the Application of Hyperparameter Importance Assessment in Deep Learning
by: Wang, Ruinan, et al.
Published: (2025)
by: Wang, Ruinan, et al.
Published: (2025)
Car Sensors Health Monitoring by Verification Based on Autoencoder and Random Forest Regression
by: Torkhesari, Sahar, et al.
Published: (2025)
by: Torkhesari, Sahar, et al.
Published: (2025)
Convergence Dynamics and Stabilization Strategies of Co-Evolving Generative Models
by: Gao, Weiguo, et al.
Published: (2025)
by: Gao, Weiguo, et al.
Published: (2025)
Weakly Supervised Learners for Correction of AI Errors with Provable Performance Guarantees
by: Tyukin, Ivan Y., et al.
Published: (2024)
by: Tyukin, Ivan Y., et al.
Published: (2024)
CircuitBuilder: From Polynomials to Circuits via Reinforcement Learning
by: Zhang, Weikun K., et al.
Published: (2026)
by: Zhang, Weikun K., et al.
Published: (2026)
Quantifying First-Order Markov Violations in Noisy Reinforcement Learning: A Causal Discovery Approach
by: Mysore, Naveen
Published: (2025)
by: Mysore, Naveen
Published: (2025)
Insights into Schizophrenia: Leveraging Machine Learning for Early Identification via EEG, ERP, and Demographic Attributes
by: Alkhalifa, Sara
Published: (2025)
by: Alkhalifa, Sara
Published: (2025)
DQN Performance with Epsilon Greedy Policies and Prioritized Experience Replay
by: Perkins, Daniel, et al.
Published: (2025)
by: Perkins, Daniel, et al.
Published: (2025)
Subset Selection for Fine-Tuning: A Utility-Diversity Balanced Approach for Mathematical Domain Adaptation
by: Kotecha, Madhav, et al.
Published: (2025)
by: Kotecha, Madhav, et al.
Published: (2025)
Machine Learning-Driven Predictive Resource Management in Complex Science Workflows
by: Chowdhury, Tasnuva, et al.
Published: (2025)
by: Chowdhury, Tasnuva, et al.
Published: (2025)
Neuro-symbolic Action Masking for Deep Reinforcement Learning
by: Han, Shuai, et al.
Published: (2026)
by: Han, Shuai, et al.
Published: (2026)
Large Language Model Meets Graph Neural Network in Knowledge Distillation
by: Hu, Shengxiang, et al.
Published: (2024)
by: Hu, Shengxiang, et al.
Published: (2024)
ASNN: Learning to Suggest Neural Architectures from Performance Distributions
by: Hong, Jinwook
Published: (2025)
by: Hong, Jinwook
Published: (2025)
A Neural Affinity Framework for Abstract Reasoning: Diagnosing the Compositional Gap in Transformer Architectures via Procedural Task Taxonomy
by: Ingram, Miguel, et al.
Published: (2025)
by: Ingram, Miguel, et al.
Published: (2025)
From Imitation to Interaction: Mastering Game of Schnapsen with Shallow Reinforcement Learning
by: Klačan, Ján, et al.
Published: (2026)
by: Klačan, Ján, et al.
Published: (2026)
Multi-Scale Graph Learning for Anti-Sparse Downscaling
by: Fan, Yingda, et al.
Published: (2025)
by: Fan, Yingda, et al.
Published: (2025)
OptPO: Optimal Rollout Allocation for Test-time Policy Optimization
by: Wang, Youkang, et al.
Published: (2025)
by: Wang, Youkang, et al.
Published: (2025)
What is the $\textit{intrinsic}$ dimension of your binary data? -- and how to compute it quickly
by: Hanika, Tom, et al.
Published: (2024)
by: Hanika, Tom, et al.
Published: (2024)
Similar Items
-
Maximally Permissive Reward Machines
by: Varricchione, Giovanni, et al.
Published: (2024) -
Pluralistic Alignment Over Time
by: Klassen, Toryn Q., et al.
Published: (2024) -
Formal Methods Meet LLMs: Auditing, Monitoring, and Intervention for Compliance of Advanced AI Systems
by: Alamdari, Parand A., et al.
Published: (2026) -
Remembering to Be Fair: Non-Markovian Fairness in Sequential Decision Making
by: Alamdari, Parand A., et al.
Published: (2023) -
Being Considerate as a Pathway Towards Pluralistic Alignment for Agentic AI
by: Alamdari, Parand A., et al.
Published: (2024)