Reward Machines for Deep RL in Noisy and Uncertain Environments
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Andrew C., Chen, Zizhao, Klassen, Toryn Q., Vaezipoor, Pashootan, Icarte, Rodrigo Toro, McIlraith, Sheila A. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Active Automata Learning with Advice
by: Fica, Michał, et al.
Published: (2025)
by: Fica, Michał, et al.
Published: (2025)
Ground-Compose-Reinforce: Grounding Language in Agentic Behaviours using Limited Data
by: Li, Andrew C., et al.
Published: (2025)
by: Li, Andrew C., et al.
Published: (2025)
Good-for-MDP State Reduction for Stochastic LTL Planning
by: Weinhuber, Christoph, et al.
Published: (2025)
by: Weinhuber, Christoph, et al.
Published: (2025)
LTL-Constrained Policy Optimization with Cycle Experience Replay
by: Shah, Ameesh, et al.
Published: (2024)
by: Shah, Ameesh, et al.
Published: (2024)
Real-Time Model Checking for Closed-Loop Robot Reactive Planning
by: Chandler, Christopher, et al.
Published: (2025)
by: Chandler, Christopher, et al.
Published: (2025)
Complex Event Recognition with Symbolic Register Transducers: Extended Technical Report
by: Alevizos, Elias, et al.
Published: (2024)
by: Alevizos, Elias, et al.
Published: (2024)
Computational methods for Dynamic Answer Set Programming
by: Hahn, Susana
Published: (2025)
by: Hahn, Susana
Published: (2025)
Exploring Large Language Models for Access Control Policy Synthesis and Summarization
by: Vatsa, Adarsh, et al.
Published: (2025)
by: Vatsa, Adarsh, et al.
Published: (2025)
Beyond Winning Strategies: Admissible and Admissible Winning Strategies for Quantitative Reachability Games
by: Muvvala, Karan, et al.
Published: (2024)
by: Muvvala, Karan, et al.
Published: (2024)
DFAMiner: Mining minimal separating DFAs from labelled samples
by: Dell'Erba, Daniele, et al.
Published: (2024)
by: Dell'Erba, Daniele, et al.
Published: (2024)
Preprint: Exploring Inevitable Waypoints for Unsolvability Explanation in Hybrid Planning Problems
by: Sarwar, Mir Md Sajid, et al.
Published: (2025)
by: Sarwar, Mir Md Sajid, et al.
Published: (2025)
Neuro-Symbolic Predictive Process Monitoring
by: Mezini, Axel, et al.
Published: (2025)
by: Mezini, Axel, et al.
Published: (2025)
Unsupervised Automata Learning via Discrete Optimization
by: Lutz, Simon, et al.
Published: (2023)
by: Lutz, Simon, et al.
Published: (2023)
Dynamically Reprogrammable Runtime Monitors for Bounded-time MTL
by: Hebballi, Chirantan, et al.
Published: (2026)
by: Hebballi, Chirantan, et al.
Published: (2026)
Digging for Decision Trees: A Case Study in Strategy Sampling and Learning
by: Budde, Carlos E., et al.
Published: (2024)
by: Budde, Carlos E., et al.
Published: (2024)
RE#: High Performance Derivative-Based Regex Matching with Intersection, Complement and Lookarounds
by: Varatalu, Ian Erik, et al.
Published: (2024)
by: Varatalu, Ian Erik, et al.
Published: (2024)
The Inclusion Depth of Pattern Languages: An Open Problem in Algorithmic Learning Theory
by: Luo, Wei
Published: (2026)
by: Luo, Wei
Published: (2026)
Reasoning and Planning with Dynamically Changing Norms
by: Olson, Taylor, et al.
Published: (2026)
by: Olson, Taylor, et al.
Published: (2026)
Policy Cards: Machine-Readable Runtime Governance for Autonomous AI Agents
by: Mavračić, Juraj
Published: (2025)
by: Mavračić, Juraj
Published: (2025)
Asymptotic Hausdorff and Language Similarity
by: Fisman, Dana, et al.
Published: (2026)
by: Fisman, Dana, et al.
Published: (2026)
Evaluating and Learning Robust Bandit Policies Under Uncertain Causal Mechanisms
by: Avery, Katherine, et al.
Published: (2025)
by: Avery, Katherine, et al.
Published: (2025)
Querying Labeled Time Series Data with Scenario Programs
by: Shanker, Devan
Published: (2024)
by: Shanker, Devan
Published: (2024)
Polynomial Bounds of CFLOBDDs against BDDs
by: Zhi, Xusheng, et al.
Published: (2024)
by: Zhi, Xusheng, et al.
Published: (2024)
Uncovering Bugs in Formal Explainers: A Case Study with PyXAI
by: Huang, Xuanxiang, et al.
Published: (2025)
by: Huang, Xuanxiang, et al.
Published: (2025)
Manipulation of Camera Sensor Data via Fault Injection for Anomaly Detection Studies in Verification and Validation Activities For AI
by: Erdogmus, Alim Kerem, et al.
Published: (2021)
by: Erdogmus, Alim Kerem, et al.
Published: (2021)
Tight Verification of Probabilistic Robustness in Bayesian Neural Networks
by: Batten, Ben, et al.
Published: (2024)
by: Batten, Ben, et al.
Published: (2024)
FSFM: A Biologically-Inspired Framework for Selective Forgetting of Agent Memory
by: Gu, Yingjie, et al.
Published: (2026)
by: Gu, Yingjie, et al.
Published: (2026)
Understanding Boolean Function Learnability on Deep Neural Networks: PAC Learning Meets Neurosymbolic Models
by: Nicolau, Marcio, et al.
Published: (2020)
by: Nicolau, Marcio, et al.
Published: (2020)
EmbodiedLGR: Integrating Lightweight Graph Representation and Retrieval for Semantic-Spatial Memory in Robotic Agents
by: Riva, Paolo, et al.
Published: (2026)
by: Riva, Paolo, et al.
Published: (2026)
Few-Shot Learning of a Graph-Based Neural Network Model Without Backpropagation
by: Lapin, Mykyta, et al.
Published: (2025)
by: Lapin, Mykyta, et al.
Published: (2025)
Learning Tree Automata with Term Rewriting
by: Kopystiański, Jakub, et al.
Published: (2026)
by: Kopystiański, Jakub, et al.
Published: (2026)
From Compactifying Lambda-Letrec Terms to Recognizing Regular-Expression Processes
by: Grabmayer, Clemens
Published: (2024)
by: Grabmayer, Clemens
Published: (2024)
Parsing Hypergraphs using Context-Free Positional Grammars
by: Costagliola, Gennaro, et al.
Published: (2026)
by: Costagliola, Gennaro, et al.
Published: (2026)
Various Types of Comet Languages and their Application in External Contextual Grammars
by: Ködding, Marvin, et al.
Published: (2024)
by: Ködding, Marvin, et al.
Published: (2024)
Samyama: A Unified Graph-Vector Database with In-Database Optimization, Agentic Enrichment, and Hardware Acceleration
by: Mandarapu, Madhulatha, et al.
Published: (2026)
by: Mandarapu, Madhulatha, et al.
Published: (2026)
Learning Joint Denoising, Demosaicing, and Compression from the Raw Natural Image Noise Dataset
by: Brummer, Benoit, et al.
Published: (2025)
by: Brummer, Benoit, et al.
Published: (2025)
Scaling Laws for State Dynamics in Large Language Models
by: Li, Jacob X, et al.
Published: (2025)
by: Li, Jacob X, et al.
Published: (2025)
SERA-H: Beyond Native Sentinel Spatial Limits for High-Resolution Canopy Height Mapping
by: Boudras, Thomas, et al.
Published: (2025)
by: Boudras, Thomas, et al.
Published: (2025)
Verification of Unbounded Client-Server Systems with Distinguishable Clients
by: Phawade, Ramchandra, et al.
Published: (2026)
by: Phawade, Ramchandra, et al.
Published: (2026)
Multi-Paradigm Agent Interaction in Practice:A Systematic Analysis of Generator-Evaluator, ReAct Loop,and Adversarial Evaluation in the buddyMe Framework
by: Wang, Xiaohua, et al.
Published: (2026)
by: Wang, Xiaohua, et al.
Published: (2026)
Similar Items
-
Active Automata Learning with Advice
by: Fica, Michał, et al.
Published: (2025) -
Ground-Compose-Reinforce: Grounding Language in Agentic Behaviours using Limited Data
by: Li, Andrew C., et al.
Published: (2025) -
Good-for-MDP State Reduction for Stochastic LTL Planning
by: Weinhuber, Christoph, et al.
Published: (2025) -
LTL-Constrained Policy Optimization with Cycle Experience Replay
by: Shah, Ameesh, et al.
Published: (2024) -
Real-Time Model Checking for Closed-Loop Robot Reactive Planning
by: Chandler, Christopher, et al.
Published: (2025)