Numeric Reward Machines
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Levina, Kristina, Pappas, Nikolaos, Karapantelakis, Athanasios, Feljan, Aneta Vulgarakis, Seipp, Jendrik |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Reinforcement Learning with Reward Machines for Sleep Control in Mobile Networks
von: Levina, Kristina, et al.
Veröffentlicht: (2026)
von: Levina, Kristina, et al.
Veröffentlicht: (2026)
Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines
von: Levina, Kristina, et al.
Veröffentlicht: (2025)
von: Levina, Kristina, et al.
Veröffentlicht: (2025)
Scale When Needed: Adaptive Neuron-level Mixed Precision Quantization Aware Training
von: Varshney, Ayush K., et al.
Veröffentlicht: (2026)
von: Varshney, Ayush K., et al.
Veröffentlicht: (2026)
Symmetry-Aware Transformer Training for Automated Planning
von: Fritzsche, Markus, et al.
Veröffentlicht: (2025)
von: Fritzsche, Markus, et al.
Veröffentlicht: (2025)
A Survey on the Integration of Generative AI for Critical Thinking in Mobile Networks
von: Karapantelakis, Athanasios, et al.
Veröffentlicht: (2024)
von: Karapantelakis, Athanasios, et al.
Veröffentlicht: (2024)
Classical Planning with LLM-Generated Heuristics: Challenging the State of the Art with Python Code
von: Corrêa, Augusto B., et al.
Veröffentlicht: (2025)
von: Corrêa, Augusto B., et al.
Veröffentlicht: (2025)
Frontier Large Language Models Rival State-of-the-Art Planners
von: Corrêa, Augusto B., et al.
Veröffentlicht: (2025)
von: Corrêa, Augusto B., et al.
Veröffentlicht: (2025)
Property-Guided LLM Program Synthesis for Planning
von: Pereira, André G., et al.
Veröffentlicht: (2026)
von: Pereira, André G., et al.
Veröffentlicht: (2026)
Multi-agent transformer-accelerated RL for satisfaction of STL specifications
von: Forsberg, Albin Larsson, et al.
Veröffentlicht: (2024)
von: Forsberg, Albin Larsson, et al.
Veröffentlicht: (2024)
When to restart? Exploring escalating restarts on convergence
von: Varshney, Ayush K., et al.
Veröffentlicht: (2026)
von: Varshney, Ayush K., et al.
Veröffentlicht: (2026)
LLM-Evolved Domain-Independent Heuristics for Symbolic AI Planning
von: Gestrin, Elliot, et al.
Veröffentlicht: (2026)
von: Gestrin, Elliot, et al.
Veröffentlicht: (2026)
Consolidating LAMA with Best-First Width Search
von: Corrêa, Augusto B., et al.
Veröffentlicht: (2024)
von: Corrêa, Augusto B., et al.
Veröffentlicht: (2024)
NL2Plan: Robust LLM-Driven Planning from Minimal Text Descriptions
von: Gestrin, Elliot, et al.
Veröffentlicht: (2024)
von: Gestrin, Elliot, et al.
Veröffentlicht: (2024)
Parallel Lifted Planning via Semi-Naive Datalog Evaluation
von: Drexler, Dominik, et al.
Veröffentlicht: (2026)
von: Drexler, Dominik, et al.
Veröffentlicht: (2026)
Dynamic Tree Databases in Automated Planning
von: Joergensen, Oliver, et al.
Veröffentlicht: (2025)
von: Joergensen, Oliver, et al.
Veröffentlicht: (2025)
Neural Reward Machines
von: Umili, Elena, et al.
Veröffentlicht: (2024)
von: Umili, Elena, et al.
Veröffentlicht: (2024)
LLM-Evolved Pattern Generators for Optimal Classical Planning
von: Phung, Windy, et al.
Veröffentlicht: (2026)
von: Phung, Windy, et al.
Veröffentlicht: (2026)
Reinforcement Learning with Symbolic Reward Machines
von: Krug, Thomas, et al.
Veröffentlicht: (2026)
von: Krug, Thomas, et al.
Veröffentlicht: (2026)
Reinforcement Learning with Stochastic Reward Machines
von: Corazza, Jan, et al.
Veröffentlicht: (2025)
von: Corazza, Jan, et al.
Veröffentlicht: (2025)
Fully Learnable Neural Reward Machines
von: Dewidar, Hazem, et al.
Veröffentlicht: (2025)
von: Dewidar, Hazem, et al.
Veröffentlicht: (2025)
Diffusion-Driven Semantic Communication for Generative Models with Bandwidth Constraints
von: Guo, Lei, et al.
Veröffentlicht: (2024)
von: Guo, Lei, et al.
Veröffentlicht: (2024)
Efficient Reinforcement Learning in Probabilistic Reward Machines
von: Lin, Xiaofeng, et al.
Veröffentlicht: (2024)
von: Lin, Xiaofeng, et al.
Veröffentlicht: (2024)
Semantic Text Transmission via Prediction with Small Language Models: Cost-Similarity Trade-off
von: Madhabhavi, Bhavani A, et al.
Veröffentlicht: (2024)
von: Madhabhavi, Bhavani A, et al.
Veröffentlicht: (2024)
Learning Robust Reward Machines from Noisy Labels
von: Parac, Roko, et al.
Veröffentlicht: (2024)
von: Parac, Roko, et al.
Veröffentlicht: (2024)
Provably Efficient Exploration in Reward Machines with Low Regret
von: Bourel, Hippolyte, et al.
Veröffentlicht: (2024)
von: Bourel, Hippolyte, et al.
Veröffentlicht: (2024)
Maximally Permissive Reward Machines
von: Varricchione, Giovanni, et al.
Veröffentlicht: (2024)
von: Varricchione, Giovanni, et al.
Veröffentlicht: (2024)
Beyond Numeric Rewards: In-Context Dueling Bandits with LLM Agents
von: Xia, Fanzeng, et al.
Veröffentlicht: (2024)
von: Xia, Fanzeng, et al.
Veröffentlicht: (2024)
Learning to Recover: Dynamic Reward Shaping with Wheel-Leg Coordination for Fallen Robots
von: Deng, Boyuan, et al.
Veröffentlicht: (2025)
von: Deng, Boyuan, et al.
Veröffentlicht: (2025)
Pushdown Reward Machines for Reinforcement Learning
von: Varricchione, Giovanni, et al.
Veröffentlicht: (2025)
von: Varricchione, Giovanni, et al.
Veröffentlicht: (2025)
Symmetries-enhanced Multi-Agent Reinforcement Learning
von: Bousias, Nikolaos, et al.
Veröffentlicht: (2025)
von: Bousias, Nikolaos, et al.
Veröffentlicht: (2025)
Conformal Prediction with Learned Features
von: Kiyani, Shayan, et al.
Veröffentlicht: (2024)
von: Kiyani, Shayan, et al.
Veröffentlicht: (2024)
Contextual Pre-planning on Reward Machine Abstractions for Enhanced Transfer in Deep Reinforcement Learning
von: Azran, Guy, et al.
Veröffentlicht: (2023)
von: Azran, Guy, et al.
Veröffentlicht: (2023)
ARM-FM: Automated Reward Machines via Foundation Models for Compositional Reinforcement Learning
von: Castanyer, Roger Creus, et al.
Veröffentlicht: (2025)
von: Castanyer, Roger Creus, et al.
Veröffentlicht: (2025)
Evaluating Supervised Machine Learning Models: Principles, Pitfalls, and Metric Selection
von: Liu, Xuanyan, et al.
Veröffentlicht: (2026)
von: Liu, Xuanyan, et al.
Veröffentlicht: (2026)
Attention-Based Reward Shaping for Sparse and Delayed Rewards
von: Holmes, Ian, et al.
Veröffentlicht: (2025)
von: Holmes, Ian, et al.
Veröffentlicht: (2025)
Reward Hacking Mitigation using Verifiable Composite Rewards
von: Tarek, Mirza Farhan Bin, et al.
Veröffentlicht: (2025)
von: Tarek, Mirza Farhan Bin, et al.
Veröffentlicht: (2025)
Repairing Reward Functions with Feedback to Mitigate Reward Hacking
von: Hatgis-Kessell, Stephane, et al.
Veröffentlicht: (2025)
von: Hatgis-Kessell, Stephane, et al.
Veröffentlicht: (2025)
Intrinsic Reward Policy Optimization for Sparse-Reward Environments
von: Cho, Minjae, et al.
Veröffentlicht: (2026)
von: Cho, Minjae, et al.
Veröffentlicht: (2026)
Reward Centering
von: Naik, Abhishek, et al.
Veröffentlicht: (2024)
von: Naik, Abhishek, et al.
Veröffentlicht: (2024)
Adversarial Reward Auditing for Active Detection and Mitigation of Reward Hacking
von: Beigi, Mohammad, et al.
Veröffentlicht: (2026)
von: Beigi, Mohammad, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Reinforcement Learning with Reward Machines for Sleep Control in Mobile Networks
von: Levina, Kristina, et al.
Veröffentlicht: (2026) -
Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines
von: Levina, Kristina, et al.
Veröffentlicht: (2025) -
Scale When Needed: Adaptive Neuron-level Mixed Precision Quantization Aware Training
von: Varshney, Ayush K., et al.
Veröffentlicht: (2026) -
Symmetry-Aware Transformer Training for Automated Planning
von: Fritzsche, Markus, et al.
Veröffentlicht: (2025) -
A Survey on the Integration of Generative AI for Critical Thinking in Mobile Networks
von: Karapantelakis, Athanasios, et al.
Veröffentlicht: (2024)