Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines
Fuente:
arXiv
Saved in:
| Main Authors: | Levina, Kristina, Pappas, Nikolaos, Karapantelakis, Athanasios, Feljan, Aneta Vulgarakis, Seipp, Jendrik |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Reinforcement Learning with Reward Machines for Sleep Control in Mobile Networks
by: Levina, Kristina, et al.
Published: (2026)
by: Levina, Kristina, et al.
Published: (2026)
Numeric Reward Machines
by: Levina, Kristina, et al.
Published: (2024)
by: Levina, Kristina, et al.
Published: (2024)
A Survey on the Integration of Generative AI for Critical Thinking in Mobile Networks
by: Karapantelakis, Athanasios, et al.
Published: (2024)
by: Karapantelakis, Athanasios, et al.
Published: (2024)
Multi-agent transformer-accelerated RL for satisfaction of STL specifications
by: Forsberg, Albin Larsson, et al.
Published: (2024)
by: Forsberg, Albin Larsson, et al.
Published: (2024)
LLM-Evolved Domain-Independent Heuristics for Symbolic AI Planning
by: Gestrin, Elliot, et al.
Published: (2026)
by: Gestrin, Elliot, et al.
Published: (2026)
Scale When Needed: Adaptive Neuron-level Mixed Precision Quantization Aware Training
by: Varshney, Ayush K., et al.
Published: (2026)
by: Varshney, Ayush K., et al.
Published: (2026)
Consolidating LAMA with Best-First Width Search
by: Corrêa, Augusto B., et al.
Published: (2024)
by: Corrêa, Augusto B., et al.
Published: (2024)
Dynamic Tree Databases in Automated Planning
by: Joergensen, Oliver, et al.
Published: (2025)
by: Joergensen, Oliver, et al.
Published: (2025)
Parallel Lifted Planning via Semi-Naive Datalog Evaluation
by: Drexler, Dominik, et al.
Published: (2026)
by: Drexler, Dominik, et al.
Published: (2026)
NL2Plan: Robust LLM-Driven Planning from Minimal Text Descriptions
by: Gestrin, Elliot, et al.
Published: (2024)
by: Gestrin, Elliot, et al.
Published: (2024)
Symmetry-Aware Transformer Training for Automated Planning
by: Fritzsche, Markus, et al.
Published: (2025)
by: Fritzsche, Markus, et al.
Published: (2025)
LLM-Evolved Pattern Generators for Optimal Classical Planning
by: Phung, Windy, et al.
Published: (2026)
by: Phung, Windy, et al.
Published: (2026)
Classical Planning with LLM-Generated Heuristics: Challenging the State of the Art with Python Code
by: Corrêa, Augusto B., et al.
Published: (2025)
by: Corrêa, Augusto B., et al.
Published: (2025)
Frontier Large Language Models Rival State-of-the-Art Planners
by: Corrêa, Augusto B., et al.
Published: (2025)
by: Corrêa, Augusto B., et al.
Published: (2025)
Property-Guided LLM Program Synthesis for Planning
by: Pereira, André G., et al.
Published: (2026)
by: Pereira, André G., et al.
Published: (2026)
CHD: Coupled Hierarchical Diffusion for Long-Horizon Tasks
by: Hao, Ce, et al.
Published: (2025)
by: Hao, Ce, et al.
Published: (2025)
When to restart? Exploring escalating restarts on convergence
by: Varshney, Ayush K., et al.
Published: (2026)
by: Varshney, Ayush K., et al.
Published: (2026)
RLVMR: Reinforcement Learning with Verifiable Meta-Reasoning Rewards for Robust Long-Horizon Agents
by: Zhang, Zijing, et al.
Published: (2025)
by: Zhang, Zijing, et al.
Published: (2025)
Research and Design on Intelligent Recognition of Unordered Targets for Robots Based on Reinforcement Learning
by: Mao, Yiting, et al.
Published: (2025)
by: Mao, Yiting, et al.
Published: (2025)
The Expressive Leaky Memory Neuron: an Efficient and Expressive Phenomenological Neuron Model Can Solve Long-Horizon Tasks
by: Spieler, Aaron, et al.
Published: (2023)
by: Spieler, Aaron, et al.
Published: (2023)
Heterogeneous Multi-Expert Reinforcement Learning for Long-Horizon Multi-Goal Tasks in Autonomous Forklifts
by: Chen, Yun, et al.
Published: (2026)
by: Chen, Yun, et al.
Published: (2026)
Learning Agent-Compatible Context Management for Long-Horizon Tasks
by: Yi, Lu, et al.
Published: (2026)
by: Yi, Lu, et al.
Published: (2026)
Actor-Critic for Continuous Action Chunks: A Reinforcement Learning Framework for Long-Horizon Robotic Manipulation with Sparse Reward
by: Yang, Jiarui, et al.
Published: (2025)
by: Yang, Jiarui, et al.
Published: (2025)
SkillTree: Explainable Skill-Based Deep Reinforcement Learning for Long-Horizon Control Tasks
by: Wen, Yongyan, et al.
Published: (2024)
by: Wen, Yongyan, et al.
Published: (2024)
Reinforcement Learning with Stochastic Reward Machines
by: Corazza, Jan, et al.
Published: (2025)
by: Corazza, Jan, et al.
Published: (2025)
Reinforcement Learning with Symbolic Reward Machines
by: Krug, Thomas, et al.
Published: (2026)
by: Krug, Thomas, et al.
Published: (2026)
Symmetries-enhanced Multi-Agent Reinforcement Learning
by: Bousias, Nikolaos, et al.
Published: (2025)
by: Bousias, Nikolaos, et al.
Published: (2025)
Reinforcement Learning for Long-Horizon Interactive LLM Agents
by: Chen, Kevin, et al.
Published: (2025)
by: Chen, Kevin, et al.
Published: (2025)
Efficient Reinforcement Learning in Probabilistic Reward Machines
by: Lin, Xiaofeng, et al.
Published: (2024)
by: Lin, Xiaofeng, et al.
Published: (2024)
Pushdown Reward Machines for Reinforcement Learning
by: Varricchione, Giovanni, et al.
Published: (2025)
by: Varricchione, Giovanni, et al.
Published: (2025)
STRUCTUREDAGENT: Planning with AND/OR Trees for Long-Horizon Web Tasks
by: Lobo, ELita, et al.
Published: (2026)
by: Lobo, ELita, et al.
Published: (2026)
Rewarding What Matters: Step-by-Step Reinforcement Learning for Task-Oriented Dialogue
by: Du, Huifang, et al.
Published: (2024)
by: Du, Huifang, et al.
Published: (2024)
Spatially Grounded Long-Horizon Task Planning in the Wild
by: Jung, Sehun, et al.
Published: (2026)
by: Jung, Sehun, et al.
Published: (2026)
SPPO: Sequence-Level PPO for Long-Horizon Reasoning Tasks
by: Wang, Tianyi, et al.
Published: (2026)
by: Wang, Tianyi, et al.
Published: (2026)
ARM: Advantage Reward Modeling for Long-Horizon Manipulation
by: Mao, Yiming, et al.
Published: (2026)
by: Mao, Yiming, et al.
Published: (2026)
Learning Reward Machines in Cooperative Multi-Agent Tasks
by: Ardon, Leo, et al.
Published: (2023)
by: Ardon, Leo, et al.
Published: (2023)
Program Machine Policy: Addressing Long-Horizon Tasks by Integrating Program Synthesis and State Machines
by: Lin, Yu-An, et al.
Published: (2023)
by: Lin, Yu-An, et al.
Published: (2023)
Toward Ultra-Long-Horizon Agentic Science: Cognitive Accumulation for Machine Learning Engineering
by: Zhu, Xinyu, et al.
Published: (2026)
by: Zhu, Xinyu, et al.
Published: (2026)
DETACH: Cross-domain Learning for Long-Horizon Tasks via Mixture of Disentangled Experts
by: Shen, Yutong, et al.
Published: (2025)
by: Shen, Yutong, et al.
Published: (2025)
Learning on the Job: An Experience-Driven Self-Evolving Agent for Long-Horizon Tasks
by: Yang, Cheng, et al.
Published: (2025)
by: Yang, Cheng, et al.
Published: (2025)
Similar Items
-
Reinforcement Learning with Reward Machines for Sleep Control in Mobile Networks
by: Levina, Kristina, et al.
Published: (2026) -
Numeric Reward Machines
by: Levina, Kristina, et al.
Published: (2024) -
A Survey on the Integration of Generative AI for Critical Thinking in Mobile Networks
by: Karapantelakis, Athanasios, et al.
Published: (2024) -
Multi-agent transformer-accelerated RL for satisfaction of STL specifications
by: Forsberg, Albin Larsson, et al.
Published: (2024) -
LLM-Evolved Domain-Independent Heuristics for Symbolic AI Planning
by: Gestrin, Elliot, et al.
Published: (2026)