ARM-FM: Automated Reward Machines via Foundation Models for Compositional Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Castanyer, Roger Creus, Mohamed, Faisal, Castro, Pablo Samuel, Neary, Cyrus, Berseth, Glen |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Surprise-Adaptive Intrinsic Motivation for Unsupervised Reinforcement Learning
by: Hugessen, Adriana, et al.
Published: (2024)
by: Hugessen, Adriana, et al.
Published: (2024)
Agentick: A Unified Benchmark for General Sequential Decision-Making Agents
by: Castanyer, Roger Creus, et al.
Published: (2026)
by: Castanyer, Roger Creus, et al.
Published: (2026)
Improving Intrinsic Exploration by Creating Stationary Objectives
by: Castanyer, Roger Creus, et al.
Published: (2023)
by: Castanyer, Roger Creus, et al.
Published: (2023)
Align and Filter: Improving Performance in Asynchronous On-Policy RL
by: Honari, Homayoun, et al.
Published: (2026)
by: Honari, Homayoun, et al.
Published: (2026)
Temporal Representations for Exploration: Learning Complex Exploratory Behavior without Extrinsic Rewards
by: Mohamed, Faisal, et al.
Published: (2026)
by: Mohamed, Faisal, et al.
Published: (2026)
Improving Pre-Trained Vision-Language-Action Policies with Model-Based Search
by: Neary, Cyrus, et al.
Published: (2025)
by: Neary, Cyrus, et al.
Published: (2025)
Is Exploration or Optimization the Problem for Deep Reinforcement Learning?
by: Berseth, Glen
Published: (2025)
by: Berseth, Glen
Published: (2025)
Improving Deep Reinforcement Learning by Reducing the Chain Effect of Value and Policy Churn
by: Tang, Hongyao, et al.
Published: (2024)
by: Tang, Hongyao, et al.
Published: (2024)
Mitigating Plasticity Loss in Continual Reinforcement Learning by Reducing Churn
by: Tang, Hongyao, et al.
Published: (2025)
by: Tang, Hongyao, et al.
Published: (2025)
SegDAC: Visual Generalization in Reinforcement Learning via Dynamic Object Tokens
by: Brown, Alexandre, et al.
Published: (2025)
by: Brown, Alexandre, et al.
Published: (2025)
Stable Gradients for Stable Learning at Scale in Deep Reinforcement Learning
by: Castanyer, Roger Creus, et al.
Published: (2025)
by: Castanyer, Roger Creus, et al.
Published: (2025)
RLeXplore: Accelerating Research in Intrinsically-Motivated Reinforcement Learning
by: Yuan, Mingqi, et al.
Published: (2024)
by: Yuan, Mingqi, et al.
Published: (2024)
Neural Port-Hamiltonian Differential Algebraic Equations for Compositional Learning of Electrical Networks
by: Neary, Cyrus, et al.
Published: (2024)
by: Neary, Cyrus, et al.
Published: (2024)
unix-ctf: Procedural Environments for Unix-Competence Reinforcement Learning
by: Bradway, Geoffrey, et al.
Published: (2026)
by: Bradway, Geoffrey, et al.
Published: (2026)
Enabling Realtime Reinforcement Learning at Scale with Staggered Asynchronous Inference
by: Riemer, Matthew, et al.
Published: (2024)
by: Riemer, Matthew, et al.
Published: (2024)
xgenius: LLM-Oriented Autonomous Research Platform for SLURM Clusters
by: Creus Castanyer, Roger
Published: (2026)
by: Creus Castanyer, Roger
Published: (2026)
Multimodal Pretrained Models for Verifiable Sequential Decision-Making: Planning, Grounding, and Perception
by: Yang, Yunhao, et al.
Published: (2023)
by: Yang, Yunhao, et al.
Published: (2023)
Non-Adversarial Inverse Reinforcement Learning via Successor Feature Matching
by: Jain, Arnav Kumar, et al.
Published: (2024)
by: Jain, Arnav Kumar, et al.
Published: (2024)
V-VLAPS: Value-Guided Planning for Vision-Language-Action Models
by: Ren, Ke, et al.
Published: (2026)
by: Ren, Ke, et al.
Published: (2026)
PopuLoRA: Co-Evolving LLM Populations for Reasoning Self-Play
by: Castanyer, Roger Creus, et al.
Published: (2026)
by: Castanyer, Roger Creus, et al.
Published: (2026)
BuilderBench: The Building Blocks of Intelligent Agents
by: Ghugare, Raj, et al.
Published: (2025)
by: Ghugare, Raj, et al.
Published: (2025)
Reinforcement Learning for Versatile, Dynamic, and Robust Bipedal Locomotion Control
by: Li, Zhongyu, et al.
Published: (2024)
by: Li, Zhongyu, et al.
Published: (2024)
The Formalism-Implementation Gap in Reinforcement Learning Research
by: Castro, Pablo Samuel
Published: (2025)
by: Castro, Pablo Samuel
Published: (2025)
Using Large Language Models to Automate and Expedite Reinforcement Learning with Reward Machine
by: Alsadat, Shayan Meshkat, et al.
Published: (2024)
by: Alsadat, Shayan Meshkat, et al.
Published: (2024)
Intelligent Switching for Reset-Free RL
by: Patil, Darshan, et al.
Published: (2024)
by: Patil, Darshan, et al.
Published: (2024)
Enhancing Agent Learning through World Dynamics Modeling
by: Sun, Zhiyuan, et al.
Published: (2024)
by: Sun, Zhiyuan, et al.
Published: (2024)
ARM: Advantage Reward Modeling for Long-Horizon Manipulation
by: Mao, Yiming, et al.
Published: (2026)
by: Mao, Yiming, et al.
Published: (2026)
Reinforcement Learning with Symbolic Reward Machines
by: Krug, Thomas, et al.
Published: (2026)
by: Krug, Thomas, et al.
Published: (2026)
Reinforcement Learning with Stochastic Reward Machines
by: Corazza, Jan, et al.
Published: (2025)
by: Corazza, Jan, et al.
Published: (2025)
A Survey of State Representation Learning for Deep Reinforcement Learning
by: Echchahed, Ayoub, et al.
Published: (2025)
by: Echchahed, Ayoub, et al.
Published: (2025)
Self-Predictive Representations for Combinatorial Generalization in Behavioral Cloning
by: Lawson, Daniel, et al.
Published: (2025)
by: Lawson, Daniel, et al.
Published: (2025)
Efficient Reinforcement Learning in Probabilistic Reward Machines
by: Lin, Xiaofeng, et al.
Published: (2024)
by: Lin, Xiaofeng, et al.
Published: (2024)
SpectrumFM: Redefining Spectrum Cognition via Foundation Modeling
by: Liu, Chunyu, et al.
Published: (2025)
by: Liu, Chunyu, et al.
Published: (2025)
Mind the GAP! The Challenges of Scale in Pixel-based Deep Reinforcement Learning
by: Sokar, Ghada, et al.
Published: (2025)
by: Sokar, Ghada, et al.
Published: (2025)
Pushdown Reward Machines for Reinforcement Learning
by: Varricchione, Giovanni, et al.
Published: (2025)
by: Varricchione, Giovanni, et al.
Published: (2025)
Residual Reward Models for Preference-based Reinforcement Learning
by: Cao, Chenyang, et al.
Published: (2025)
by: Cao, Chenyang, et al.
Published: (2025)
DISCOVER: Automated Curricula for Sparse-Reward Reinforcement Learning
by: Diaz-Bone, Leander, et al.
Published: (2025)
by: Diaz-Bone, Leander, et al.
Published: (2025)
DomusFM: A Foundation Model for Smart-Home Sensor Data
by: Fiori, Michele, et al.
Published: (2026)
by: Fiori, Michele, et al.
Published: (2026)
BoostAPR: Boosting Automated Program Repair via Execution-Grounded Reinforcement Learning with Dual Reward Models
by: Li, Yuanhao, et al.
Published: (2026)
by: Li, Yuanhao, et al.
Published: (2026)
Human-centric Reward Optimization for Reinforcement Learning-based Automated Driving using Large Language Models
by: Zhou, Ziqi, et al.
Published: (2024)
by: Zhou, Ziqi, et al.
Published: (2024)
Similar Items
-
Surprise-Adaptive Intrinsic Motivation for Unsupervised Reinforcement Learning
by: Hugessen, Adriana, et al.
Published: (2024) -
Agentick: A Unified Benchmark for General Sequential Decision-Making Agents
by: Castanyer, Roger Creus, et al.
Published: (2026) -
Improving Intrinsic Exploration by Creating Stationary Objectives
by: Castanyer, Roger Creus, et al.
Published: (2023) -
Align and Filter: Improving Performance in Asynchronous On-Policy RL
by: Honari, Homayoun, et al.
Published: (2026) -
Temporal Representations for Exploration: Learning Complex Exploratory Behavior without Extrinsic Rewards
by: Mohamed, Faisal, et al.
Published: (2026)