Breaking the Grid: Distance-Guided Reinforcement Learning in Large Discrete Action Spaces
Fuente:
arXiv
Saved in:
| Main Authors: | Hoppe, Heiko, Akkerman, Fabian, van Heeswijk, Wouter, Schiffer, Maximilian |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Dynamic Neighborhood Construction for Structured Large Discrete Action Spaces
by: Akkerman, Fabian, et al.
Published: (2023)
by: Akkerman, Fabian, et al.
Published: (2023)
Target-Aligned Reinforcement Learning
by: Pleiss, Leonard S., et al.
Published: (2026)
by: Pleiss, Leonard S., et al.
Published: (2026)
Learning Dynamic Selection and Pricing of Out-of-Home Deliveries
by: Akkerman, Fabian, et al.
Published: (2023)
by: Akkerman, Fabian, et al.
Published: (2023)
Discretizing Continuous Action Space with Unimodal Probability Distributions for On-Policy Reinforcement Learning
by: Zhu, Yuanyang, et al.
Published: (2024)
by: Zhu, Yuanyang, et al.
Published: (2024)
Solving Dual Sourcing Problems with Supply Mode Dependent Failure Rates
by: Akkerman, Fabian, et al.
Published: (2024)
by: Akkerman, Fabian, et al.
Published: (2024)
Decoupled Hierarchical Reinforcement Learning with State Abstraction for Discrete Grids
by: Xiao, Qingyu, et al.
Published: (2025)
by: Xiao, Qingyu, et al.
Published: (2025)
Guided Flow Policy: Learning from High-Value Actions in Offline Reinforcement Learning
by: Tiofack, Franki Nguimatsia, et al.
Published: (2025)
by: Tiofack, Franki Nguimatsia, et al.
Published: (2025)
An Advantage-based Optimization Method for Reinforcement Learning in Large Action Space
by: Lin, Hai, et al.
Published: (2024)
by: Lin, Hai, et al.
Published: (2024)
Constrained Latent Action Policies for Model-Based Offline Reinforcement Learning
by: Alles, Marvin, et al.
Published: (2024)
by: Alles, Marvin, et al.
Published: (2024)
Global Rewards in Multi-Agent Deep Reinforcement Learning for Autonomous Mobility on Demand Systems
by: Hoppe, Heiko, et al.
Published: (2023)
by: Hoppe, Heiko, et al.
Published: (2023)
In-Context Reinforcement Learning for Variable Action Spaces
by: Sinii, Viacheslav, et al.
Published: (2023)
by: Sinii, Viacheslav, et al.
Published: (2023)
Stochastic Q-learning for Large Discrete Action Spaces
by: Fourati, Fares, et al.
Published: (2024)
by: Fourati, Fares, et al.
Published: (2024)
Structured Reinforcement Learning for Combinatorial Decision-Making
by: Hoppe, Heiko, et al.
Published: (2025)
by: Hoppe, Heiko, et al.
Published: (2025)
Model-based Reinforcement Learning for Parameterized Action Spaces
by: Zhang, Renhao, et al.
Published: (2024)
by: Zhang, Renhao, et al.
Published: (2024)
On the Geometry of Reinforcement Learning in Continuous State and Action Spaces
by: Tiwari, Saket, et al.
Published: (2022)
by: Tiwari, Saket, et al.
Published: (2022)
Neural Cluster First, Route Second: One-Shot Capacitated Vehicle Routing via Differentiable Optimal Transport
by: Chin, Samuel J. K., et al.
Published: (2026)
by: Chin, Samuel J. K., et al.
Published: (2026)
Learning The Minimum Action Distance
by: Steccanella, Lorenzo, et al.
Published: (2025)
by: Steccanella, Lorenzo, et al.
Published: (2025)
State-Novelty Guided Action Persistence in Deep Reinforcement Learning
by: Hu, Jianshu, et al.
Published: (2024)
by: Hu, Jianshu, et al.
Published: (2024)
Geometry of Neural Reinforcement Learning in Continuous State and Action Spaces
by: Tiwari, Saket, et al.
Published: (2025)
by: Tiwari, Saket, et al.
Published: (2025)
Non-Linear Reinforcement Learning in Large Action Spaces: Structural Conditions and Sample-efficiency of Posterior Sampling
by: Agarwal, Alekh, et al.
Published: (2022)
by: Agarwal, Alekh, et al.
Published: (2022)
Model-Based Reinforcement Learning in Discrete-Action Non-Markovian Reward Decision Processes
by: Trapasso, Alessandro, et al.
Published: (2025)
by: Trapasso, Alessandro, et al.
Published: (2025)
SAINT: Attention-Based Policies for Discrete Combinatorial Action Spaces
by: Landers, Matthew, et al.
Published: (2025)
by: Landers, Matthew, et al.
Published: (2025)
Reinforcement Learning with LLM-Guided Action Spaces for Synthesizable Lead Optimization
by: Li, Tao, et al.
Published: (2026)
by: Li, Tao, et al.
Published: (2026)
Transfer Reinforcement Learning in Heterogeneous Action Spaces using Subgoal Mapping
by: Sivakumar, Kavinayan P., et al.
Published: (2024)
by: Sivakumar, Kavinayan P., et al.
Published: (2024)
Robust Deep Reinforcement Learning with Adaptive Adversarial Perturbations in Action Space
by: Liu, Qianmei, et al.
Published: (2024)
by: Liu, Qianmei, et al.
Published: (2024)
REACT: Revealing Evolutionary Action Consequence Trajectories for Interpretable Reinforcement Learning
by: Altmann, Philipp, et al.
Published: (2024)
by: Altmann, Philipp, et al.
Published: (2024)
Bayesian Off-Policy Evaluation and Learning for Large Action Spaces
by: Aouali, Imad, et al.
Published: (2024)
by: Aouali, Imad, et al.
Published: (2024)
Adapting the Behavior of Reinforcement Learning Agents to Changing Action Spaces and Reward Functions
by: de la Rosa, Raul, et al.
Published: (2026)
by: de la Rosa, Raul, et al.
Published: (2026)
RL2Grid: Benchmarking Reinforcement Learning in Power Grid Operations
by: Marchesini, Enrico, et al.
Published: (2025)
by: Marchesini, Enrico, et al.
Published: (2025)
Preference-aware compensation policies for crowdsourced on-demand services
by: Nouli, Georgina, et al.
Published: (2025)
by: Nouli, Georgina, et al.
Published: (2025)
Surrogate Fitness Metrics for Interpretable Reinforcement Learning
by: Altmann, Philipp, et al.
Published: (2025)
by: Altmann, Philipp, et al.
Published: (2025)
Offline Reinforcement Learning in Large State Spaces: Algorithms and Guarantees
by: Jiang, Nan, et al.
Published: (2025)
by: Jiang, Nan, et al.
Published: (2025)
CAMEL: Continuous Action Masking Enabled by Large Language Models for Reinforcement Learning
by: Zhao, Yanxiao, et al.
Published: (2025)
by: Zhao, Yanxiao, et al.
Published: (2025)
Harnessing Discrete Representations For Continual Reinforcement Learning
by: Meyer, Edan, et al.
Published: (2023)
by: Meyer, Edan, et al.
Published: (2023)
Distributionally Robust Model-based Reinforcement Learning with Large State Spaces
by: Ramesh, Shyam Sundhar, et al.
Published: (2023)
by: Ramesh, Shyam Sundhar, et al.
Published: (2023)
Grid-Mapping Pseudo-Count Constraint for Offline Reinforcement Learning
by: Shen, Yi, et al.
Published: (2024)
by: Shen, Yi, et al.
Published: (2024)
Optimizing Power Grid Topologies with Reinforcement Learning: A Survey of Methods and Challenges
by: van der Sar, Erica, et al.
Published: (2025)
by: van der Sar, Erica, et al.
Published: (2025)
Discrete Flow Matching for Offline-to-Online Reinforcement Learning
by: Khan, Fairoz Nower, et al.
Published: (2026)
by: Khan, Fairoz Nower, et al.
Published: (2026)
Reinforcement Learning with Action Chunking
by: Li, Qiyang, et al.
Published: (2025)
by: Li, Qiyang, et al.
Published: (2025)
Curriculum-Guided Antifragile Reinforcement Learning for Secure UAV Deconfliction under Observation-Space Attacks
by: Panda, Deepak Kumar, et al.
Published: (2025)
by: Panda, Deepak Kumar, et al.
Published: (2025)
Similar Items
-
Dynamic Neighborhood Construction for Structured Large Discrete Action Spaces
by: Akkerman, Fabian, et al.
Published: (2023) -
Target-Aligned Reinforcement Learning
by: Pleiss, Leonard S., et al.
Published: (2026) -
Learning Dynamic Selection and Pricing of Out-of-Home Deliveries
by: Akkerman, Fabian, et al.
Published: (2023) -
Discretizing Continuous Action Space with Unimodal Probability Distributions for On-Policy Reinforcement Learning
by: Zhu, Yuanyang, et al.
Published: (2024) -
Solving Dual Sourcing Problems with Supply Mode Dependent Failure Rates
by: Akkerman, Fabian, et al.
Published: (2024)