An Investigation of Offline Reinforcement Learning in Factorisable Action Spaces
Fuente:
arXiv
Guardado en:
| Autores principales: | Beeson, Alex, Ireland, David, Montana, Giovanni |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
REValueD: Regularised Ensemble Value-Decomposition for Factorisable Markov Decision Processes
por: Ireland, David, et al.
Publicado: (2024)
por: Ireland, David, et al.
Publicado: (2024)
Goal-conditioned Offline Reinforcement Learning through State Space Partitioning
por: Wang, Mianchu, et al.
Publicado: (2023)
por: Wang, Mianchu, et al.
Publicado: (2023)
Learning Partial Action Replacement in Offline MARL
por: Jin, Yue, et al.
Publicado: (2026)
por: Jin, Yue, et al.
Publicado: (2026)
Learning on One Mode: Addressing Multi-modality in Offline Reinforcement Learning
por: Wang, Mianchu, et al.
Publicado: (2024)
por: Wang, Mianchu, et al.
Publicado: (2024)
State-Constrained Offline Reinforcement Learning
por: Hepburn, Charles A., et al.
Publicado: (2024)
por: Hepburn, Charles A., et al.
Publicado: (2024)
Partial Action Replacement: Tackling Distribution Shift in Offline MARL
por: Jin, Yue, et al.
Publicado: (2025)
por: Jin, Yue, et al.
Publicado: (2025)
Evaluation-Time Policy Switching for Offline Reinforcement Learning
por: Neggatu, Natinael Solomon, et al.
Publicado: (2025)
por: Neggatu, Natinael Solomon, et al.
Publicado: (2025)
Action-Free Offline-to-Online RL via Discretised State Policies
por: Neggatu, Natinael Solomon, et al.
Publicado: (2026)
por: Neggatu, Natinael Solomon, et al.
Publicado: (2026)
GOPlan: Goal-conditioned Offline Reinforcement Learning by Planning with Learned Models
por: Wang, Mianchu, et al.
Publicado: (2023)
por: Wang, Mianchu, et al.
Publicado: (2023)
BraVE: Offline Reinforcement Learning for Discrete Combinatorial Action Spaces
por: Landers, Matthew, et al.
Publicado: (2024)
por: Landers, Matthew, et al.
Publicado: (2024)
Flow Matching for Offline Reinforcement Learning with Discrete Actions
por: Khan, Fairoz Nower, et al.
Publicado: (2026)
por: Khan, Fairoz Nower, et al.
Publicado: (2026)
Offline Reinforcement Learning with Penalized Action Noise Injection
por: Oh, JunHyeok, et al.
Publicado: (2025)
por: Oh, JunHyeok, et al.
Publicado: (2025)
Achieving Collective Welfare in Multi-Agent Reinforcement Learning via Suggestion Sharing
por: Jin, Yue, et al.
Publicado: (2024)
por: Jin, Yue, et al.
Publicado: (2024)
Mitigating Relative Over-Generalization in Multi-Agent Reinforcement Learning
por: Zhu, Ting, et al.
Publicado: (2024)
por: Zhu, Ting, et al.
Publicado: (2024)
Taming OOD Actions for Offline Reinforcement Learning: An Advantage-Based Approach
por: Chen, Xuyang, et al.
Publicado: (2025)
por: Chen, Xuyang, et al.
Publicado: (2025)
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction
por: Durkin, Alex, et al.
Publicado: (2025)
por: Durkin, Alex, et al.
Publicado: (2025)
Two-Step Offline Preference-Based Reinforcement Learning with Constrained Actions
por: Xu, Yinglun, et al.
Publicado: (2023)
por: Xu, Yinglun, et al.
Publicado: (2023)
Preferred-Action-Optimized Diffusion Policies for Offline Reinforcement Learning
por: Zhang, Tianle, et al.
Publicado: (2024)
por: Zhang, Tianle, et al.
Publicado: (2024)
Investigating Relational State Abstraction in Collaborative MARL
por: Utke, Sharlin, et al.
Publicado: (2024)
por: Utke, Sharlin, et al.
Publicado: (2024)
SAMG: Offline-to-Online Reinforcement Learning via State-Action-Conditional Offline Model Guidance
por: Zhang, Liyu, et al.
Publicado: (2024)
por: Zhang, Liyu, et al.
Publicado: (2024)
Equivariant Offline Reinforcement Learning
por: Tangri, Arsh, et al.
Publicado: (2024)
por: Tangri, Arsh, et al.
Publicado: (2024)
Offline Reinforcement Learning with OOD State Correction and OOD Action Suppression
por: Mao, Yixiu, et al.
Publicado: (2024)
por: Mao, Yixiu, et al.
Publicado: (2024)
Constrained Latent Action Policies for Model-Based Offline Reinforcement Learning
por: Alles, Marvin, et al.
Publicado: (2024)
por: Alles, Marvin, et al.
Publicado: (2024)
Penalizing Infeasible Actions and Reward Scaling in Reinforcement Learning with Offline Data
por: Kim, Jeonghye, et al.
Publicado: (2025)
por: Kim, Jeonghye, et al.
Publicado: (2025)
Robustness Evaluation of Offline Reinforcement Learning for Robot Control Against Action Perturbations
por: Ayabe, Shingo, et al.
Publicado: (2024)
por: Ayabe, Shingo, et al.
Publicado: (2024)
Uncertainty-Based Smooth Policy Regularisation for Reinforcement Learning with Few Demonstrations
por: Zhu, Yujie, et al.
Publicado: (2025)
por: Zhu, Yujie, et al.
Publicado: (2025)
The Signal in the Noise: OOD Detection Through Goodness-of-Fit Testing in Factorised Latent Spaces
por: Bomatter, Philipp, et al.
Publicado: (2026)
por: Bomatter, Philipp, et al.
Publicado: (2026)
DAWM: Diffusion Action World Models for Offline Reinforcement Learning via Action-Inferred Transitions
por: Li, Zongyue, et al.
Publicado: (2025)
por: Li, Zongyue, et al.
Publicado: (2025)
Offline Reinforcement Learning in Large State Spaces: Algorithms and Guarantees
por: Jiang, Nan, et al.
Publicado: (2025)
por: Jiang, Nan, et al.
Publicado: (2025)
Disentanglement as Identifiable Pushforward Factorisation
por: Allen, Carl
Publicado: (2024)
por: Allen, Carl
Publicado: (2024)
Guided Flow Policy: Learning from High-Value Actions in Offline Reinforcement Learning
por: Tiofack, Franki Nguimatsia, et al.
Publicado: (2025)
por: Tiofack, Franki Nguimatsia, et al.
Publicado: (2025)
Proximal Action Replacement for Behavior Cloning Actor-Critic in Offline Reinforcement Learning
por: Dong, Jinzong, et al.
Publicado: (2026)
por: Dong, Jinzong, et al.
Publicado: (2026)
Switching the Loss Reduces the Cost in Batch (Offline) Reinforcement Learning
por: Ayoub, Alex, et al.
Publicado: (2024)
por: Ayoub, Alex, et al.
Publicado: (2024)
Beyond Non-Expert Demonstrations: Outcome-Driven Action Constraint for Offline Reinforcement Learning
por: Jiang, Ke, et al.
Publicado: (2025)
por: Jiang, Ke, et al.
Publicado: (2025)
Offline Trajectory Optimization for Offline Reinforcement Learning
por: Zhao, Ziqi, et al.
Publicado: (2024)
por: Zhao, Ziqi, et al.
Publicado: (2024)
Investigating Action Encodings in Recurrent Neural Networks in Reinforcement Learning
por: Schlegel, Matthew, et al.
Publicado: (2026)
por: Schlegel, Matthew, et al.
Publicado: (2026)
Federated Offline Reinforcement Learning
por: Zhou, Doudou, et al.
Publicado: (2022)
por: Zhou, Doudou, et al.
Publicado: (2022)
Epistemic Robust Offline Reinforcement Learning
por: Chenreddy, Abhilash Reddy, et al.
Publicado: (2026)
por: Chenreddy, Abhilash Reddy, et al.
Publicado: (2026)
In-Context Reinforcement Learning for Variable Action Spaces
por: Sinii, Viacheslav, et al.
Publicado: (2023)
por: Sinii, Viacheslav, et al.
Publicado: (2023)
Offline Multitask Representation Learning for Reinforcement Learning
por: Ishfaq, Haque, et al.
Publicado: (2024)
por: Ishfaq, Haque, et al.
Publicado: (2024)
Ejemplares similares
-
REValueD: Regularised Ensemble Value-Decomposition for Factorisable Markov Decision Processes
por: Ireland, David, et al.
Publicado: (2024) -
Goal-conditioned Offline Reinforcement Learning through State Space Partitioning
por: Wang, Mianchu, et al.
Publicado: (2023) -
Learning Partial Action Replacement in Offline MARL
por: Jin, Yue, et al.
Publicado: (2026) -
Learning on One Mode: Addressing Multi-modality in Offline Reinforcement Learning
por: Wang, Mianchu, et al.
Publicado: (2024) -
State-Constrained Offline Reinforcement Learning
por: Hepburn, Charles A., et al.
Publicado: (2024)