Reset-free Reinforcement Learning with World Models
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Zhao, Moerland, Thomas M., Preuss, Mike, Plaat, Aske, Hu, Edward S. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Guiding Skill Discovery with Foundation Models
by: Yang, Zhao, et al.
Published: (2025)
by: Yang, Zhao, et al.
Published: (2025)
Towards a Practical Understanding of Lagrangian Methods in Safe Reinforcement Learning
by: Spoor, Lindsay, et al.
Published: (2025)
by: Spoor, Lindsay, et al.
Published: (2025)
Explicitly Disentangled Representations in Object-Centric Learning
by: Majellaro, Riccardo, et al.
Published: (2024)
by: Majellaro, Riccardo, et al.
Published: (2024)
Research Re: search & Re-search
by: Plaat, Aske
Published: (2024)
by: Plaat, Aske
Published: (2024)
Chargax: A JAX Accelerated EV Charging Simulator
by: Ponse, Koen, et al.
Published: (2025)
by: Ponse, Koen, et al.
Published: (2025)
Reinforcement Learning for Sustainable Energy: A Survey
by: Ponse, Koen, et al.
Published: (2024)
by: Ponse, Koen, et al.
Published: (2024)
Slot Structured World Models
by: Collu, Jonathan, et al.
Published: (2024)
by: Collu, Jonathan, et al.
Published: (2024)
EduGym: An Environment and Notebook Suite for Reinforcement Learning Education
by: Moerland, Thomas M., et al.
Published: (2023)
by: Moerland, Thomas M., et al.
Published: (2023)
Solving Deep Reinforcement Learning Tasks with Evolution Strategies and Linear Policy Networks
by: Wong, Annie, et al.
Published: (2024)
by: Wong, Annie, et al.
Published: (2024)
Agentic Large Language Models, a survey
by: Plaat, Aske, et al.
Published: (2025)
by: Plaat, Aske, et al.
Published: (2025)
Analysis of Bluffing by DQN and CFR in Leduc Hold'em Poker
by: Zaciragic, Tarik, et al.
Published: (2025)
by: Zaciragic, Tarik, et al.
Published: (2025)
Mirror Mode in Fire Emblem: Beating Players at their own Game with Imitation and Reinforcement Learning
by: Smid, Yanna Elizabeth, et al.
Published: (2025)
by: Smid, Yanna Elizabeth, et al.
Published: (2025)
A Unified Framework for Zero-Shot Reinforcement Learning
by: Di Ventura, Jacopo, et al.
Published: (2025)
by: Di Ventura, Jacopo, et al.
Published: (2025)
CORE: Towards Scalable and Efficient Causal Discovery with Reinforcement Learning
by: Sauter, Andreas W. M., et al.
Published: (2024)
by: Sauter, Andreas W. M., et al.
Published: (2024)
On the Effect of Regularization in Policy Mirror Descent
by: Kleuker, Jan Felix, et al.
Published: (2025)
by: Kleuker, Jan Felix, et al.
Published: (2025)
How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding
by: Chen, Xi, et al.
Published: (2025)
by: Chen, Xi, et al.
Published: (2025)
Baba is LLM: Reasoning in a Game with Dynamic Rules
by: van Wetten, Fien, et al.
Published: (2025)
by: van Wetten, Fien, et al.
Published: (2025)
CausalPlayground: Addressing Data-Generation Requirements in Cutting-Edge Causality Research
by: Sauter, Andreas W M, et al.
Published: (2024)
by: Sauter, Andreas W M, et al.
Published: (2024)
A Benchmark Study of Deep Reinforcement Learning Algorithms for the Container Stowage Planning Problem
by: Huang, Yunqi, et al.
Published: (2025)
by: Huang, Yunqi, et al.
Published: (2025)
ACTIVA: Amortized Causal Effect Estimation via Transformer-based Variational Autoencoder
by: Sauter, Andreas, et al.
Published: (2025)
by: Sauter, Andreas, et al.
Published: (2025)
Reinforcement learning for Quantum Tiq-Taq-Toe
by: Dinu, Catalin-Viorel, et al.
Published: (2024)
by: Dinu, Catalin-Viorel, et al.
Published: (2024)
CoComposer: LLM Multi-agent Collaborative Music Composition
by: Xing, Peiwen, et al.
Published: (2025)
by: Xing, Peiwen, et al.
Published: (2025)
Reasoning Capabilities of Large Language Models on Dynamic Tasks
by: Wong, Annie, et al.
Published: (2025)
by: Wong, Annie, et al.
Published: (2025)
State Design Matters: How Representations Shape Dynamic Reasoning in Large Language Models
by: Wong, Annie, et al.
Published: (2026)
by: Wong, Annie, et al.
Published: (2026)
Multi-Step Reasoning with Large Language Models, a Survey
by: Plaat, Aske, et al.
Published: (2024)
by: Plaat, Aske, et al.
Published: (2024)
EconoJax: A Fast & Scalable Economic Simulation in Jax
by: Ponse, Koen, et al.
Published: (2024)
by: Ponse, Koen, et al.
Published: (2024)
Assessing Reproducibility in Evolutionary Computation: A Case Study using Human- and LLM-based Assessment
by: Da Ros, Francesca, et al.
Published: (2026)
by: Da Ros, Francesca, et al.
Published: (2026)
The Power of Resets in Online Reinforcement Learning
by: Mhammedi, Zakaria, et al.
Published: (2024)
by: Mhammedi, Zakaria, et al.
Published: (2024)
A Hybrid Intelligence Method for Argument Mining
by: van der Meer, Michiel, et al.
Published: (2024)
by: van der Meer, Michiel, et al.
Published: (2024)
Reset & Distill: A Recipe for Overcoming Negative Transfer in Continual Reinforcement Learning
by: Ahn, Hongjoon, et al.
Published: (2024)
by: Ahn, Hongjoon, et al.
Published: (2024)
Diverse Feature Learning by Self-distillation and Reset
by: Park, Sejik
Published: (2024)
by: Park, Sejik
Published: (2024)
A KL-regularization Framework for Learning to Plan with Adaptive Priors
by: Serra-Gomez, Álvaro, et al.
Published: (2025)
by: Serra-Gomez, Álvaro, et al.
Published: (2025)
Credit Assignment with Resets in Language Model Reasoning
by: Samanta, Ankur, et al.
Published: (2026)
by: Samanta, Ankur, et al.
Published: (2026)
A Reinforcement Learning based Reset Policy for CDCL SAT Solvers
by: Li, Chunxiao, et al.
Published: (2024)
by: Li, Chunxiao, et al.
Published: (2024)
Single-Reset Divide & Conquer Imitation Learning
by: Chenu, Alexandre, et al.
Published: (2024)
by: Chenu, Alexandre, et al.
Published: (2024)
World-Gymnast: Training Robots with Reinforcement Learning in a World Model
by: Sharma, Ansh Kumar, et al.
Published: (2026)
by: Sharma, Ansh Kumar, et al.
Published: (2026)
Real-World Reinforcement Learning of Active Perception Behaviors
by: Hu, Edward S., et al.
Published: (2025)
by: Hu, Edward S., et al.
Published: (2025)
Self-Normalized Resets for Plasticity in Continual Learning
by: Farias, Vivek F., et al.
Published: (2024)
by: Farias, Vivek F., et al.
Published: (2024)
RLVR-World: Training World Models with Reinforcement Learning
by: Wu, Jialong, et al.
Published: (2025)
by: Wu, Jialong, et al.
Published: (2025)
Explainable Reinforcement Learning Agents Using World Models
by: Singh, Madhuri, et al.
Published: (2025)
by: Singh, Madhuri, et al.
Published: (2025)
Similar Items
-
Guiding Skill Discovery with Foundation Models
by: Yang, Zhao, et al.
Published: (2025) -
Towards a Practical Understanding of Lagrangian Methods in Safe Reinforcement Learning
by: Spoor, Lindsay, et al.
Published: (2025) -
Explicitly Disentangled Representations in Object-Centric Learning
by: Majellaro, Riccardo, et al.
Published: (2024) -
Research Re: search & Re-search
by: Plaat, Aske
Published: (2024) -
Chargax: A JAX Accelerated EV Charging Simulator
by: Ponse, Koen, et al.
Published: (2025)