CALE: Continuous Arcade Learning Environment
Fuente:
arXiv
Guardado en:
| Autores principales: | Farebrother, Jesse, Castro, Pablo Samuel |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Non-Adversarial Inverse Reinforcement Learning via Successor Feature Matching
por: Jain, Arnav Kumar, et al.
Publicado: (2024)
por: Jain, Arnav Kumar, et al.
Publicado: (2024)
Mixtures of Experts Unlock Parameter Scaling for Deep RL
por: Obando-Ceron, Johan, et al.
Publicado: (2024)
por: Obando-Ceron, Johan, et al.
Publicado: (2024)
Temporal Difference Flows
por: Farebrother, Jesse, et al.
Publicado: (2025)
por: Farebrother, Jesse, et al.
Publicado: (2025)
Compositional Planning with Jumpy World Models
por: Farebrother, Jesse, et al.
Publicado: (2026)
por: Farebrother, Jesse, et al.
Publicado: (2026)
The Formalism-Implementation Gap in Reinforcement Learning Research
por: Castro, Pablo Samuel
Publicado: (2025)
por: Castro, Pablo Samuel
Publicado: (2025)
Investigating Memory in Model-Free RL with POPGym Arcade
por: Wang, Zekang, et al.
Publicado: (2025)
por: Wang, Zekang, et al.
Publicado: (2025)
Stop Regressing: Training Value Functions via Classification for Scalable Deep RL
por: Farebrother, Jesse, et al.
Publicado: (2024)
por: Farebrother, Jesse, et al.
Publicado: (2024)
A Distributional Analogue to the Successor Representation
por: Wiltzer, Harley, et al.
Publicado: (2024)
por: Wiltzer, Harley, et al.
Publicado: (2024)
A Survey of State Representation Learning for Deep Reinforcement Learning
por: Echchahed, Ayoub, et al.
Publicado: (2025)
por: Echchahed, Ayoub, et al.
Publicado: (2025)
Mitigating Plasticity Loss in Continual Reinforcement Learning by Reducing Churn
por: Tang, Hongyao, et al.
Publicado: (2025)
por: Tang, Hongyao, et al.
Publicado: (2025)
Mind the GAP! The Challenges of Scale in Pixel-based Deep Reinforcement Learning
por: Sokar, Ghada, et al.
Publicado: (2025)
por: Sokar, Ghada, et al.
Publicado: (2025)
NAVIX: Scaling MiniGrid Environments with JAX
por: Pignatelli, Eduardo, et al.
Publicado: (2024)
por: Pignatelli, Eduardo, et al.
Publicado: (2024)
CLeAN: Continual Learning Adaptive Normalization in Dynamic Environments
por: Marasco, Isabella, et al.
Publicado: (2026)
por: Marasco, Isabella, et al.
Publicado: (2026)
HackAtari: Atari Learning Environments for Robust and Continual Reinforcement Learning
por: Delfosse, Quentin, et al.
Publicado: (2024)
por: Delfosse, Quentin, et al.
Publicado: (2024)
The Impact of On-Policy Parallelized Data Collection on Deep Reinforcement Learning Networks
por: Mayor, Walter, et al.
Publicado: (2025)
por: Mayor, Walter, et al.
Publicado: (2025)
Continual Learning for Adaptable Car-Following in Dynamic Traffic Environments
por: Chen, Xianda, et al.
Publicado: (2024)
por: Chen, Xianda, et al.
Publicado: (2024)
Natural Mitigation of Catastrophic Interference: Continual Learning in Power-Law Learning Environments
por: Gandhi, Atith, et al.
Publicado: (2024)
por: Gandhi, Atith, et al.
Publicado: (2024)
Explorative Imitation Learning: A Path Signature Approach for Continuous Environments
por: Gavenski, Nathan, et al.
Publicado: (2024)
por: Gavenski, Nathan, et al.
Publicado: (2024)
ARM-FM: Automated Reward Machines via Foundation Models for Compositional Reinforcement Learning
por: Castanyer, Roger Creus, et al.
Publicado: (2025)
por: Castanyer, Roger Creus, et al.
Publicado: (2025)
Continual Reinforcement Learning via Autoencoder-Driven Task and New Environment Recognition
por: Erden, Zeki Doruk, et al.
Publicado: (2025)
por: Erden, Zeki Doruk, et al.
Publicado: (2025)
Stable Deep Reinforcement Learning via Isotropic Gaussian Representations
por: Pasand, Ali Saheb, et al.
Publicado: (2026)
por: Pasand, Ali Saheb, et al.
Publicado: (2026)
In value-based deep reinforcement learning, a pruned network is a good network
por: Obando-Ceron, Johan, et al.
Publicado: (2024)
por: Obando-Ceron, Johan, et al.
Publicado: (2024)
Studying the Interplay Between the Actor and Critic Representations in Reinforcement Learning
por: Garcin, Samuel, et al.
Publicado: (2025)
por: Garcin, Samuel, et al.
Publicado: (2025)
On Sequential Bayesian Inference for Continual Learning
por: Kessler, Samuel, et al.
Publicado: (2023)
por: Kessler, Samuel, et al.
Publicado: (2023)
Multi-Task Reinforcement Learning Enables Parameter Scaling
por: McLean, Reginald, et al.
Publicado: (2025)
por: McLean, Reginald, et al.
Publicado: (2025)
Continual Learning in Vision-Language Models via Aligned Model Merging
por: Sokar, Ghada, et al.
Publicado: (2025)
por: Sokar, Ghada, et al.
Publicado: (2025)
Action Mapping for Reinforcement Learning in Continuous Environments with Constraints
por: Theile, Mirco, et al.
Publicado: (2024)
por: Theile, Mirco, et al.
Publicado: (2024)
DRED: Zero-Shot Transfer in Reinforcement Learning via Data-Regularised Environment Design
por: Garcin, Samuel, et al.
Publicado: (2024)
por: Garcin, Samuel, et al.
Publicado: (2024)
Offline Reinforcement Learning for Learning to Dispatch for Job Shop Scheduling
por: van Remmerden, Jesse, et al.
Publicado: (2024)
por: van Remmerden, Jesse, et al.
Publicado: (2024)
CRAFT: Forgetting-Aware Intervention-Based Adaptation for Continual Learning
por: Hossen, Md Anwar, et al.
Publicado: (2026)
por: Hossen, Md Anwar, et al.
Publicado: (2026)
Generalizing Beyond Suboptimality: Offline Reinforcement Learning Learns Effective Scheduling through Random Data
por: van Remmerden, Jesse, et al.
Publicado: (2025)
por: van Remmerden, Jesse, et al.
Publicado: (2025)
ED2: Environment Dynamics Decomposition World Models for Continuous Control
por: Hao, Jianye, et al.
Publicado: (2021)
por: Hao, Jianye, et al.
Publicado: (2021)
On the consistency of hyper-parameter selection in value-based deep reinforcement learning
por: Obando-Ceron, Johan, et al.
Publicado: (2024)
por: Obando-Ceron, Johan, et al.
Publicado: (2024)
Don't flatten, tokenize! Unlocking the key to SoftMoE's efficacy in deep RL
por: Sokar, Ghada, et al.
Publicado: (2024)
por: Sokar, Ghada, et al.
Publicado: (2024)
Mixture of Experts in a Mixture of RL settings
por: Willi, Timon, et al.
Publicado: (2024)
por: Willi, Timon, et al.
Publicado: (2024)
Learning to Compile Programs to Neural Networks
por: Weber, Logan, et al.
Publicado: (2024)
por: Weber, Logan, et al.
Publicado: (2024)
Leveraging Approximate Model-based Shielding for Probabilistic Safety Guarantees in Continuous Environments
por: Goodall, Alexander W., et al.
Publicado: (2024)
por: Goodall, Alexander W., et al.
Publicado: (2024)
STEP: Learning STructured Embeddings for Progressive Time Series
por: Thil, Lucas, et al.
Publicado: (2026)
por: Thil, Lucas, et al.
Publicado: (2026)
Refining Minimax Regret for Unsupervised Environment Design
por: Beukman, Michael, et al.
Publicado: (2024)
por: Beukman, Michael, et al.
Publicado: (2024)
Scaling Algorithm Distillation for Continuous Control with Mamba
por: Beaussant, Samuel, et al.
Publicado: (2025)
por: Beaussant, Samuel, et al.
Publicado: (2025)
Ejemplares similares
-
Non-Adversarial Inverse Reinforcement Learning via Successor Feature Matching
por: Jain, Arnav Kumar, et al.
Publicado: (2024) -
Mixtures of Experts Unlock Parameter Scaling for Deep RL
por: Obando-Ceron, Johan, et al.
Publicado: (2024) -
Temporal Difference Flows
por: Farebrother, Jesse, et al.
Publicado: (2025) -
Compositional Planning with Jumpy World Models
por: Farebrother, Jesse, et al.
Publicado: (2026) -
The Formalism-Implementation Gap in Reinforcement Learning Research
por: Castro, Pablo Samuel
Publicado: (2025)