NAVIX: Scaling MiniGrid Environments with JAX
Fuente:
arXiv
Guardado en:
| Autores principales: | Pignatelli, Eduardo, Liesen, Jarek, Lange, Robert Tjarko, Lu, Chris, Castro, Pablo Samuel, Toni, Laura |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Behaviour Distillation
por: Lupu, Andrei, et al.
Publicado: (2024)
por: Lupu, Andrei, et al.
Publicado: (2024)
XLand-MiniGrid: Scalable Meta-Reinforcement Learning Environments in JAX
por: Nikulin, Alexander, et al.
Publicado: (2023)
por: Nikulin, Alexander, et al.
Publicado: (2023)
The impact of intrinsic rewards on exploration in Reinforcement Learning
por: Kayal, Aya, et al.
Publicado: (2025)
por: Kayal, Aya, et al.
Publicado: (2025)
JaxMARL: Multi-Agent RL Environments and Algorithms in JAX
por: Rutherford, Alexander, et al.
Publicado: (2023)
por: Rutherford, Alexander, et al.
Publicado: (2023)
Assessing the Zero-Shot Capabilities of LLMs for Action Evaluation in RL
por: Pignatelli, Eduardo, et al.
Publicado: (2024)
por: Pignatelli, Eduardo, et al.
Publicado: (2024)
Discovering Minimal Reinforcement Learning Environments
por: Liesen, Jarek, et al.
Publicado: (2024)
por: Liesen, Jarek, et al.
Publicado: (2024)
The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery
por: Lu, Chris, et al.
Publicado: (2024)
por: Lu, Chris, et al.
Publicado: (2024)
Discovering Temporally-Aware Reinforcement Learning Algorithms
por: Jackson, Matthew Thomas, et al.
Publicado: (2024)
por: Jackson, Matthew Thomas, et al.
Publicado: (2024)
Text-to-LoRA: Instant Transformer Adaption
por: Charakorn, Rujikorn, et al.
Publicado: (2025)
por: Charakorn, Rujikorn, et al.
Publicado: (2025)
A Clean Slate for Offline Reinforcement Learning
por: Jackson, Matthew Thomas, et al.
Publicado: (2025)
por: Jackson, Matthew Thomas, et al.
Publicado: (2025)
Large Language Models As Evolution Strategies
por: Lange, Robert Tjarko, et al.
Publicado: (2024)
por: Lange, Robert Tjarko, et al.
Publicado: (2024)
A Survey of Temporal Credit Assignment in Deep Reinforcement Learning
por: Pignatelli, Eduardo, et al.
Publicado: (2023)
por: Pignatelli, Eduardo, et al.
Publicado: (2023)
The AI Scientist-v2: Workshop-Level Automated Scientific Discovery via Agentic Tree Search
por: Yamada, Yutaro, et al.
Publicado: (2025)
por: Yamada, Yutaro, et al.
Publicado: (2025)
CALE: Continuous Arcade Learning Environment
por: Farebrother, Jesse, et al.
Publicado: (2024)
por: Farebrother, Jesse, et al.
Publicado: (2024)
Jumanji: a Diverse Suite of Scalable Reinforcement Learning Environments in JAX
por: Bonnet, Clément, et al.
Publicado: (2023)
por: Bonnet, Clément, et al.
Publicado: (2023)
Imagined Autocurricula
por: Güzel, Ahmet H., et al.
Publicado: (2025)
por: Güzel, Ahmet H., et al.
Publicado: (2025)
Mind the GAP! The Challenges of Scale in Pixel-based Deep Reinforcement Learning
por: Sokar, Ghada, et al.
Publicado: (2025)
por: Sokar, Ghada, et al.
Publicado: (2025)
JaxARC: A High-Performance JAX-based Environment for Abstraction and Reasoning Research
por: Aadam, et al.
Publicado: (2026)
por: Aadam, et al.
Publicado: (2026)
laplax -- Laplace Approximations with JAX
por: Weber, Tobias, et al.
Publicado: (2025)
por: Weber, Tobias, et al.
Publicado: (2025)
Towards Robust Agentic CUDA Kernel Benchmarking, Verification, and Optimization
por: Lange, Robert Tjarko, et al.
Publicado: (2025)
por: Lange, Robert Tjarko, et al.
Publicado: (2025)
CAX: Cellular Automata Accelerated in JAX
por: Faldor, Maxence, et al.
Publicado: (2024)
por: Faldor, Maxence, et al.
Publicado: (2024)
minimax: Efficient Baselines for Autocurricula in JAX
por: Jiang, Minqi, et al.
Publicado: (2023)
por: Jiang, Minqi, et al.
Publicado: (2023)
Scaling Is All You Need: Autonomous Driving with JAX-Accelerated Reinforcement Learning
por: Harmel, Moritz, et al.
Publicado: (2023)
por: Harmel, Moritz, et al.
Publicado: (2023)
Position: Leverage Foundational Models for Black-Box Optimization
por: Song, Xingyou, et al.
Publicado: (2024)
por: Song, Xingyou, et al.
Publicado: (2024)
The Formalism-Implementation Gap in Reinforcement Learning Research
por: Castro, Pablo Samuel
Publicado: (2025)
por: Castro, Pablo Samuel
Publicado: (2025)
PuzzleJAX: A Benchmark for Reasoning and Learning
por: Earle, Sam, et al.
Publicado: (2025)
por: Earle, Sam, et al.
Publicado: (2025)
Learning Bug Context for PyTorch-to-JAX Translation with LLMs
por: Phan, Hung, et al.
Publicado: (2025)
por: Phan, Hung, et al.
Publicado: (2025)
A Survey of State Representation Learning for Deep Reinforcement Learning
por: Echchahed, Ayoub, et al.
Publicado: (2025)
por: Echchahed, Ayoub, et al.
Publicado: (2025)
Mahjax: A GPU-Accelerated Mahjong Simulator for Reinforcement Learning in JAX
por: Nishimori, Soichiro, et al.
Publicado: (2026)
por: Nishimori, Soichiro, et al.
Publicado: (2026)
Multi-Task Reinforcement Learning Enables Parameter Scaling
por: McLean, Reginald, et al.
Publicado: (2025)
por: McLean, Reginald, et al.
Publicado: (2025)
Jasmine: A Simple, Performant and Scalable JAX-based World Modeling Codebase
por: Mahajan, Mihir, et al.
Publicado: (2025)
por: Mahajan, Mihir, et al.
Publicado: (2025)
Comparing BFGS and OGR for Second-Order Optimization
por: Przybysz, Adrian, et al.
Publicado: (2025)
por: Przybysz, Adrian, et al.
Publicado: (2025)
Chargax: A JAX Accelerated EV Charging Simulator
por: Ponse, Koen, et al.
Publicado: (2025)
por: Ponse, Koen, et al.
Publicado: (2025)
$\texttt{MiniMol}$: A Parameter-Efficient Foundation Model for Molecular Learning
por: Kläser, Kerstin, et al.
Publicado: (2024)
por: Kläser, Kerstin, et al.
Publicado: (2024)
ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance
por: Risukhin, Andrey, et al.
Publicado: (2025)
por: Risukhin, Andrey, et al.
Publicado: (2025)
In value-based deep reinforcement learning, a pruned network is a good network
por: Obando-Ceron, Johan, et al.
Publicado: (2024)
por: Obando-Ceron, Johan, et al.
Publicado: (2024)
Heterogeneous Graph Structure Learning through the Lens of Data-generating Processes
por: Jiang, Keyue, et al.
Publicado: (2025)
por: Jiang, Keyue, et al.
Publicado: (2025)
From In Silico to In Vitro: Evaluating Molecule Generative Models for Hit Generation
por: Osman, Nagham, et al.
Publicado: (2025)
por: Osman, Nagham, et al.
Publicado: (2025)
Bures-Wasserstein Flow Matching for Graph Generation
por: Jiang, Keyue, et al.
Publicado: (2025)
por: Jiang, Keyue, et al.
Publicado: (2025)
Mixtures of Experts Unlock Parameter Scaling for Deep RL
por: Obando-Ceron, Johan, et al.
Publicado: (2024)
por: Obando-Ceron, Johan, et al.
Publicado: (2024)
Ejemplares similares
-
Behaviour Distillation
por: Lupu, Andrei, et al.
Publicado: (2024) -
XLand-MiniGrid: Scalable Meta-Reinforcement Learning Environments in JAX
por: Nikulin, Alexander, et al.
Publicado: (2023) -
The impact of intrinsic rewards on exploration in Reinforcement Learning
por: Kayal, Aya, et al.
Publicado: (2025) -
JaxMARL: Multi-Agent RL Environments and Algorithms in JAX
por: Rutherford, Alexander, et al.
Publicado: (2023) -
Assessing the Zero-Shot Capabilities of LLMs for Action Evaluation in RL
por: Pignatelli, Eduardo, et al.
Publicado: (2024)