XLand-MiniGrid: Scalable Meta-Reinforcement Learning Environments in JAX
Fuente:
arXiv
Saved in:
| Main Authors: | Nikulin, Alexander, Kurenkov, Vladislav, Zisman, Ilya, Agarkov, Artem, Sinii, Viacheslav, Kolesnikov, Sergey |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Emergence of In-Context Reinforcement Learning from Noise Distillation
by: Zisman, Ilya, et al.
Published: (2023)
by: Zisman, Ilya, et al.
Published: (2023)
In-Context Reinforcement Learning for Variable Action Spaces
by: Sinii, Viacheslav, et al.
Published: (2023)
by: Sinii, Viacheslav, et al.
Published: (2023)
XLand-100B: A Large-Scale Multi-Task Dataset for In-Context Reinforcement Learning
by: Nikulin, Alexander, et al.
Published: (2024)
by: Nikulin, Alexander, et al.
Published: (2024)
NAVIX: Scaling MiniGrid Environments with JAX
by: Pignatelli, Eduardo, et al.
Published: (2024)
by: Pignatelli, Eduardo, et al.
Published: (2024)
N-Gram Induction Heads for In-Context RL: Improving Stability and Reducing Data Needs
by: Zisman, Ilya, et al.
Published: (2024)
by: Zisman, Ilya, et al.
Published: (2024)
Zero-Shot Adaptation of Behavioral Foundation Models to Unseen Dynamics
by: Bobrin, Maksim, et al.
Published: (2025)
by: Bobrin, Maksim, et al.
Published: (2025)
Vintix: Action Model via In-Context Reinforcement Learning
by: Polubarov, Andrey, et al.
Published: (2025)
by: Polubarov, Andrey, et al.
Published: (2025)
Latent Action Learning Requires Supervision in the Presence of Distractors
by: Nikulin, Alexander, et al.
Published: (2025)
by: Nikulin, Alexander, et al.
Published: (2025)
cadrille: Multi-modal CAD Reconstruction with Reinforcement Learning
by: Kolodiazhnyi, Maksim, et al.
Published: (2025)
by: Kolodiazhnyi, Maksim, et al.
Published: (2025)
gfnx: Fast and Scalable Library for Generative Flow Networks in JAX
by: Tiapkin, Daniil, et al.
Published: (2025)
by: Tiapkin, Daniil, et al.
Published: (2025)
Vision-Language Models Unlock Task-Centric Latent Actions
by: Nikulin, Alexander, et al.
Published: (2026)
by: Nikulin, Alexander, et al.
Published: (2026)
NinA: Normalizing Flows in Action. Training VLA Models with Normalizing Flows
by: Tarasov, Denis, et al.
Published: (2025)
by: Tarasov, Denis, et al.
Published: (2025)
Yes, Q-learning Helps Offline In-Context RL
by: Tarasov, Denis, et al.
Published: (2025)
by: Tarasov, Denis, et al.
Published: (2025)
Vintix II: Decision Pre-Trained Transformer is a Scalable In-Context Reinforcement Learner
by: Polubarov, Andrei, et al.
Published: (2026)
by: Polubarov, Andrei, et al.
Published: (2026)
Jumanji: a Diverse Suite of Scalable Reinforcement Learning Environments in JAX
by: Bonnet, Clément, et al.
Published: (2023)
by: Bonnet, Clément, et al.
Published: (2023)
Enhancing Vision-Language Model Training with Reinforcement Learning in Synthetic Worlds for Real-World Success
by: Bredis, George, et al.
Published: (2025)
by: Bredis, George, et al.
Published: (2025)
Electrostatics from Laplacian Eigenbasis for Neural Network Interatomic Potentials
by: Zhdanov, Maksim, et al.
Published: (2025)
by: Zhdanov, Maksim, et al.
Published: (2025)
Steering LLM Reasoning Through Bias-Only Adaptation
by: Sinii, Viacheslav, et al.
Published: (2025)
by: Sinii, Viacheslav, et al.
Published: (2025)
Object-Centric Latent Action Learning
by: Klepach, Albina, et al.
Published: (2025)
by: Klepach, Albina, et al.
Published: (2025)
Octax: Accelerated CHIP-8 Arcade Environments for Reinforcement Learning in JAX
by: Radji, Waris, et al.
Published: (2025)
by: Radji, Waris, et al.
Published: (2025)
The Differences Between Direct Alignment Algorithms are a Blur
by: Gorbatovski, Alexey, et al.
Published: (2025)
by: Gorbatovski, Alexey, et al.
Published: (2025)
ESSA: Evolutionary Strategies for Scalable Alignment
by: Korotyshova, Daria, et al.
Published: (2025)
by: Korotyshova, Daria, et al.
Published: (2025)
You Do Not Fully Utilize Transformer's Representation Capacity
by: Gerasimov, Gleb, et al.
Published: (2025)
by: Gerasimov, Gleb, et al.
Published: (2025)
DrJAX: Scalable and Differentiable MapReduce Primitives in JAX
by: Rush, Keith, et al.
Published: (2024)
by: Rush, Keith, et al.
Published: (2024)
F-GRPO: Don't Let Your Policy Learn the Obvious and Forget the Rare
by: Plyusov, Daniil, et al.
Published: (2026)
by: Plyusov, Daniil, et al.
Published: (2026)
MPX: Mixed Precision Training for JAX
by: Gräfe, Alexander, et al.
Published: (2025)
by: Gräfe, Alexander, et al.
Published: (2025)
Mini Honor of Kings: A Lightweight Environment for Multi-Agent Reinforcement Learning
by: Liu, Lin, et al.
Published: (2024)
by: Liu, Lin, et al.
Published: (2024)
Mahjax: A GPU-Accelerated Mahjong Simulator for Reinforcement Learning in JAX
by: Nishimori, Soichiro, et al.
Published: (2026)
by: Nishimori, Soichiro, et al.
Published: (2026)
Multistep Quasimetric Learning for Scalable Goal-conditioned Reinforcement Learning
by: Zheng, Bill Chunyuan, et al.
Published: (2025)
by: Zheng, Bill Chunyuan, et al.
Published: (2025)
JaxMARL: Multi-Agent RL Environments and Algorithms in JAX
by: Rutherford, Alexander, et al.
Published: (2023)
by: Rutherford, Alexander, et al.
Published: (2023)
Small Vectors, Big Effects: A Mechanistic Study of RL-Induced Reasoning via Steering Vectors
by: Sinii, Viacheslav, et al.
Published: (2025)
by: Sinii, Viacheslav, et al.
Published: (2025)
Scaling Is All You Need: Autonomous Driving with JAX-Accelerated Reinforcement Learning
by: Harmel, Moritz, et al.
Published: (2023)
by: Harmel, Moritz, et al.
Published: (2023)
BlackJAX: Composable Bayesian inference in JAX
by: Cabezas, Alberto, et al.
Published: (2024)
by: Cabezas, Alberto, et al.
Published: (2024)
UVIP: Model-Free Approach to Evaluate Reinforcement Learning Algorithms
by: Belomestny, Denis, et al.
Published: (2021)
by: Belomestny, Denis, et al.
Published: (2021)
RealStats: A Rigorous Real-Only Statistical Framework for Fake Image Detection
by: Zisman, Haim, et al.
Published: (2026)
by: Zisman, Haim, et al.
Published: (2026)
Jasmine: A Simple, Performant and Scalable JAX-based World Modeling Codebase
by: Mahajan, Mihir, et al.
Published: (2025)
by: Mahajan, Mihir, et al.
Published: (2025)
Scalable Multi-Objective and Meta Reinforcement Learning via Gradient Estimation
by: Zhang, Zhenshuo, et al.
Published: (2025)
by: Zhang, Zhenshuo, et al.
Published: (2025)
PuzzleJAX: A Benchmark for Reasoning and Learning
by: Earle, Sam, et al.
Published: (2025)
by: Earle, Sam, et al.
Published: (2025)
SE(3)-Hyena Operator for Scalable Equivariant Learning
by: Moskalev, Artem, et al.
Published: (2024)
by: Moskalev, Artem, et al.
Published: (2024)
BindGPT: A Scalable Framework for 3D Molecular Design via Language Modeling and Reinforcement Learning
by: Zholus, Artem, et al.
Published: (2024)
by: Zholus, Artem, et al.
Published: (2024)
Similar Items
-
Emergence of In-Context Reinforcement Learning from Noise Distillation
by: Zisman, Ilya, et al.
Published: (2023) -
In-Context Reinforcement Learning for Variable Action Spaces
by: Sinii, Viacheslav, et al.
Published: (2023) -
XLand-100B: A Large-Scale Multi-Task Dataset for In-Context Reinforcement Learning
by: Nikulin, Alexander, et al.
Published: (2024) -
NAVIX: Scaling MiniGrid Environments with JAX
by: Pignatelli, Eduardo, et al.
Published: (2024) -
N-Gram Induction Heads for In-Context RL: Improving Stability and Reducing Data Needs
by: Zisman, Ilya, et al.
Published: (2024)