MARL-GPT: Foundation Model for Multi-Agent Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Nesterova, Maria, Kolosov, Mikhail, Andreychuk, Anton, Cherepanov, Egor, Bulichev, Oleg, Kovalev, Alexey, Yakovlev, Konstantin, Panov, Aleksandr, Skrynnik, Alexey |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MAPF-GPT: Imitation Learning for Multi-Agent Pathfinding at Scale
von: Andreychuk, Anton, et al.
Veröffentlicht: (2024)
von: Andreychuk, Anton, et al.
Veröffentlicht: (2024)
Advancing Learnable Multi-Agent Pathfinding Solvers with Active Fine-Tuning
von: Andreychuk, Anton, et al.
Veröffentlicht: (2025)
von: Andreychuk, Anton, et al.
Veröffentlicht: (2025)
Learning to Communicate Locally for Large-Scale Multi-Agent Pathfinding
von: Vyaltsev, Valeriy, et al.
Veröffentlicht: (2026)
von: Vyaltsev, Valeriy, et al.
Veröffentlicht: (2026)
POGEMA: A Benchmark Platform for Cooperative Multi-Agent Pathfinding
von: Skrynnik, Alexey, et al.
Veröffentlicht: (2024)
von: Skrynnik, Alexey, et al.
Veröffentlicht: (2024)
Memory Retention Is Not Enough to Master Memory Tasks in Reinforcement Learning
von: Shchendrigin, Oleg, et al.
Veröffentlicht: (2026)
von: Shchendrigin, Oleg, et al.
Veröffentlicht: (2026)
KAGE-Bench: Fast Known-Axis Visual Generalization Evaluation for Reinforcement Learning
von: Cherepanov, Egor, et al.
Veröffentlicht: (2026)
von: Cherepanov, Egor, et al.
Veröffentlicht: (2026)
Memory, Benchmark & Robots: A Benchmark for Solving Complex Tasks with Reinforcement Learning
von: Cherepanov, Egor, et al.
Veröffentlicht: (2025)
von: Cherepanov, Egor, et al.
Veröffentlicht: (2025)
ELMUR: External Layer Memory with Update/Rewrite for Long-Horizon RL Problems
von: Cherepanov, Egor, et al.
Veröffentlicht: (2025)
von: Cherepanov, Egor, et al.
Veröffentlicht: (2025)
Recurrent Action Transformer with Memory
von: Cherepanov, Egor, et al.
Veröffentlicht: (2023)
von: Cherepanov, Egor, et al.
Veröffentlicht: (2023)
CAMAR: Continuous Actions Multi-Agent Routing
von: Pshenitsyn, Artem, et al.
Veröffentlicht: (2025)
von: Pshenitsyn, Artem, et al.
Veröffentlicht: (2025)
CoRL-MPPI: Enhancing MPPI With Learnable Behaviours For Efficient And Provably-Safe Multi-Robot Collision Avoidance
von: Dergachev, Stepan, et al.
Veröffentlicht: (2025)
von: Dergachev, Stepan, et al.
Veröffentlicht: (2025)
Re:Frame -- Retrieving Experience From Associative Memory
von: Zelezetsky, Daniil, et al.
Veröffentlicht: (2025)
von: Zelezetsky, Daniil, et al.
Veröffentlicht: (2025)
Unraveling the Complexity of Memory in RL Agents: an Approach for Classification and Evaluation
von: Cherepanov, Egor, et al.
Veröffentlicht: (2024)
von: Cherepanov, Egor, et al.
Veröffentlicht: (2024)
Instruction Following with Goal-Conditioned Reinforcement Learning in Virtual Environments
von: Volovikova, Zoya, et al.
Veröffentlicht: (2024)
von: Volovikova, Zoya, et al.
Veröffentlicht: (2024)
A New Perspective on Transformers in Online Reinforcement Learning for Continuous Control
von: Kachaev, Nikita, et al.
Veröffentlicht: (2025)
von: Kachaev, Nikita, et al.
Veröffentlicht: (2025)
Don't Blind Your VLA: Aligning Visual Representations for OOD Generalization
von: Kachaev, Nikita, et al.
Veröffentlicht: (2025)
von: Kachaev, Nikita, et al.
Veröffentlicht: (2025)
Self-Guided Plan Extraction for Instruction-Following Tasks with Goal-Conditional Reinforcement Learning
von: Volovikova, Zoya, et al.
Veröffentlicht: (2026)
von: Volovikova, Zoya, et al.
Veröffentlicht: (2026)
Revisiting Tree Search for LLMs: Gumbel and Sequential Halving for Budget-Scalable Reasoning
von: Ugadiarov, Leonid, et al.
Veröffentlicht: (2026)
von: Ugadiarov, Leonid, et al.
Veröffentlicht: (2026)
Optimal and Bounded Suboptimal Any-Angle Multi-agent Pathfinding
von: Yakovlev, Konstantin, et al.
Veröffentlicht: (2024)
von: Yakovlev, Konstantin, et al.
Veröffentlicht: (2024)
Enhancing PIBT via Multi-Action Operations
von: Yukhnevich, Egor, et al.
Veröffentlicht: (2025)
von: Yukhnevich, Egor, et al.
Veröffentlicht: (2025)
VerifyLLM: LLM-Based Pre-Execution Task Plan Verification for Robots
von: Grigorev, Danil S., et al.
Veröffentlicht: (2025)
von: Grigorev, Danil S., et al.
Veröffentlicht: (2025)
Accelerating Transformers in Online RL
von: Zelezetsky, Daniil, et al.
Veröffentlicht: (2025)
von: Zelezetsky, Daniil, et al.
Veröffentlicht: (2025)
Symbolic Disentangled Representations for Images
von: Korchemnyi, Alexandr, et al.
Veröffentlicht: (2024)
von: Korchemnyi, Alexandr, et al.
Veröffentlicht: (2024)
Spatial Traces: Enhancing VLA Models with Spatial-Temporal Understanding
von: Patratskiy, Maxim A., et al.
Veröffentlicht: (2025)
von: Patratskiy, Maxim A., et al.
Veröffentlicht: (2025)
CrafText Benchmark: Advancing Instruction Following in Complex Multimodal Open-Ended World
von: Volovikova, Zoya, et al.
Veröffentlicht: (2025)
von: Volovikova, Zoya, et al.
Veröffentlicht: (2025)
Object-Centric Learning with Slot Mixture Module
von: Kirilenko, Daniil, et al.
Veröffentlicht: (2023)
von: Kirilenko, Daniil, et al.
Veröffentlicht: (2023)
LookPlanGraph: Embodied Instruction Following Method with VLM Graph Augmentation
von: Onishchenko, Anatoly O., et al.
Veröffentlicht: (2025)
von: Onishchenko, Anatoly O., et al.
Veröffentlicht: (2025)
Spectroscopic orbits for three SB2s and one hierarchical triple using SALT data
von: Kovalev, Mikhail, et al.
Veröffentlicht: (2026)
von: Kovalev, Mikhail, et al.
Veröffentlicht: (2026)
Gradual Optimization Learning for Conformational Energy Minimization
von: Tsypin, Artem, et al.
Veröffentlicht: (2023)
von: Tsypin, Artem, et al.
Veröffentlicht: (2023)
Safe Policy Exploration Improvement via Subgoals
von: Angulo, Brian, et al.
Veröffentlicht: (2024)
von: Angulo, Brian, et al.
Veröffentlicht: (2024)
AmbiK: Dataset of Ambiguous Tasks in Kitchen Environment
von: Ivanova, Anastasiia, et al.
Veröffentlicht: (2025)
von: Ivanova, Anastasiia, et al.
Veröffentlicht: (2025)
PRISM-Loc: a Lightweight Long-range LiDAR Localization in Urban Environments with Topological Maps
von: Muravyev, Kirill, et al.
Veröffentlicht: (2025)
von: Muravyev, Kirill, et al.
Veröffentlicht: (2025)
Integrated Pipeline for Coronary Angiography With Automated Lesion Profiling, Virtual Stenting, and 100-Vessel FFR Validation
von: Kopanitsa, Georgy, et al.
Veröffentlicht: (2025)
von: Kopanitsa, Georgy, et al.
Veröffentlicht: (2025)
Model-based Policy Optimization using Symbolic World Model
von: Gorodetskiy, Andrey, et al.
Veröffentlicht: (2024)
von: Gorodetskiy, Andrey, et al.
Veröffentlicht: (2024)
Multi-Agent Path Finding For Large Agents Is Intractable
von: Agafonov, Artem, et al.
Veröffentlicht: (2025)
von: Agafonov, Artem, et al.
Veröffentlicht: (2025)
HELP: Hierarchical Embodied Language Planner for Household Tasks
von: Korchemnyi, Alexandr V., et al.
Veröffentlicht: (2025)
von: Korchemnyi, Alexandr V., et al.
Veröffentlicht: (2025)
Decentralized Uncertainty-Aware Multi-Agent Collision Avoidance with Model Predictive Path Integral
von: Dergachev, Stepan, et al.
Veröffentlicht: (2025)
von: Dergachev, Stepan, et al.
Veröffentlicht: (2025)
BenchMARL: Benchmarking Multi-Agent Reinforcement Learning
von: Bettini, Matteo, et al.
Veröffentlicht: (2023)
von: Bettini, Matteo, et al.
Veröffentlicht: (2023)
Decentralized Unlabeled Multi-Agent Navigation in Continuous Space
von: Dergachev, Stepan, et al.
Veröffentlicht: (2024)
von: Dergachev, Stepan, et al.
Veröffentlicht: (2024)
LangMARL: Natural Language Multi-Agent Reinforcement Learning
von: Yao, Huaiyuan, et al.
Veröffentlicht: (2026)
von: Yao, Huaiyuan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
MAPF-GPT: Imitation Learning for Multi-Agent Pathfinding at Scale
von: Andreychuk, Anton, et al.
Veröffentlicht: (2024) -
Advancing Learnable Multi-Agent Pathfinding Solvers with Active Fine-Tuning
von: Andreychuk, Anton, et al.
Veröffentlicht: (2025) -
Learning to Communicate Locally for Large-Scale Multi-Agent Pathfinding
von: Vyaltsev, Valeriy, et al.
Veröffentlicht: (2026) -
POGEMA: A Benchmark Platform for Cooperative Multi-Agent Pathfinding
von: Skrynnik, Alexey, et al.
Veröffentlicht: (2024) -
Memory Retention Is Not Enough to Master Memory Tasks in Reinforcement Learning
von: Shchendrigin, Oleg, et al.
Veröffentlicht: (2026)