Unraveling the Complexity of Memory in RL Agents: an Approach for Classification and Evaluation
Fuente:
arXiv
Saved in:
| Main Authors: | Cherepanov, Egor, Kachaev, Nikita, Zholus, Artem, Kovalev, Alexey K., Panov, Aleksandr I. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Memory, Benchmark & Robots: A Benchmark for Solving Complex Tasks with Reinforcement Learning
by: Cherepanov, Egor, et al.
Published: (2025)
by: Cherepanov, Egor, et al.
Published: (2025)
ELMUR: External Layer Memory with Update/Rewrite for Long-Horizon RL Problems
by: Cherepanov, Egor, et al.
Published: (2025)
by: Cherepanov, Egor, et al.
Published: (2025)
Recurrent Action Transformer with Memory
by: Cherepanov, Egor, et al.
Published: (2023)
by: Cherepanov, Egor, et al.
Published: (2023)
Memory Retention Is Not Enough to Master Memory Tasks in Reinforcement Learning
by: Shchendrigin, Oleg, et al.
Published: (2026)
by: Shchendrigin, Oleg, et al.
Published: (2026)
A New Perspective on Transformers in Online Reinforcement Learning for Continuous Control
by: Kachaev, Nikita, et al.
Published: (2025)
by: Kachaev, Nikita, et al.
Published: (2025)
Re:Frame -- Retrieving Experience From Associative Memory
by: Zelezetsky, Daniil, et al.
Published: (2025)
by: Zelezetsky, Daniil, et al.
Published: (2025)
KAGE-Bench: Fast Known-Axis Visual Generalization Evaluation for Reinforcement Learning
by: Cherepanov, Egor, et al.
Published: (2026)
by: Cherepanov, Egor, et al.
Published: (2026)
Don't Blind Your VLA: Aligning Visual Representations for OOD Generalization
by: Kachaev, Nikita, et al.
Published: (2025)
by: Kachaev, Nikita, et al.
Published: (2025)
CAMAR: Continuous Actions Multi-Agent Routing
by: Pshenitsyn, Artem, et al.
Published: (2025)
by: Pshenitsyn, Artem, et al.
Published: (2025)
LookPlanGraph: Embodied Instruction Following Method with VLM Graph Augmentation
by: Onishchenko, Anatoly O., et al.
Published: (2025)
by: Onishchenko, Anatoly O., et al.
Published: (2025)
Symbolic Disentangled Representations for Images
by: Korchemnyi, Alexandr, et al.
Published: (2024)
by: Korchemnyi, Alexandr, et al.
Published: (2024)
Accelerating Transformers in Online RL
by: Zelezetsky, Daniil, et al.
Published: (2025)
by: Zelezetsky, Daniil, et al.
Published: (2025)
Object-Centric Learning with Slot Mixture Module
by: Kirilenko, Daniil, et al.
Published: (2023)
by: Kirilenko, Daniil, et al.
Published: (2023)
AmbiK: Dataset of Ambiguous Tasks in Kitchen Environment
by: Ivanova, Anastasiia, et al.
Published: (2025)
by: Ivanova, Anastasiia, et al.
Published: (2025)
Steering LLM Reasoning Through Bias-Only Adaptation
by: Sinii, Viacheslav, et al.
Published: (2025)
by: Sinii, Viacheslav, et al.
Published: (2025)
MARL-GPT: Foundation Model for Multi-Agent Reinforcement Learning
by: Nesterova, Maria, et al.
Published: (2026)
by: Nesterova, Maria, et al.
Published: (2026)
HELP: Hierarchical Embodied Language Planner for Household Tasks
by: Korchemnyi, Alexandr V., et al.
Published: (2025)
by: Korchemnyi, Alexandr V., et al.
Published: (2025)
MAPF-GPT: Imitation Learning for Multi-Agent Pathfinding at Scale
by: Andreychuk, Anton, et al.
Published: (2024)
by: Andreychuk, Anton, et al.
Published: (2024)
Advancing Learnable Multi-Agent Pathfinding Solvers with Active Fine-Tuning
by: Andreychuk, Anton, et al.
Published: (2025)
by: Andreychuk, Anton, et al.
Published: (2025)
Revisiting Tree Search for LLMs: Gumbel and Sequential Halving for Budget-Scalable Reasoning
by: Ugadiarov, Leonid, et al.
Published: (2026)
by: Ugadiarov, Leonid, et al.
Published: (2026)
IDAT: A Multi-Modal Dataset and Toolkit for Building and Evaluating Interactive Task-Solving Agents
by: Mohanty, Shrestha, et al.
Published: (2024)
by: Mohanty, Shrestha, et al.
Published: (2024)
Learning Successor Features with Distributed Hebbian Temporal Memory
by: Dzhivelikian, Evgenii, et al.
Published: (2023)
by: Dzhivelikian, Evgenii, et al.
Published: (2023)
Mind and Motion Aligned: A Joint Evaluation IsaacSim Benchmark for Task Planning and Low-Level Policies in Mobile Manipulation
by: Kachaev, Nikita, et al.
Published: (2025)
by: Kachaev, Nikita, et al.
Published: (2025)
VerifyLLM: LLM-Based Pre-Execution Task Plan Verification for Robots
by: Grigorev, Danil S., et al.
Published: (2025)
by: Grigorev, Danil S., et al.
Published: (2025)
POGEMA: A Benchmark Platform for Cooperative Multi-Agent Pathfinding
by: Skrynnik, Alexey, et al.
Published: (2024)
by: Skrynnik, Alexey, et al.
Published: (2024)
Relational Object-Centric Actor-Critic
by: Ugadiarov, Leonid, et al.
Published: (2023)
by: Ugadiarov, Leonid, et al.
Published: (2023)
Object-Centric World Models Meet Monte Carlo Tree Search
by: Vakhitov, Rodion, et al.
Published: (2026)
by: Vakhitov, Rodion, et al.
Published: (2026)
Learning to Communicate Locally for Large-Scale Multi-Agent Pathfinding
by: Vyaltsev, Valeriy, et al.
Published: (2026)
by: Vyaltsev, Valeriy, et al.
Published: (2026)
Spatial Traces: Enhancing VLA Models with Spatial-Temporal Understanding
by: Patratskiy, Maxim A., et al.
Published: (2025)
by: Patratskiy, Maxim A., et al.
Published: (2025)
Adaptive Conformal Prediction for Improving Factuality of Generations by Large Language Models
by: Rubashevskii, Aleksandr, et al.
Published: (2026)
by: Rubashevskii, Aleksandr, et al.
Published: (2026)
GIFT-SW: Gaussian noise Injected Fine-Tuning of Salient Weights for LLMs
by: Zhelnin, Maxim, et al.
Published: (2024)
by: Zhelnin, Maxim, et al.
Published: (2024)
Knowledge-Guided Manipulation Using Multi-Task Reinforcement Learning
by: Narendra, Aditya, et al.
Published: (2026)
by: Narendra, Aditya, et al.
Published: (2026)
A theoretical framework for self-supervised contrastive learning for continuous dependent data
by: Marusov, Alexander, et al.
Published: (2025)
by: Marusov, Alexander, et al.
Published: (2025)
Safe Planning and Policy Optimization via World Model Learning
by: Latyshev, Artem, et al.
Published: (2025)
by: Latyshev, Artem, et al.
Published: (2025)
HCVR: A Hybrid Approach with Correlation-aware Voting Rules for Feature Selection
by: Bhedasgaonkar, Nikita, et al.
Published: (2025)
by: Bhedasgaonkar, Nikita, et al.
Published: (2025)
MemAgent: Reshaping Long-Context LLM with Multi-Conv RL-based Memory Agent
by: Yu, Hongli, et al.
Published: (2025)
by: Yu, Hongli, et al.
Published: (2025)
Continuous Histogram Loss: Beyond Neural Similarity
by: Zholus, Artem, et al.
Published: (2020)
by: Zholus, Artem, et al.
Published: (2020)
Stack Trace Deduplication: Faster, More Accurately, and in More Realistic Scenarios
by: Shibaev, Egor, et al.
Published: (2024)
by: Shibaev, Egor, et al.
Published: (2024)
Uniformly Safe RL with Objective Suppression for Multi-Constraint Safety-Critical Applications
by: Zhou, Zihan, et al.
Published: (2024)
by: Zhou, Zihan, et al.
Published: (2024)
IDLM: Inverse-distilled Diffusion Language Models
by: Li, David, et al.
Published: (2026)
by: Li, David, et al.
Published: (2026)
Similar Items
-
Memory, Benchmark & Robots: A Benchmark for Solving Complex Tasks with Reinforcement Learning
by: Cherepanov, Egor, et al.
Published: (2025) -
ELMUR: External Layer Memory with Update/Rewrite for Long-Horizon RL Problems
by: Cherepanov, Egor, et al.
Published: (2025) -
Recurrent Action Transformer with Memory
by: Cherepanov, Egor, et al.
Published: (2023) -
Memory Retention Is Not Enough to Master Memory Tasks in Reinforcement Learning
by: Shchendrigin, Oleg, et al.
Published: (2026) -
A New Perspective on Transformers in Online Reinforcement Learning for Continuous Control
by: Kachaev, Nikita, et al.
Published: (2025)