Recurrent Action Transformer with Memory
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cherepanov, Egor, Staroverov, Alexey, Kovalev, Alexey K., Panov, Aleksandr I. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ELMUR: External Layer Memory with Update/Rewrite for Long-Horizon RL Problems
von: Cherepanov, Egor, et al.
Veröffentlicht: (2025)
von: Cherepanov, Egor, et al.
Veröffentlicht: (2025)
Memory Retention Is Not Enough to Master Memory Tasks in Reinforcement Learning
von: Shchendrigin, Oleg, et al.
Veröffentlicht: (2026)
von: Shchendrigin, Oleg, et al.
Veröffentlicht: (2026)
Re:Frame -- Retrieving Experience From Associative Memory
von: Zelezetsky, Daniil, et al.
Veröffentlicht: (2025)
von: Zelezetsky, Daniil, et al.
Veröffentlicht: (2025)
Memory, Benchmark & Robots: A Benchmark for Solving Complex Tasks with Reinforcement Learning
von: Cherepanov, Egor, et al.
Veröffentlicht: (2025)
von: Cherepanov, Egor, et al.
Veröffentlicht: (2025)
Unraveling the Complexity of Memory in RL Agents: an Approach for Classification and Evaluation
von: Cherepanov, Egor, et al.
Veröffentlicht: (2024)
von: Cherepanov, Egor, et al.
Veröffentlicht: (2024)
KAGE-Bench: Fast Known-Axis Visual Generalization Evaluation for Reinforcement Learning
von: Cherepanov, Egor, et al.
Veröffentlicht: (2026)
von: Cherepanov, Egor, et al.
Veröffentlicht: (2026)
A New Perspective on Transformers in Online Reinforcement Learning for Continuous Control
von: Kachaev, Nikita, et al.
Veröffentlicht: (2025)
von: Kachaev, Nikita, et al.
Veröffentlicht: (2025)
LookPlanGraph: Embodied Instruction Following Method with VLM Graph Augmentation
von: Onishchenko, Anatoly O., et al.
Veröffentlicht: (2025)
von: Onishchenko, Anatoly O., et al.
Veröffentlicht: (2025)
Symbolic Disentangled Representations for Images
von: Korchemnyi, Alexandr, et al.
Veröffentlicht: (2024)
von: Korchemnyi, Alexandr, et al.
Veröffentlicht: (2024)
CAMAR: Continuous Actions Multi-Agent Routing
von: Pshenitsyn, Artem, et al.
Veröffentlicht: (2025)
von: Pshenitsyn, Artem, et al.
Veröffentlicht: (2025)
Object-Centric Learning with Slot Mixture Module
von: Kirilenko, Daniil, et al.
Veröffentlicht: (2023)
von: Kirilenko, Daniil, et al.
Veröffentlicht: (2023)
Don't Blind Your VLA: Aligning Visual Representations for OOD Generalization
von: Kachaev, Nikita, et al.
Veröffentlicht: (2025)
von: Kachaev, Nikita, et al.
Veröffentlicht: (2025)
AmbiK: Dataset of Ambiguous Tasks in Kitchen Environment
von: Ivanova, Anastasiia, et al.
Veröffentlicht: (2025)
von: Ivanova, Anastasiia, et al.
Veröffentlicht: (2025)
Accelerating Transformers in Online RL
von: Zelezetsky, Daniil, et al.
Veröffentlicht: (2025)
von: Zelezetsky, Daniil, et al.
Veröffentlicht: (2025)
HELP: Hierarchical Embodied Language Planner for Household Tasks
von: Korchemnyi, Alexandr V., et al.
Veröffentlicht: (2025)
von: Korchemnyi, Alexandr V., et al.
Veröffentlicht: (2025)
Revisiting Tree Search for LLMs: Gumbel and Sequential Halving for Budget-Scalable Reasoning
von: Ugadiarov, Leonid, et al.
Veröffentlicht: (2026)
von: Ugadiarov, Leonid, et al.
Veröffentlicht: (2026)
MARL-GPT: Foundation Model for Multi-Agent Reinforcement Learning
von: Nesterova, Maria, et al.
Veröffentlicht: (2026)
von: Nesterova, Maria, et al.
Veröffentlicht: (2026)
VerifyLLM: LLM-Based Pre-Execution Task Plan Verification for Robots
von: Grigorev, Danil S., et al.
Veröffentlicht: (2025)
von: Grigorev, Danil S., et al.
Veröffentlicht: (2025)
Advancing Learnable Multi-Agent Pathfinding Solvers with Active Fine-Tuning
von: Andreychuk, Anton, et al.
Veröffentlicht: (2025)
von: Andreychuk, Anton, et al.
Veröffentlicht: (2025)
MAPF-GPT: Imitation Learning for Multi-Agent Pathfinding at Scale
von: Andreychuk, Anton, et al.
Veröffentlicht: (2024)
von: Andreychuk, Anton, et al.
Veröffentlicht: (2024)
Spatial Traces: Enhancing VLA Models with Spatial-Temporal Understanding
von: Patratskiy, Maxim A., et al.
Veröffentlicht: (2025)
von: Patratskiy, Maxim A., et al.
Veröffentlicht: (2025)
Learning Successor Features with Distributed Hebbian Temporal Memory
von: Dzhivelikian, Evgenii, et al.
Veröffentlicht: (2023)
von: Dzhivelikian, Evgenii, et al.
Veröffentlicht: (2023)
POGEMA: A Benchmark Platform for Cooperative Multi-Agent Pathfinding
von: Skrynnik, Alexey, et al.
Veröffentlicht: (2024)
von: Skrynnik, Alexey, et al.
Veröffentlicht: (2024)
A theoretical framework for self-supervised contrastive learning for continuous dependent data
von: Marusov, Alexander, et al.
Veröffentlicht: (2025)
von: Marusov, Alexander, et al.
Veröffentlicht: (2025)
Steering LLM Reasoning Through Bias-Only Adaptation
von: Sinii, Viacheslav, et al.
Veröffentlicht: (2025)
von: Sinii, Viacheslav, et al.
Veröffentlicht: (2025)
Relational Object-Centric Actor-Critic
von: Ugadiarov, Leonid, et al.
Veröffentlicht: (2023)
von: Ugadiarov, Leonid, et al.
Veröffentlicht: (2023)
Object-Centric World Models Meet Monte Carlo Tree Search
von: Vakhitov, Rodion, et al.
Veröffentlicht: (2026)
von: Vakhitov, Rodion, et al.
Veröffentlicht: (2026)
Learning to Communicate Locally for Large-Scale Multi-Agent Pathfinding
von: Vyaltsev, Valeriy, et al.
Veröffentlicht: (2026)
von: Vyaltsev, Valeriy, et al.
Veröffentlicht: (2026)
Dynamic Neural Potential Field: Online Trajectory Optimization in the Presence of Moving Obstacles
von: Staroverov, Aleksei, et al.
Veröffentlicht: (2024)
von: Staroverov, Aleksei, et al.
Veröffentlicht: (2024)
Thinking like a CHEMIST: Combined Heterogeneous Embedding Model Integrating Structure and Tokens
von: Rekut, Nikolai, et al.
Veröffentlicht: (2025)
von: Rekut, Nikolai, et al.
Veröffentlicht: (2025)
Adaptive Conformal Prediction for Improving Factuality of Generations by Large Language Models
von: Rubashevskii, Aleksandr, et al.
Veröffentlicht: (2026)
von: Rubashevskii, Aleksandr, et al.
Veröffentlicht: (2026)
Instruction Following with Goal-Conditioned Reinforcement Learning in Virtual Environments
von: Volovikova, Zoya, et al.
Veröffentlicht: (2024)
von: Volovikova, Zoya, et al.
Veröffentlicht: (2024)
GIFT-SW: Gaussian noise Injected Fine-Tuning of Salient Weights for LLMs
von: Zhelnin, Maxim, et al.
Veröffentlicht: (2024)
von: Zhelnin, Maxim, et al.
Veröffentlicht: (2024)
Knowledge-Guided Manipulation Using Multi-Task Reinforcement Learning
von: Narendra, Aditya, et al.
Veröffentlicht: (2026)
von: Narendra, Aditya, et al.
Veröffentlicht: (2026)
Parameter-Efficient Neural CDEs via Implicit Function Jacobians
von: Kuleshov, Ilya, et al.
Veröffentlicht: (2025)
von: Kuleshov, Ilya, et al.
Veröffentlicht: (2025)
Associative Recurrent Memory Transformer
von: Rodkin, Ivan, et al.
Veröffentlicht: (2024)
von: Rodkin, Ivan, et al.
Veröffentlicht: (2024)
Recurrence-Complete Frame-based Action Models
von: Keiblinger, Michael
Veröffentlicht: (2025)
von: Keiblinger, Michael
Veröffentlicht: (2025)
Stack Trace Deduplication: Faster, More Accurately, and in More Realistic Scenarios
von: Shibaev, Egor, et al.
Veröffentlicht: (2024)
von: Shibaev, Egor, et al.
Veröffentlicht: (2024)
Trust-Region Behavior Blending for On-Policy Distillation
von: Plyusov, Daniil, et al.
Veröffentlicht: (2026)
von: Plyusov, Daniil, et al.
Veröffentlicht: (2026)
F-GRPO: Don't Let Your Policy Learn the Obvious and Forget the Rare
von: Plyusov, Daniil, et al.
Veröffentlicht: (2026)
von: Plyusov, Daniil, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
ELMUR: External Layer Memory with Update/Rewrite for Long-Horizon RL Problems
von: Cherepanov, Egor, et al.
Veröffentlicht: (2025) -
Memory Retention Is Not Enough to Master Memory Tasks in Reinforcement Learning
von: Shchendrigin, Oleg, et al.
Veröffentlicht: (2026) -
Re:Frame -- Retrieving Experience From Associative Memory
von: Zelezetsky, Daniil, et al.
Veröffentlicht: (2025) -
Memory, Benchmark & Robots: A Benchmark for Solving Complex Tasks with Reinforcement Learning
von: Cherepanov, Egor, et al.
Veröffentlicht: (2025) -
Unraveling the Complexity of Memory in RL Agents: an Approach for Classification and Evaluation
von: Cherepanov, Egor, et al.
Veröffentlicht: (2024)