A New Perspective on Transformers in Online Reinforcement Learning for Continuous Control
Fuente:
arXiv
Salvato in:
| Autori principali: | Kachaev, Nikita, Zelezetsky, Daniil, Cherepanov, Egor, Kovelev, Alexey K., Panov, Aleksandr I. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Memory, Benchmark & Robots: A Benchmark for Solving Complex Tasks with Reinforcement Learning
di: Cherepanov, Egor, et al.
Pubblicazione: (2025)
di: Cherepanov, Egor, et al.
Pubblicazione: (2025)
Don't Blind Your VLA: Aligning Visual Representations for OOD Generalization
di: Kachaev, Nikita, et al.
Pubblicazione: (2025)
di: Kachaev, Nikita, et al.
Pubblicazione: (2025)
KAGE-Bench: Fast Known-Axis Visual Generalization Evaluation for Reinforcement Learning
di: Cherepanov, Egor, et al.
Pubblicazione: (2026)
di: Cherepanov, Egor, et al.
Pubblicazione: (2026)
Re:Frame -- Retrieving Experience From Associative Memory
di: Zelezetsky, Daniil, et al.
Pubblicazione: (2025)
di: Zelezetsky, Daniil, et al.
Pubblicazione: (2025)
ELMUR: External Layer Memory with Update/Rewrite for Long-Horizon RL Problems
di: Cherepanov, Egor, et al.
Pubblicazione: (2025)
di: Cherepanov, Egor, et al.
Pubblicazione: (2025)
Unraveling the Complexity of Memory in RL Agents: an Approach for Classification and Evaluation
di: Cherepanov, Egor, et al.
Pubblicazione: (2024)
di: Cherepanov, Egor, et al.
Pubblicazione: (2024)
Recurrent Action Transformer with Memory
di: Cherepanov, Egor, et al.
Pubblicazione: (2023)
di: Cherepanov, Egor, et al.
Pubblicazione: (2023)
Memory Retention Is Not Enough to Master Memory Tasks in Reinforcement Learning
di: Shchendrigin, Oleg, et al.
Pubblicazione: (2026)
di: Shchendrigin, Oleg, et al.
Pubblicazione: (2026)
Accelerating Transformers in Online RL
di: Zelezetsky, Daniil, et al.
Pubblicazione: (2025)
di: Zelezetsky, Daniil, et al.
Pubblicazione: (2025)
LookPlanGraph: Embodied Instruction Following Method with VLM Graph Augmentation
di: Onishchenko, Anatoly O., et al.
Pubblicazione: (2025)
di: Onishchenko, Anatoly O., et al.
Pubblicazione: (2025)
Knowledge-Guided Manipulation Using Multi-Task Reinforcement Learning
di: Narendra, Aditya, et al.
Pubblicazione: (2026)
di: Narendra, Aditya, et al.
Pubblicazione: (2026)
Relational Object-Centric Actor-Critic
di: Ugadiarov, Leonid, et al.
Pubblicazione: (2023)
di: Ugadiarov, Leonid, et al.
Pubblicazione: (2023)
Vintix II: Decision Pre-Trained Transformer is a Scalable In-Context Reinforcement Learner
di: Polubarov, Andrei, et al.
Pubblicazione: (2026)
di: Polubarov, Andrei, et al.
Pubblicazione: (2026)
HELP: Hierarchical Embodied Language Planner for Household Tasks
di: Korchemnyi, Alexandr V., et al.
Pubblicazione: (2025)
di: Korchemnyi, Alexandr V., et al.
Pubblicazione: (2025)
AmbiK: Dataset of Ambiguous Tasks in Kitchen Environment
di: Ivanova, Anastasiia, et al.
Pubblicazione: (2025)
di: Ivanova, Anastasiia, et al.
Pubblicazione: (2025)
Object-Centric World Models Meet Monte Carlo Tree Search
di: Vakhitov, Rodion, et al.
Pubblicazione: (2026)
di: Vakhitov, Rodion, et al.
Pubblicazione: (2026)
Mind and Motion Aligned: A Joint Evaluation IsaacSim Benchmark for Task Planning and Low-Level Policies in Mobile Manipulation
di: Kachaev, Nikita, et al.
Pubblicazione: (2025)
di: Kachaev, Nikita, et al.
Pubblicazione: (2025)
Object-Centric Learning with Slot Mixture Module
di: Kirilenko, Daniil, et al.
Pubblicazione: (2023)
di: Kirilenko, Daniil, et al.
Pubblicazione: (2023)
VerifyLLM: LLM-Based Pre-Execution Task Plan Verification for Robots
di: Grigorev, Danil S., et al.
Pubblicazione: (2025)
di: Grigorev, Danil S., et al.
Pubblicazione: (2025)
CAMAR: Continuous Actions Multi-Agent Routing
di: Pshenitsyn, Artem, et al.
Pubblicazione: (2025)
di: Pshenitsyn, Artem, et al.
Pubblicazione: (2025)
Steering LLM Reasoning Through Bias-Only Adaptation
di: Sinii, Viacheslav, et al.
Pubblicazione: (2025)
di: Sinii, Viacheslav, et al.
Pubblicazione: (2025)
Continuous Control Reinforcement Learning: Distributed Distributional DrQ Algorithms
di: Zhou, Zehao
Pubblicazione: (2024)
di: Zhou, Zehao
Pubblicazione: (2024)
SGN-CIRL: Scene Graph-based Navigation with Curriculum, Imitation, and Reinforcement Learning
di: Oskolkov, Nikita, et al.
Pubblicazione: (2025)
di: Oskolkov, Nikita, et al.
Pubblicazione: (2025)
Visualizing Critic Match Loss Landscapes for Interpretation of Online Reinforcement Learning Control Algorithms
di: Liu, Jingyi, et al.
Pubblicazione: (2026)
di: Liu, Jingyi, et al.
Pubblicazione: (2026)
RLtools: A Fast, Portable Deep Reinforcement Learning Library for Continuous Control
di: Eschmann, Jonas, et al.
Pubblicazione: (2023)
di: Eschmann, Jonas, et al.
Pubblicazione: (2023)
GOPT: Generalizable Online 3D Bin Packing via Transformer-based Deep Reinforcement Learning
di: Xiong, Heng, et al.
Pubblicazione: (2024)
di: Xiong, Heng, et al.
Pubblicazione: (2024)
Continual Policy Distillation of Reinforcement Learning-based Controllers for Soft Robotic In-Hand Manipulation
di: Li, Lanpei, et al.
Pubblicazione: (2024)
di: Li, Lanpei, et al.
Pubblicazione: (2024)
Online Decision MetaMorphFormer: A Casual Transformer-Based Reinforcement Learning Framework of Universal Embodied Intelligence
di: Ji, Luo, et al.
Pubblicazione: (2024)
di: Ji, Luo, et al.
Pubblicazione: (2024)
Online Training and Pruning of Deep Reinforcement Learning Networks
di: Guenter, Valentin Frank Ingmar, et al.
Pubblicazione: (2025)
di: Guenter, Valentin Frank Ingmar, et al.
Pubblicazione: (2025)
Online Continual Learning For Interactive Instruction Following Agents
di: Kim, Byeonghwi, et al.
Pubblicazione: (2024)
di: Kim, Byeonghwi, et al.
Pubblicazione: (2024)
Continual Model-Based Reinforcement Learning with Hypernetworks
di: Huang, Yizhou, et al.
Pubblicazione: (2020)
di: Huang, Yizhou, et al.
Pubblicazione: (2020)
Vintix: Action Model via In-Context Reinforcement Learning
di: Polubarov, Andrey, et al.
Pubblicazione: (2025)
di: Polubarov, Andrey, et al.
Pubblicazione: (2025)
ENOTO: Improving Offline-to-Online Reinforcement Learning with Q-Ensembles
di: Zhao, Kai, et al.
Pubblicazione: (2023)
di: Zhao, Kai, et al.
Pubblicazione: (2023)
Accelerated Online Reinforcement Learning using Auxiliary Start State Distributions
di: Mehra, Aman, et al.
Pubblicazione: (2025)
di: Mehra, Aman, et al.
Pubblicazione: (2025)
Robot Policy Transfer with Online Demonstrations: An Active Reinforcement Learning Approach
di: Hou, Muhan, et al.
Pubblicazione: (2025)
di: Hou, Muhan, et al.
Pubblicazione: (2025)
Sample Complexity of Distributionally Robust Off-Dynamics Reinforcement Learning with Online Interaction
di: He, Yiting, et al.
Pubblicazione: (2025)
di: He, Yiting, et al.
Pubblicazione: (2025)
Return Augmented Decision Transformer for Off-Dynamics Reinforcement Learning
di: Wang, Ruhan, et al.
Pubblicazione: (2024)
di: Wang, Ruhan, et al.
Pubblicazione: (2024)
Tiny Reinforcement Learning for Quadruped Locomotion using Decision Transformers
di: Akgün, Orhan Eren, et al.
Pubblicazione: (2024)
di: Akgün, Orhan Eren, et al.
Pubblicazione: (2024)
Online Action-Stacking Improves Reinforcement Learning Performance for Air Traffic Control
di: Carvell, Ben, et al.
Pubblicazione: (2026)
di: Carvell, Ben, et al.
Pubblicazione: (2026)
Gradual Transition from Bellman Optimality Operator to Bellman Operator in Online Reinforcement Learning
di: Omura, Motoki, et al.
Pubblicazione: (2025)
di: Omura, Motoki, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Memory, Benchmark & Robots: A Benchmark for Solving Complex Tasks with Reinforcement Learning
di: Cherepanov, Egor, et al.
Pubblicazione: (2025) -
Don't Blind Your VLA: Aligning Visual Representations for OOD Generalization
di: Kachaev, Nikita, et al.
Pubblicazione: (2025) -
KAGE-Bench: Fast Known-Axis Visual Generalization Evaluation for Reinforcement Learning
di: Cherepanov, Egor, et al.
Pubblicazione: (2026) -
Re:Frame -- Retrieving Experience From Associative Memory
di: Zelezetsky, Daniil, et al.
Pubblicazione: (2025) -
ELMUR: External Layer Memory with Update/Rewrite for Long-Horizon RL Problems
di: Cherepanov, Egor, et al.
Pubblicazione: (2025)