Retrieval-Augmented Decision Transformer: External Memory for In-context RL
Fuente:
arXiv
Guardado en:
| Autores principales: | Schmied, Thomas, Paischer, Fabian, Patil, Vihang, Hofmarcher, Markus, Pascanu, Razvan, Hochreiter, Sepp |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Contrastive Abstraction for Reinforcement Learning
por: Patil, Vihang, et al.
Publicado: (2024)
por: Patil, Vihang, et al.
Publicado: (2024)
A Large Recurrent Action Model: xLSTM enables Fast Inference for Robotics Tasks
por: Schmied, Thomas, et al.
Publicado: (2024)
por: Schmied, Thomas, et al.
Publicado: (2024)
LLMs are Greedy Agents: Effects of RL Fine-tuning on Decision-Making Abilities
por: Schmied, Thomas, et al.
Publicado: (2025)
por: Schmied, Thomas, et al.
Publicado: (2025)
Linear Alignment of Vision-language Models for Image Captioning
por: Paischer, Fabian, et al.
Publicado: (2023)
por: Paischer, Fabian, et al.
Publicado: (2023)
Parameter Efficient Fine-tuning via Explained Variance Adaptation
por: Paischer, Fabian, et al.
Publicado: (2024)
por: Paischer, Fabian, et al.
Publicado: (2024)
Lattice: Learning to Efficiently Compress the Memory
por: Karami, Mahdi, et al.
Publicado: (2025)
por: Karami, Mahdi, et al.
Publicado: (2025)
Large Language Models Can Self-Improve At Web Agent Tasks
por: Patel, Ajay, et al.
Publicado: (2024)
por: Patel, Ajay, et al.
Publicado: (2024)
Promoting Exploration in Memory-Augmented Adam using Critical Momenta
por: Malviya, Pranshu, et al.
Publicado: (2023)
por: Malviya, Pranshu, et al.
Publicado: (2023)
FlashRNN: I/O-Aware Optimization of Traditional RNNs on modern hardware
por: Pöppel, Korbinian, et al.
Publicado: (2024)
por: Pöppel, Korbinian, et al.
Publicado: (2024)
State Soup: In-Context Skill Learning, Retrieval and Mixing
por: Pióro, Maciej, et al.
Publicado: (2024)
por: Pióro, Maciej, et al.
Publicado: (2024)
Simplified priors for Object-Centric Learning
por: Patil, Vihang, et al.
Publicado: (2024)
por: Patil, Vihang, et al.
Publicado: (2024)
Meta-learning how to Share Credit among Macro-Actions
por: Hosu, Ionel-Alexandru, et al.
Publicado: (2025)
por: Hosu, Ionel-Alexandru, et al.
Publicado: (2025)
Revisiting Adam for Streaming Reinforcement Learning
por: Gogianu, Florin, et al.
Publicado: (2026)
por: Gogianu, Florin, et al.
Publicado: (2026)
xLSTM: Extended Long Short-Term Memory
por: Beck, Maximilian, et al.
Publicado: (2024)
por: Beck, Maximilian, et al.
Publicado: (2024)
A Diffusion Model Framework for Unsupervised Neural Combinatorial Optimization
por: Sanokowski, Sebastian, et al.
Publicado: (2024)
por: Sanokowski, Sebastian, et al.
Publicado: (2024)
On Information-Theoretic Measures of Predictive Uncertainty
por: Schweighofer, Kajetan, et al.
Publicado: (2024)
por: Schweighofer, Kajetan, et al.
Publicado: (2024)
Improving Uncertainty Estimation through Semantically Diverse Language Generation
por: Aichberger, Lukas, et al.
Publicado: (2024)
por: Aichberger, Lukas, et al.
Publicado: (2024)
Addressing Pitfalls in the Evaluation of Uncertainty Estimation Methods for Natural Language Generation
por: Ielanskyi, Mykyta, et al.
Publicado: (2025)
por: Ielanskyi, Mykyta, et al.
Publicado: (2025)
Tiled Flash Linear Attention: More Efficient Linear RNN and xLSTM Kernels
por: Beck, Maximilian, et al.
Publicado: (2025)
por: Beck, Maximilian, et al.
Publicado: (2025)
The Disparate Benefits of Deep Ensembles
por: Schweighofer, Kajetan, et al.
Publicado: (2024)
por: Schweighofer, Kajetan, et al.
Publicado: (2024)
Rethinking Losses for Diffusion Bridge Samplers
por: Sanokowski, Sebastian, et al.
Publicado: (2025)
por: Sanokowski, Sebastian, et al.
Publicado: (2025)
Softmax is not Enough (for Sharp Size Generalisation)
por: Veličković, Petar, et al.
Publicado: (2024)
por: Veličković, Petar, et al.
Publicado: (2024)
Diffusion-Augmented Markov Decision Processes for Maximum Entropy Reinforcement Learning
por: Sanokowski, Sebastian, et al.
Publicado: (2025)
por: Sanokowski, Sebastian, et al.
Publicado: (2025)
Mining Generalizable Activation Functions
por: Vitvitskyi, Alex, et al.
Publicado: (2026)
por: Vitvitskyi, Alex, et al.
Publicado: (2026)
Online Finetuning Decision Transformers with Pure RL Gradients
por: Luo, Junkai, et al.
Publicado: (2026)
por: Luo, Junkai, et al.
Publicado: (2026)
ELMUR: External Layer Memory with Update/Rewrite for Long-Horizon RL Problems
por: Cherepanov, Egor, et al.
Publicado: (2025)
por: Cherepanov, Egor, et al.
Publicado: (2025)
pLSTM: parallelizable Linear Source Transition Mark networks
por: Pöppel, Korbinian, et al.
Publicado: (2025)
por: Pöppel, Korbinian, et al.
Publicado: (2025)
Bio-xLSTM: Generative modeling, representation and in-context learning of biological and chemical sequences
por: Schmidinger, Niklas, et al.
Publicado: (2024)
por: Schmidinger, Niklas, et al.
Publicado: (2024)
Perplexity Cannot Always Tell Right from Wrong
por: Veličković, Petar, et al.
Publicado: (2026)
por: Veličković, Petar, et al.
Publicado: (2026)
MIM-Refiner: A Contrastive Learning Boost from Intermediate Pre-Trained Representations
por: Alkin, Benedikt, et al.
Publicado: (2024)
por: Alkin, Benedikt, et al.
Publicado: (2024)
SymbolicAI: A framework for logic-based approaches combining generative models and solvers
por: Dinu, Marius-Constantin, et al.
Publicado: (2024)
por: Dinu, Marius-Constantin, et al.
Publicado: (2024)
Layerwise LQR for Geometry-Aware Optimization of Deep Networks
por: Dufort-Labbé, Simon, et al.
Publicado: (2026)
por: Dufort-Labbé, Simon, et al.
Publicado: (2026)
Navigating Potholes with Geometry-Aware Sharpness Minimization
por: Dufort-Labbé, Simon, et al.
Publicado: (2026)
por: Dufort-Labbé, Simon, et al.
Publicado: (2026)
VN-EGNN: E(3)-Equivariant Graph Neural Networks with Virtual Nodes Enhance Protein Binding Site Identification
por: Sestak, Florian, et al.
Publicado: (2024)
por: Sestak, Florian, et al.
Publicado: (2024)
Unlocking the Working Memory of Large Language Models for Latent Reasoning
por: Aichberger, Lukas, et al.
Publicado: (2026)
por: Aichberger, Lukas, et al.
Publicado: (2026)
Normalization and effective learning rates in reinforcement learning
por: Lyle, Clare, et al.
Publicado: (2024)
por: Lyle, Clare, et al.
Publicado: (2024)
Fine-Tuned In-Context Learners for Efficient Adaptation
por: Bornschein, Jorg, et al.
Publicado: (2025)
por: Bornschein, Jorg, et al.
Publicado: (2025)
Maxwell's Demon at Work: Efficient Pruning by Leveraging Saturation of Neurons
por: Dufort-Labbé, Simon, et al.
Publicado: (2024)
por: Dufort-Labbé, Simon, et al.
Publicado: (2024)
Kalman Filter for Online Classification of Non-Stationary Data
por: Titsias, Michalis K., et al.
Publicado: (2023)
por: Titsias, Michalis K., et al.
Publicado: (2023)
Toward Explainable Offline RL: Analyzing Representations in Intrinsically Motivated Decision Transformers
por: Guiducci, Leonardo, et al.
Publicado: (2025)
por: Guiducci, Leonardo, et al.
Publicado: (2025)
Ejemplares similares
-
Contrastive Abstraction for Reinforcement Learning
por: Patil, Vihang, et al.
Publicado: (2024) -
A Large Recurrent Action Model: xLSTM enables Fast Inference for Robotics Tasks
por: Schmied, Thomas, et al.
Publicado: (2024) -
LLMs are Greedy Agents: Effects of RL Fine-tuning on Decision-Making Abilities
por: Schmied, Thomas, et al.
Publicado: (2025) -
Linear Alignment of Vision-language Models for Image Captioning
por: Paischer, Fabian, et al.
Publicado: (2023) -
Parameter Efficient Fine-tuning via Explained Variance Adaptation
por: Paischer, Fabian, et al.
Publicado: (2024)