Beyond Sliding Windows: Learning to Manage Memory in Non-Markovian Environments
Fuente:
arXiv
Saved in:
| Main Authors: | Tasse, Geraud Nangue, Riemer, Matthew, Rosman, Benjamin, Klinger, Tim |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Counting Reward Automata: Sample Efficient Reinforcement Learning Through the Exploitation of Reward Function Structure
by: Bester, Tristan, et al.
Published: (2023)
by: Bester, Tristan, et al.
Published: (2023)
Skill Machines: Temporal Logic Skill Composition in Reinforcement Learning
by: Tasse, Geraud Nangue, et al.
Published: (2022)
by: Tasse, Geraud Nangue, et al.
Published: (2022)
MoralityGym: A Benchmark for Evaluating Hierarchical Moral Alignment in Sequential Decision-Making Agents
by: Rosen, Simon, et al.
Published: (2026)
by: Rosen, Simon, et al.
Published: (2026)
RobocupGym: A challenging continuous control benchmark in Robocup
by: Beukman, Michael, et al.
Published: (2024)
by: Beukman, Michael, et al.
Published: (2024)
Unsupervised Hierarchical Skill Discovery
by: Harvey, Damion, et al.
Published: (2026)
by: Harvey, Damion, et al.
Published: (2026)
Compositional Instruction Following with Language Models and Reinforcement Learning
by: Cohen, Vanya, et al.
Published: (2025)
by: Cohen, Vanya, et al.
Published: (2025)
Deep RL With Information Constrained Policies: Generalization in Continuous Control
by: Malloy, Tailia, et al.
Published: (2020)
by: Malloy, Tailia, et al.
Published: (2020)
Policy Dispersion in Non-Markovian Environment
by: Qu, Bohao, et al.
Published: (2023)
by: Qu, Bohao, et al.
Published: (2023)
Enabling Realtime Reinforcement Learning at Scale with Staggered Asynchronous Inference
by: Riemer, Matthew, et al.
Published: (2024)
by: Riemer, Matthew, et al.
Published: (2024)
VPWEM: Non-Markovian Visuomotor Policy with Working and Episodic Memory
by: Lei, Yuheng, et al.
Published: (2026)
by: Lei, Yuheng, et al.
Published: (2026)
Safe Reinforcement Learning with Learned Non-Markovian Safety Constraints
by: Low, Siow Meng, et al.
Published: (2024)
by: Low, Siow Meng, et al.
Published: (2024)
Policy Gradient Methods for Non-Markovian Reinforcement Learning
by: Kar, Avik, et al.
Published: (2026)
by: Kar, Avik, et al.
Published: (2026)
What makes Models Compositional? A Theoretical View: With Supplement
by: Ram, Parikshit, et al.
Published: (2024)
by: Ram, Parikshit, et al.
Published: (2024)
On The Specialization of Neural Modules
by: Jarvis, Devon, et al.
Published: (2024)
by: Jarvis, Devon, et al.
Published: (2024)
Spatial-Temporal Reinforcement Learning for Network Routing with Non-Markovian Traffic
by: Wang, Molly, et al.
Published: (2025)
by: Wang, Molly, et al.
Published: (2025)
Sliding Window Attention Training for Efficient Large Language Models
by: Fu, Zichuan, et al.
Published: (2025)
by: Fu, Zichuan, et al.
Published: (2025)
The Effectiveness of Approximate Regularized Replay for Efficient Supervised Fine-Tuning of Large Language Models
by: Riemer, Matthew, et al.
Published: (2025)
by: Riemer, Matthew, et al.
Published: (2025)
ParMod: A Parallel and Modular Framework for Learning Non-Markovian Tasks
by: Miao, Ruixuan, et al.
Published: (2024)
by: Miao, Ruixuan, et al.
Published: (2024)
Handling Delay in Real-Time Reinforcement Learning
by: Anokhin, Ivan, et al.
Published: (2025)
by: Anokhin, Ivan, et al.
Published: (2025)
Model-Based Reinforcement Learning in Discrete-Action Non-Markovian Reward Decision Processes
by: Trapasso, Alessandro, et al.
Published: (2025)
by: Trapasso, Alessandro, et al.
Published: (2025)
Process-Tensor Tomography of SGD: Measuring Non-Markovian Memory via Back-Flow of Distinguishability
by: Sevetlidis, Vasileios, et al.
Published: (2026)
by: Sevetlidis, Vasileios, et al.
Published: (2026)
Learning to Plan, Planning to Learn: Adaptive Hierarchical RL-MPC for Sample-Efficient Decision Making
by: Hori, Toshiaki, et al.
Published: (2025)
by: Hori, Toshiaki, et al.
Published: (2025)
Adaptive Memory Crystallization for Autonomous AI Agent Learning in Dynamic Environments
by: Khanda, Rajat, et al.
Published: (2026)
by: Khanda, Rajat, et al.
Published: (2026)
Momentum Boosted Episodic Memory for Improving Learning in Long-Tailed RL Environments
by: Fernandes, Dolton, et al.
Published: (2025)
by: Fernandes, Dolton, et al.
Published: (2025)
SketchOGD: Memory-Efficient Continual Learning
by: Min, Youngjae, et al.
Published: (2023)
by: Min, Youngjae, et al.
Published: (2023)
Online Reinforcement Learning in Non-Stationary Context-Driven Environments
by: Hamadanian, Pouya, et al.
Published: (2023)
by: Hamadanian, Pouya, et al.
Published: (2023)
Spectral Convolution on Orbifolds for Geometric Deep Learning
by: Mangliers, Tim, et al.
Published: (2026)
by: Mangliers, Tim, et al.
Published: (2026)
Task-Core Memory Management and Consolidation for Long-term Continual Learning
by: Huai, Tianyu, et al.
Published: (2025)
by: Huai, Tianyu, et al.
Published: (2025)
The ODE Method for Stochastic Approximation and Reinforcement Learning with Markovian Noise
by: Liu, Shuze Daniel, et al.
Published: (2024)
by: Liu, Shuze Daniel, et al.
Published: (2024)
Raising the Bar in Graph OOD Generalization: Invariant Learning Beyond Explicit Environment Modeling
by: Shen, Xu, et al.
Published: (2025)
by: Shen, Xu, et al.
Published: (2025)
Foundation World Models for Agents that Learn, Verify, and Adapt Reliably Beyond Static Environments
by: Delgrange, Florent
Published: (2026)
by: Delgrange, Florent
Published: (2026)
Non-Markovian Discrete Diffusion with Causal Language Models
by: Zhang, Yangtian, et al.
Published: (2025)
by: Zhang, Yangtian, et al.
Published: (2025)
ABC: Any-Subset Autoregression via Non-Markovian Diffusion Bridges in Continuous Time and Space
by: Guo, Gabe, et al.
Published: (2026)
by: Guo, Gabe, et al.
Published: (2026)
MARLINE: Multi-Source Mapping Transfer Learning for Non-Stationary Environments
by: Du, Honghui, et al.
Published: (2025)
by: Du, Honghui, et al.
Published: (2025)
Context Distillation as Latent Memory Management
by: Zheng, Ziyang, et al.
Published: (2026)
by: Zheng, Ziyang, et al.
Published: (2026)
Sophisticated Learning: A novel algorithm for active learning during model-based planning
by: Hodson, Rowan, et al.
Published: (2023)
by: Hodson, Rowan, et al.
Published: (2023)
Forecasting in Offline Reinforcement Learning for Non-stationary Environments
by: Ada, Suzan Ece, et al.
Published: (2025)
by: Ada, Suzan Ece, et al.
Published: (2025)
Learning to Solve Orienteering Problem with Time Windows and Variable Profits
by: Gao, Songqun, et al.
Published: (2026)
by: Gao, Songqun, et al.
Published: (2026)
Job Shop Scheduling Benchmark: Environments and Instances for Learning and Non-learning Methods
by: Reijnen, Robbert, et al.
Published: (2023)
by: Reijnen, Robbert, et al.
Published: (2023)
Honest Lying: Understanding Memory Confabulation in Reflexive Agents
by: Dixit, Prakhar, et al.
Published: (2026)
by: Dixit, Prakhar, et al.
Published: (2026)
Similar Items
-
Counting Reward Automata: Sample Efficient Reinforcement Learning Through the Exploitation of Reward Function Structure
by: Bester, Tristan, et al.
Published: (2023) -
Skill Machines: Temporal Logic Skill Composition in Reinforcement Learning
by: Tasse, Geraud Nangue, et al.
Published: (2022) -
MoralityGym: A Benchmark for Evaluating Hierarchical Moral Alignment in Sequential Decision-Making Agents
by: Rosen, Simon, et al.
Published: (2026) -
RobocupGym: A challenging continuous control benchmark in Robocup
by: Beukman, Michael, et al.
Published: (2024) -
Unsupervised Hierarchical Skill Discovery
by: Harvey, Damion, et al.
Published: (2026)