Synthetic POMDPs to Challenge Memory-Augmented RL: Memory Demand Structure Modeling
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Yongyi, Li, Lingfeng, Chen, Bozhou, Li, Ang, Liu, Hanyu, Zheng, Qirui, Yang, Xionghui, Li, Wenxin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Decoupling Return-to-Go for Efficient Decision Transformer
by: Wang, Yongyi, et al.
Published: (2026)
by: Wang, Yongyi, et al.
Published: (2026)
Beyond Autoregressive RTG: Conditioning via Injection Outside Sequential Modeling in Decision Transformer
by: Wang, Yongyi, et al.
Published: (2026)
by: Wang, Yongyi, et al.
Published: (2026)
ShuttleEnv: An Interactive Data-Driven RL Environment for Badminton Strategy Modeling
by: Li, Ang, et al.
Published: (2026)
by: Li, Ang, et al.
Published: (2026)
BotzoneBench: Scalable LLM Evaluation via Graded AI Anchors
by: Li, Lingfeng, et al.
Published: (2026)
by: Li, Lingfeng, et al.
Published: (2026)
Style-Preserving Policy Optimization for Game Agents
by: Li, Lingfeng, et al.
Published: (2025)
by: Li, Lingfeng, et al.
Published: (2025)
Mxplainer: Explain and Learn Insights by Imitating Mahjong Agents
by: Li, Lingfeng, et al.
Published: (2025)
by: Li, Lingfeng, et al.
Published: (2025)
Constructing Non-Markovian Decision Process via History Aggregator
by: Wang, Yongyi, et al.
Published: (2025)
by: Wang, Yongyi, et al.
Published: (2025)
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs
by: Galesloot, Maris F. L., et al.
Published: (2025)
by: Galesloot, Maris F. L., et al.
Published: (2025)
Adapting Rules of Official International Mahjong for Online Players
by: Wang, Chucai, et al.
Published: (2026)
by: Wang, Chucai, et al.
Published: (2026)
HAGE: Harnessing Agentic Memory via RL-Driven Weighted Graph Evolution
by: Jiang, Dongming, et al.
Published: (2026)
by: Jiang, Dongming, et al.
Published: (2026)
Generalizing Multi-Step Inverse Models for Representation Learning to Finite-Memory POMDPs
by: Wu, Lili, et al.
Published: (2024)
by: Wu, Lili, et al.
Published: (2024)
MindBridge: Scalable and Cross-Model Knowledge Editing via Memory-Augmented Modality
by: Li, Shuaike, et al.
Published: (2025)
by: Li, Shuaike, et al.
Published: (2025)
MemLong: Memory-Augmented Retrieval for Long Text Modeling
by: Liu, Weijie, et al.
Published: (2024)
by: Liu, Weijie, et al.
Published: (2024)
Scalable Policy-Based RL Algorithms for POMDPs
by: Anjarlekar, Ameya, et al.
Published: (2025)
by: Anjarlekar, Ameya, et al.
Published: (2025)
R-Debater: Retrieval-Augmented Debate Generation through Argumentative Memory
by: Li, Maoyuan, et al.
Published: (2025)
by: Li, Maoyuan, et al.
Published: (2025)
Retrieval-Augmented Decision Transformer: External Memory for In-context RL
by: Schmied, Thomas, et al.
Published: (2024)
by: Schmied, Thomas, et al.
Published: (2024)
Explainable Representation of Finite-Memory Policies for POMDPs using Decision Trees
by: Azeem, Muqsit, et al.
Published: (2024)
by: Azeem, Muqsit, et al.
Published: (2024)
SafeHarbor: Hierarchical Memory-Augmented Guardrail for LLM Agent Safety
by: Liu, Zhe, et al.
Published: (2026)
by: Liu, Zhe, et al.
Published: (2026)
Investigating Memory in Model-Free RL with POPGym Arcade
by: Wang, Zekang, et al.
Published: (2025)
by: Wang, Zekang, et al.
Published: (2025)
GRAVITY: Architecture-Agnostic Structured Anchoring for Long-Horizon Conversational Memory
by: Sun, Yushi, et al.
Published: (2026)
by: Sun, Yushi, et al.
Published: (2026)
Memory in Large Language Models: Mechanisms, Evaluation and Evolution
by: Zhang, Dianxing, et al.
Published: (2025)
by: Zhang, Dianxing, et al.
Published: (2025)
MemPO: Self-Memory Policy Optimization for Long-Horizon Agents
by: Li, Ruoran, et al.
Published: (2026)
by: Li, Ruoran, et al.
Published: (2026)
Exploratory Memory-Augmented LLM Agent via Hybrid On- and Off-Policy Optimization
by: Liu, Zeyuan, et al.
Published: (2026)
by: Liu, Zeyuan, et al.
Published: (2026)
Scalable Solution Methods for Dec-POMDPs with Deterministic Dynamics
by: You, Yang, et al.
Published: (2025)
by: You, Yang, et al.
Published: (2025)
Contrastive Augmented Graph2Graph Memory Interaction for Few Shot Continual Learning
by: Qi, Biqing, et al.
Published: (2024)
by: Qi, Biqing, et al.
Published: (2024)
Pareto-guided Pipeline for Distilling Featherweight AI Agents in Mobile MOBA Games
by: Yang, Xionghui, et al.
Published: (2026)
by: Yang, Xionghui, et al.
Published: (2026)
Boundary-Guided Policy Optimization for Memory-efficient RL of Diffusion Large Language Models
by: Lin, Nianyi, et al.
Published: (2025)
by: Lin, Nianyi, et al.
Published: (2025)
MemoryMamba: Memory-Augmented State Space Model for Defect Recognition
by: Wang, Qianning, et al.
Published: (2024)
by: Wang, Qianning, et al.
Published: (2024)
ODMA: On-Demand Memory Allocation Strategy for LLM Serving on LPDDR-Class Accelerators
by: Zou, Guoqiang, et al.
Published: (2025)
by: Zou, Guoqiang, et al.
Published: (2025)
LifeAlign: Lifelong Alignment for Large Language Models with Memory-Augmented Focalized Preference Optimization
by: Li, Junsong, et al.
Published: (2025)
by: Li, Junsong, et al.
Published: (2025)
Generalizable End-to-End Tool-Use RL with Synthetic CodeGym
by: Du, Weihua, et al.
Published: (2025)
by: Du, Weihua, et al.
Published: (2025)
Automated Reformulation of Robust Optimization via Memory-Augmented Large Language Models
by: Chen, Jinbiao, et al.
Published: (2026)
by: Chen, Jinbiao, et al.
Published: (2026)
MemoryKT: An Integrative Memory-and-Forgetting Method for Knowledge Tracing
by: Lin, Mingrong, et al.
Published: (2025)
by: Lin, Mingrong, et al.
Published: (2025)
State Contamination in Memory-Augmented LLM Agents
by: Wang, Yian, et al.
Published: (2026)
by: Wang, Yian, et al.
Published: (2026)
MemRec: Collaborative Memory-Augmented Agentic Recommender System
by: Chen, Weixin, et al.
Published: (2026)
by: Chen, Weixin, et al.
Published: (2026)
Symmetry-Guided Memory Augmentation for Efficient Locomotion Learning
by: Bao, Kaixi, et al.
Published: (2025)
by: Bao, Kaixi, et al.
Published: (2025)
Memory, Consciousness and Large Language Model
by: Li, Jitang, et al.
Published: (2024)
by: Li, Jitang, et al.
Published: (2024)
MetaState: Persistent Working Memory Enhances Reasoning in Discrete Diffusion Language Models
by: Xia, Kejing, et al.
Published: (2026)
by: Xia, Kejing, et al.
Published: (2026)
$δ$-mem: Efficient Online Memory for Large Language Models
by: Lei, Jingdi, et al.
Published: (2026)
by: Lei, Jingdi, et al.
Published: (2026)
Coinvisor: An RL-Enhanced Chatbot Agent for Interactive Cryptocurrency Investment Analysis
by: Chen, Chong, et al.
Published: (2025)
by: Chen, Chong, et al.
Published: (2025)
Similar Items
-
Decoupling Return-to-Go for Efficient Decision Transformer
by: Wang, Yongyi, et al.
Published: (2026) -
Beyond Autoregressive RTG: Conditioning via Injection Outside Sequential Modeling in Decision Transformer
by: Wang, Yongyi, et al.
Published: (2026) -
ShuttleEnv: An Interactive Data-Driven RL Environment for Badminton Strategy Modeling
by: Li, Ang, et al.
Published: (2026) -
BotzoneBench: Scalable LLM Evaluation via Graded AI Anchors
by: Li, Lingfeng, et al.
Published: (2026) -
Style-Preserving Policy Optimization for Game Agents
by: Li, Lingfeng, et al.
Published: (2025)