Neural Paging: Learning Context Management Policies for Turing-Complete Agents
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Liang, Liu, Qi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Position: The Turing-Completeness of Autoregressive Transformers Relies Heavily on Context Management
von: Cui, Guanyu, et al.
Veröffentlicht: (2026)
von: Cui, Guanyu, et al.
Veröffentlicht: (2026)
Transformers Provably Implement In-Context Reinforcement Learning with Policy Improvement
von: Liang, Haodong, et al.
Veröffentlicht: (2026)
von: Liang, Haodong, et al.
Veröffentlicht: (2026)
Turing Test on Screen: A Benchmark for Mobile GUI Agent Humanization
von: Zhu, Jiachen, et al.
Veröffentlicht: (2026)
von: Zhu, Jiachen, et al.
Veröffentlicht: (2026)
Efficient On-Device Agents via Adaptive Context Management
von: Vijayvargiya, Sanidhya, et al.
Veröffentlicht: (2025)
von: Vijayvargiya, Sanidhya, et al.
Veröffentlicht: (2025)
Random Policy Enables In-Context Reinforcement Learning within Trust Horizons
von: Chen, Weiqin, et al.
Veröffentlicht: (2024)
von: Chen, Weiqin, et al.
Veröffentlicht: (2024)
Offline Multi-Agent Reinforcement Learning via In-Sample Sequential Policy Optimization
von: Liu, Zongkai, et al.
Veröffentlicht: (2024)
von: Liu, Zongkai, et al.
Veröffentlicht: (2024)
PolicyLong: Towards On-Policy Context Extension
von: Jia, Junlong, et al.
Veröffentlicht: (2026)
von: Jia, Junlong, et al.
Veröffentlicht: (2026)
Paged Attention Meets FlexAttention: Unlocking Long-Context Efficiency in Deployed Inference
von: Joshi, Thomas, et al.
Veröffentlicht: (2025)
von: Joshi, Thomas, et al.
Veröffentlicht: (2025)
Learning to Learn-at-Test-Time: Language Agents with Learnable Adaptation Policies
von: Lou, Zhanzhi, et al.
Veröffentlicht: (2026)
von: Lou, Zhanzhi, et al.
Veröffentlicht: (2026)
DeepStock: Reinforcement Learning with Policy Regularizations for Inventory Management
von: Xie, Yaqi, et al.
Veröffentlicht: (2026)
von: Xie, Yaqi, et al.
Veröffentlicht: (2026)
AgentFold: Long-Horizon Web Agents with Proactive Context Management
von: Ye, Rui, et al.
Veröffentlicht: (2025)
von: Ye, Rui, et al.
Veröffentlicht: (2025)
Policy and World Modeling Co-Training for Language Agents
von: Lu, Ning, et al.
Veröffentlicht: (2026)
von: Lu, Ning, et al.
Veröffentlicht: (2026)
TypeBandit: Type-Level Context Allocation and Reweighting for Effective Attribute Completion in Heterogeneous Graph Neural Networks
von: Wang, Ta-Yang, et al.
Veröffentlicht: (2026)
von: Wang, Ta-Yang, et al.
Veröffentlicht: (2026)
Learning Long-Context Diffusion Policies via Past-Token Prediction
von: Torne, Marcel, et al.
Veröffentlicht: (2025)
von: Torne, Marcel, et al.
Veröffentlicht: (2025)
Learning to Remember: End-to-End Training of Memory Agents for Long-Context Reasoning
von: Zhang, Kehao, et al.
Veröffentlicht: (2026)
von: Zhang, Kehao, et al.
Veröffentlicht: (2026)
Agent-Omit: Adaptive Context Omission for Efficient LLM Agents
von: Ning, Yansong, et al.
Veröffentlicht: (2026)
von: Ning, Yansong, et al.
Veröffentlicht: (2026)
The PokeAgent Challenge: Competitive and Long-Context Learning at Scale
von: Karten, Seth, et al.
Veröffentlicht: (2026)
von: Karten, Seth, et al.
Veröffentlicht: (2026)
Context Distillation as Latent Memory Management
von: Zheng, Ziyang, et al.
Veröffentlicht: (2026)
von: Zheng, Ziyang, et al.
Veröffentlicht: (2026)
Scalable In-Context Q-Learning
von: Liu, Jinmei, et al.
Veröffentlicht: (2025)
von: Liu, Jinmei, et al.
Veröffentlicht: (2025)
Clustering Context in Off-Policy Evaluation
von: Guzman-Olivares, Daniel, et al.
Veröffentlicht: (2025)
von: Guzman-Olivares, Daniel, et al.
Veröffentlicht: (2025)
MEMENTO: Teaching LLMs to Manage Their Own Context
von: Kontonis, Vasilis, et al.
Veröffentlicht: (2026)
von: Kontonis, Vasilis, et al.
Veröffentlicht: (2026)
Metric-Gradient Projection for Stable Multi-Agent Policy Learning
von: Zhang, Zuyuan, et al.
Veröffentlicht: (2026)
von: Zhang, Zuyuan, et al.
Veröffentlicht: (2026)
$K$-Level Policy Gradients for Multi-Agent Reinforcement Learning
von: Reddi, Aryaman, et al.
Veröffentlicht: (2025)
von: Reddi, Aryaman, et al.
Veröffentlicht: (2025)
Policy Induction: Predicting Startup Success via Explainable Memory-Augmented In-Context Learning
von: Mu, Xianling, et al.
Veröffentlicht: (2025)
von: Mu, Xianling, et al.
Veröffentlicht: (2025)
Dual Turing Test: A Framework for Detecting and Mitigating Undetectable AI
von: Messina, Alberto
Veröffentlicht: (2025)
von: Messina, Alberto
Veröffentlicht: (2025)
Group-in-Group Policy Optimization for LLM Agent Training
von: Feng, Lang, et al.
Veröffentlicht: (2025)
von: Feng, Lang, et al.
Veröffentlicht: (2025)
Context and Diversity Matter: The Emergence of In-Context Learning in World Models
von: Wang, Fan, et al.
Veröffentlicht: (2025)
von: Wang, Fan, et al.
Veröffentlicht: (2025)
Context Learning for Multi-Agent Discussion
von: Hua, Xingyuan, et al.
Veröffentlicht: (2026)
von: Hua, Xingyuan, et al.
Veröffentlicht: (2026)
Reinforcement Learning for Machine Learning Engineering Agents
von: Yang, Sherry, et al.
Veröffentlicht: (2025)
von: Yang, Sherry, et al.
Veröffentlicht: (2025)
BlendRL: A Framework for Merging Symbolic and Neural Policy Learning
von: Shindo, Hikaru, et al.
Veröffentlicht: (2024)
von: Shindo, Hikaru, et al.
Veröffentlicht: (2024)
Conjugate Learning Theory: Uncovering the Mechanisms of Trainability and Generalization in Deep Neural Networks
von: Qi, Binchuan
Veröffentlicht: (2026)
von: Qi, Binchuan
Veröffentlicht: (2026)
ContextEvolve: Multi-Agent Context Compression for Systems Code Optimization
von: Su, Hongyuan, et al.
Veröffentlicht: (2026)
von: Su, Hongyuan, et al.
Veröffentlicht: (2026)
Mixture-of-Experts Meets In-Context Reinforcement Learning
von: Wu, Wenhao, et al.
Veröffentlicht: (2025)
von: Wu, Wenhao, et al.
Veröffentlicht: (2025)
UniFluids: Unified Neural Operator Learning with Conditional Flow-matching
von: Li, Haosen, et al.
Veröffentlicht: (2026)
von: Li, Haosen, et al.
Veröffentlicht: (2026)
Towards Lifecycle Unlearning Commitment Management: Measuring Sample-level Unlearning Completeness
von: Wang, Cheng-Long, et al.
Veröffentlicht: (2025)
von: Wang, Cheng-Long, et al.
Veröffentlicht: (2025)
Descent-Guided Policy Gradient for Scalable Cooperative Multi-Agent Learning
von: Yang, Shan, et al.
Veröffentlicht: (2026)
von: Yang, Shan, et al.
Veröffentlicht: (2026)
Provable and Practical In-Context Policy Optimization for Self-Improvement
von: Yu, Tianrun, et al.
Veröffentlicht: (2026)
von: Yu, Tianrun, et al.
Veröffentlicht: (2026)
Self-cross Feature based Spiking Neural Networks for Efficient Few-shot Learning
von: Xu, Qi, et al.
Veröffentlicht: (2025)
von: Xu, Qi, et al.
Veröffentlicht: (2025)
Challenge on Optimization of Context Collection for Code Completion
von: Ustalov, Dmitry, et al.
Veröffentlicht: (2025)
von: Ustalov, Dmitry, et al.
Veröffentlicht: (2025)
RLSynC: Offline-Online Reinforcement Learning for Synthon Completion
von: Baker, Frazier N., et al.
Veröffentlicht: (2023)
von: Baker, Frazier N., et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Position: The Turing-Completeness of Autoregressive Transformers Relies Heavily on Context Management
von: Cui, Guanyu, et al.
Veröffentlicht: (2026) -
Transformers Provably Implement In-Context Reinforcement Learning with Policy Improvement
von: Liang, Haodong, et al.
Veröffentlicht: (2026) -
Turing Test on Screen: A Benchmark for Mobile GUI Agent Humanization
von: Zhu, Jiachen, et al.
Veröffentlicht: (2026) -
Efficient On-Device Agents via Adaptive Context Management
von: Vijayvargiya, Sanidhya, et al.
Veröffentlicht: (2025) -
Random Policy Enables In-Context Reinforcement Learning within Trust Horizons
von: Chen, Weiqin, et al.
Veröffentlicht: (2024)