InfiniteICL: Breaking the Limit of Context Window Size via Long Short-term Memory Transformation
Fuente:
arXiv
Guardado en:
| Autores principales: | Cao, Bowen, Cai, Deng, Lam, Wai |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
GRAVITY: Architecture-Agnostic Structured Anchoring for Long-Horizon Conversational Memory
por: Sun, Yushi, et al.
Publicado: (2026)
por: Sun, Yushi, et al.
Publicado: (2026)
On the Worst Prompt Performance of Large Language Models
por: Cao, Bowen, et al.
Publicado: (2024)
por: Cao, Bowen, et al.
Publicado: (2024)
Stress-Testing Long-Context Language Models with Lifelong ICL and Task Haystack
por: Xu, Xiaoyue, et al.
Publicado: (2024)
por: Xu, Xiaoyue, et al.
Publicado: (2024)
InftyThink: Breaking the Length Limits of Long-Context Reasoning in Large Language Models
por: Yan, Yuchen, et al.
Publicado: (2025)
por: Yan, Yuchen, et al.
Publicado: (2025)
Commonsense-augmented Memory Construction and Management in Long-term Conversations via Context-aware Persona Refinement
por: Kim, Hana, et al.
Publicado: (2024)
por: Kim, Hana, et al.
Publicado: (2024)
HGMEM: Hypergraph-based Working Memory to Improve Multi-step RAG for Long-Context Complex Relational Modeling
por: Zhou, Chulun, et al.
Publicado: (2025)
por: Zhou, Chulun, et al.
Publicado: (2025)
DepWiGNN: A Depth-wise Graph Neural Network for Multi-hop Spatial Reasoning in Text
por: Li, Shuaiyi, et al.
Publicado: (2023)
por: Li, Shuaiyi, et al.
Publicado: (2023)
Auto-ICL: In-Context Learning without Human Supervision
por: Yang, Jinghan, et al.
Publicado: (2023)
por: Yang, Jinghan, et al.
Publicado: (2023)
Gradual Forgetting: Logarithmic Compression for Extending Transformer Context Windows
por: Dickson, Billy, et al.
Publicado: (2025)
por: Dickson, Billy, et al.
Publicado: (2025)
RetICL: Sequential Retrieval of In-Context Examples with Reinforcement Learning
por: Scarlatos, Alexander, et al.
Publicado: (2023)
por: Scarlatos, Alexander, et al.
Publicado: (2023)
Opt-ICL at LeWiDi-2025: Maximizing In-Context Signal from Rater Examples via Meta-Learning
por: Sorensen, Taylor, et al.
Publicado: (2025)
por: Sorensen, Taylor, et al.
Publicado: (2025)
On the Transformations across Reward Model, Parameter Update, and In-Context Prompt
por: Cai, Deng, et al.
Publicado: (2024)
por: Cai, Deng, et al.
Publicado: (2024)
Towards Infinite-Long Prefix in Transformer
por: Liang, Yingyu, et al.
Publicado: (2024)
por: Liang, Yingyu, et al.
Publicado: (2024)
SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization
por: Sun, Huashan, et al.
Publicado: (2025)
por: Sun, Huashan, et al.
Publicado: (2025)
OWL: Overcoming Window Length-Dependence in Speculative Decoding for Long-Context Inputs
por: Lee, Jaeseong, et al.
Publicado: (2025)
por: Lee, Jaeseong, et al.
Publicado: (2025)
LongCodeBench: Evaluating Coding LLMs at 1M Context Windows
por: Rando, Stefano, et al.
Publicado: (2025)
por: Rando, Stefano, et al.
Publicado: (2025)
Look Back to Reason Forward: Revisitable Memory for Long-Context LLM Agents
por: Shi, Yaorui, et al.
Publicado: (2025)
por: Shi, Yaorui, et al.
Publicado: (2025)
YaRN: Efficient Context Window Extension of Large Language Models
por: Peng, Bowen, et al.
Publicado: (2023)
por: Peng, Bowen, et al.
Publicado: (2023)
Evaluating Language Model Context Windows: A "Working Memory" Test and Inference-time Correction
por: Dsouza, Amanda, et al.
Publicado: (2024)
por: Dsouza, Amanda, et al.
Publicado: (2024)
SWAA: Sliding Window Attention Adaptation for Efficient and Quality Preserving Long Context Processing
por: Yu, Yijiong, et al.
Publicado: (2025)
por: Yu, Yijiong, et al.
Publicado: (2025)
CoCA: Fusing Position Embedding with Collinear Constrained Attention in Transformers for Long Context Window Extending
por: Zhu, Shiyi, et al.
Publicado: (2023)
por: Zhu, Shiyi, et al.
Publicado: (2023)
Generalizing From Short to Long: Effective Data Synthesis for Long-Context Instruction Tuning
por: Zhu, Wenhao, et al.
Publicado: (2025)
por: Zhu, Wenhao, et al.
Publicado: (2025)
MemReread: Enhancing Agentic Long-Context Reasoning via Memory-Guided Rereading
por: Ji, Baibei, et al.
Publicado: (2026)
por: Ji, Baibei, et al.
Publicado: (2026)
Simple Hack for Transformers against Heavy Long-Text Classification on a Time- and Memory-Limited GPU Service
por: Mutasodirin, Mirza Alim, et al.
Publicado: (2024)
por: Mutasodirin, Mirza Alim, et al.
Publicado: (2024)
Contexts are Never Long Enough: Structured Reasoning for Scalable Question Answering over Long Document Sets
por: Joshi, Harshit, et al.
Publicado: (2026)
por: Joshi, Harshit, et al.
Publicado: (2026)
Vector-ICL: In-context Learning with Continuous Vector Representations
por: Zhuang, Yufan, et al.
Publicado: (2024)
por: Zhuang, Yufan, et al.
Publicado: (2024)
Towards Generalization of Block Attention via Automatic Segmentation and Block Distillation
por: Li, Shuaiyi, et al.
Publicado: (2026)
por: Li, Shuaiyi, et al.
Publicado: (2026)
Recurrent Context Compression: Efficiently Expanding the Context Window of LLM
por: Huang, Chensen, et al.
Publicado: (2024)
por: Huang, Chensen, et al.
Publicado: (2024)
Dynamic Long Context Reasoning over Compressed Memory via End-to-End Reinforcement Learning
por: Chen, Zhuoen, et al.
Publicado: (2026)
por: Chen, Zhuoen, et al.
Publicado: (2026)
TSUBASA: Improving Long-Horizon Personalization via Evolving Memory and Self-Learning with Context Distillation
por: Zhang, Xinliang Frederick, et al.
Publicado: (2026)
por: Zhang, Xinliang Frederick, et al.
Publicado: (2026)
Human-inspired Episodic Memory for Infinite Context LLMs
por: Fountas, Zafeirios, et al.
Publicado: (2024)
por: Fountas, Zafeirios, et al.
Publicado: (2024)
Dynamic Long Short-Term Memory Based Memory Storage For Long Horizon LLM Interaction
por: Lou, Yuyang, et al.
Publicado: (2025)
por: Lou, Yuyang, et al.
Publicado: (2025)
AMemGym: Interactive Memory Benchmarking for Assistants in Long-Horizon Conversations
por: Jiayang, Cheng, et al.
Publicado: (2026)
por: Jiayang, Cheng, et al.
Publicado: (2026)
Leave No Context Behind: Efficient Infinite Context Transformers with Infini-attention
por: Munkhdalai, Tsendsuren, et al.
Publicado: (2024)
por: Munkhdalai, Tsendsuren, et al.
Publicado: (2024)
Periodic RoPE for Infinite Context LLMs
por: Huo, Simin
Publicado: (2026)
por: Huo, Simin
Publicado: (2026)
Context Is What You Need: The Maximum Effective Context Window for Real World Limits of LLMs
por: Paulsen, Norman
Publicado: (2025)
por: Paulsen, Norman
Publicado: (2025)
BABILong: Testing the Limits of LLMs with Long Context Reasoning-in-a-Haystack
por: Kuratov, Yuri, et al.
Publicado: (2024)
por: Kuratov, Yuri, et al.
Publicado: (2024)
Hardware-aligned Hierarchical Sparse Attention for Efficient Long-term Memory Access
por: Hu, Xiang, et al.
Publicado: (2025)
por: Hu, Xiang, et al.
Publicado: (2025)
In Prospect and Retrospect: Reflective Memory Management for Long-term Personalized Dialogue Agents
por: Tan, Zhen, et al.
Publicado: (2025)
por: Tan, Zhen, et al.
Publicado: (2025)
PSC: Extending Context Window of Large Language Models via Phase Shift Calibration
por: Zhu, Wenqiao, et al.
Publicado: (2025)
por: Zhu, Wenqiao, et al.
Publicado: (2025)
Ejemplares similares
-
GRAVITY: Architecture-Agnostic Structured Anchoring for Long-Horizon Conversational Memory
por: Sun, Yushi, et al.
Publicado: (2026) -
On the Worst Prompt Performance of Large Language Models
por: Cao, Bowen, et al.
Publicado: (2024) -
Stress-Testing Long-Context Language Models with Lifelong ICL and Task Haystack
por: Xu, Xiaoyue, et al.
Publicado: (2024) -
InftyThink: Breaking the Length Limits of Long-Context Reasoning in Large Language Models
por: Yan, Yuchen, et al.
Publicado: (2025) -
Commonsense-augmented Memory Construction and Management in Long-term Conversations via Context-aware Persona Refinement
por: Kim, Hana, et al.
Publicado: (2024)