Learning What to Remember: Adaptive Probabilistic Memory Retention for Memory-Efficient Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Rafiuddin, S M, Khan, Muntaha Nujat |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Structured Token Retention and Computational Memory Paths in Large Language Models
di: Delena, Jonathan, et al.
Pubblicazione: (2025)
di: Delena, Jonathan, et al.
Pubblicazione: (2025)
Ranking of Bangla Word Graph using Graph-based Ranking Algorithms
di: Rafiuddin, S M
Pubblicazione: (2025)
di: Rafiuddin, S M
Pubblicazione: (2025)
A Long Short-Term Memory (LSTM) Model for Business Sentiment Analysis Based on Recurrent Neural Network
di: Razin, Md. Jahidul Islam, et al.
Pubblicazione: (2025)
di: Razin, Md. Jahidul Islam, et al.
Pubblicazione: (2025)
Performance Analysis of Supervised Machine Learning Algorithms for Text Classification
di: Mishu, Sadia Zaman, et al.
Pubblicazione: (2025)
di: Mishu, Sadia Zaman, et al.
Pubblicazione: (2025)
Retention Consequence in Lifecycle Memory Control
di: Han, Jiarui
Pubblicazione: (2026)
di: Han, Jiarui
Pubblicazione: (2026)
Rethinking How to Remember: Beyond Atomic Facts in Lifelong LLM Agent Memory
di: Sun, Jingwei, et al.
Pubblicazione: (2026)
di: Sun, Jingwei, et al.
Pubblicazione: (2026)
Exploiting Adaptive Contextual Masking for Aspect-Based Sentiment Analysis
di: Rafiuddin, S M, et al.
Pubblicazione: (2024)
di: Rafiuddin, S M, et al.
Pubblicazione: (2024)
Mem-$π$: Adaptive Memory through Learning When and What to Generate
di: Wang, Xiaoqiang, et al.
Pubblicazione: (2026)
di: Wang, Xiaoqiang, et al.
Pubblicazione: (2026)
Adaptive Focus Memory for Language Models
di: Cruz, Christopher
Pubblicazione: (2025)
di: Cruz, Christopher
Pubblicazione: (2025)
Language Model Memory and Memory Models for Language
di: Badger, Benjamin L.
Pubblicazione: (2026)
di: Badger, Benjamin L.
Pubblicazione: (2026)
If Attention Serves as a Cognitive Model of Human Memory Retrieval, What is the Plausible Memory Representation?
di: Yoshida, Ryo, et al.
Pubblicazione: (2025)
di: Yoshida, Ryo, et al.
Pubblicazione: (2025)
Remembering More, Risking More: Longitudinal Safety Risks in Memory-Equipped LLM Agents
di: Al-Tawaha, Ahmad, et al.
Pubblicazione: (2026)
di: Al-Tawaha, Ahmad, et al.
Pubblicazione: (2026)
AdaptiSent: Context-Aware Adaptive Attention for Multimodal Aspect-Based Sentiment Analysis
di: Rafiuddin, S M, et al.
Pubblicazione: (2025)
di: Rafiuddin, S M, et al.
Pubblicazione: (2025)
MLP Memory: A Retriever-Pretrained Memory for Large Language Models
di: Wei, Rubin, et al.
Pubblicazione: (2025)
di: Wei, Rubin, et al.
Pubblicazione: (2025)
MINI-LLM: Memory-Efficient Structured Pruning for Large Language Models
di: Cheng, Hongrong, et al.
Pubblicazione: (2024)
di: Cheng, Hongrong, et al.
Pubblicazione: (2024)
KVPruner: Structural Pruning for Faster and Memory-Efficient Large Language Models
di: Lv, Bo, et al.
Pubblicazione: (2024)
di: Lv, Bo, et al.
Pubblicazione: (2024)
What Languages are Easy to Language-Model? A Perspective from Learning Probabilistic Regular Languages
di: Borenstein, Nadav, et al.
Pubblicazione: (2024)
di: Borenstein, Nadav, et al.
Pubblicazione: (2024)
Memory-based Language Models: An Efficient, Explainable, and Eco-friendly Approach to Large Language Modeling
di: Bosch, Antal van den, et al.
Pubblicazione: (2025)
di: Bosch, Antal van den, et al.
Pubblicazione: (2025)
Memory-Efficient Looped Transformer: Decoupling Compute from Memory in Looped Language Models
di: Vendrell, Victor Conchello, et al.
Pubblicazione: (2026)
di: Vendrell, Victor Conchello, et al.
Pubblicazione: (2026)
In-Memory Learning: A Declarative Learning Framework for Large Language Models
di: Wang, Bo, et al.
Pubblicazione: (2024)
di: Wang, Bo, et al.
Pubblicazione: (2024)
How LoRA Remembers? A Parametric Memory Law for LLM Finetuning
di: Xu, Ziwen, et al.
Pubblicazione: (2026)
di: Xu, Ziwen, et al.
Pubblicazione: (2026)
Agentic Memory: Learning Unified Long-Term and Short-Term Memory Management for Large Language Model Agents
di: Yu, Yi, et al.
Pubblicazione: (2026)
di: Yu, Yi, et al.
Pubblicazione: (2026)
What Are the Odds? Language Models Are Capable of Probabilistic Reasoning
di: Paruchuri, Akshay, et al.
Pubblicazione: (2024)
di: Paruchuri, Akshay, et al.
Pubblicazione: (2024)
Adaptive Memory Replay for Continual Learning
di: Smith, James Seale, et al.
Pubblicazione: (2024)
di: Smith, James Seale, et al.
Pubblicazione: (2024)
Schrodinger's Memory: Large Language Models
di: Wang, Wei, et al.
Pubblicazione: (2024)
di: Wang, Wei, et al.
Pubblicazione: (2024)
How Do Multilingual Language Models Remember Facts?
di: Fierro, Constanza, et al.
Pubblicazione: (2024)
di: Fierro, Constanza, et al.
Pubblicazione: (2024)
Remember Me, Refine Me: A Dynamic Procedural Memory Framework for Experience-Driven Agent Evolution
di: Cao, Zouying, et al.
Pubblicazione: (2025)
di: Cao, Zouying, et al.
Pubblicazione: (2025)
Memory Grafting: Scaling Language Model Pre-training via Offline Conditional Memory
di: Cheng, Runxi, et al.
Pubblicazione: (2026)
di: Cheng, Runxi, et al.
Pubblicazione: (2026)
Language Modeling With Factorization Memory
di: Xiong, Lee, et al.
Pubblicazione: (2025)
di: Xiong, Lee, et al.
Pubblicazione: (2025)
R$^3$Mem: Bridging Memory Retention and Retrieval via Reversible Compression
di: Wang, Xiaoqiang, et al.
Pubblicazione: (2025)
di: Wang, Xiaoqiang, et al.
Pubblicazione: (2025)
Dynamic Cheatsheet: Test-Time Learning with Adaptive Memory
di: Suzgun, Mirac, et al.
Pubblicazione: (2025)
di: Suzgun, Mirac, et al.
Pubblicazione: (2025)
Text2Mem: A Unified Memory Operation Language for Memory Operating System
di: Wang, Yi, et al.
Pubblicazione: (2025)
di: Wang, Yi, et al.
Pubblicazione: (2025)
Memory-R1: Enhancing Large Language Model Agents to Manage and Utilize Memories via Reinforcement Learning
di: Yan, Sikuan, et al.
Pubblicazione: (2025)
di: Yan, Sikuan, et al.
Pubblicazione: (2025)
Memory Decoder: A Pretrained, Plug-and-Play Memory for Large Language Models
di: Cao, Jiaqi, et al.
Pubblicazione: (2025)
di: Cao, Jiaqi, et al.
Pubblicazione: (2025)
TernaryLM: Memory-Efficient Language Modeling via Native 1.5-Bit Quantization with Adaptive Layer-wise Scaling
di: Nargund, Nisharg, et al.
Pubblicazione: (2026)
di: Nargund, Nisharg, et al.
Pubblicazione: (2026)
Cognitive Memory in Large Language Models
di: Shan, Lianlei, et al.
Pubblicazione: (2025)
di: Shan, Lianlei, et al.
Pubblicazione: (2025)
The Mosaic Memory of Large Language Models
di: Shilov, Igor, et al.
Pubblicazione: (2024)
di: Shilov, Igor, et al.
Pubblicazione: (2024)
What Training Data Teaches RL Memory Agents: An Empirical Study of Curriculum Effects in Memory-Augmented QA
di: He, Xinjie, et al.
Pubblicazione: (2026)
di: He, Xinjie, et al.
Pubblicazione: (2026)
Interweaving Memories of a Siamese Large Language Model
di: Song, Xin, et al.
Pubblicazione: (2024)
di: Song, Xin, et al.
Pubblicazione: (2024)
Disentangling Memory and Reasoning Ability in Large Language Models
di: Jin, Mingyu, et al.
Pubblicazione: (2024)
di: Jin, Mingyu, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Structured Token Retention and Computational Memory Paths in Large Language Models
di: Delena, Jonathan, et al.
Pubblicazione: (2025) -
Ranking of Bangla Word Graph using Graph-based Ranking Algorithms
di: Rafiuddin, S M
Pubblicazione: (2025) -
A Long Short-Term Memory (LSTM) Model for Business Sentiment Analysis Based on Recurrent Neural Network
di: Razin, Md. Jahidul Islam, et al.
Pubblicazione: (2025) -
Performance Analysis of Supervised Machine Learning Algorithms for Text Classification
di: Mishu, Sadia Zaman, et al.
Pubblicazione: (2025) -
Retention Consequence in Lifecycle Memory Control
di: Han, Jiarui
Pubblicazione: (2026)