MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems
Fuente:
arXiv
Saved in:
| Main Authors: | Ai, Qingyao, Tang, Yichen, Wang, Changyue, Long, Jianming, Su, Weihang, Liu, Yiqun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
JuDGE: Benchmarking Judgment Document Generation for Chinese Legal System
by: Su, Weihang, et al.
Published: (2025)
by: Su, Weihang, et al.
Published: (2025)
SurGE: A Benchmark and Evaluation Framework for Scientific Survey Generation
by: Su, Weihang, et al.
Published: (2025)
by: Su, Weihang, et al.
Published: (2025)
Mitigating Entity-Level Hallucination in Large Language Models
by: Su, Weihang, et al.
Published: (2024)
by: Su, Weihang, et al.
Published: (2024)
Enhancing Judgment Document Generation via Agentic Legal Information Collection and Rubric-Guided Optimization
by: Su, Weihang, et al.
Published: (2026)
by: Su, Weihang, et al.
Published: (2026)
Analytical Search
by: Tu, Yiteng, et al.
Published: (2026)
by: Tu, Yiteng, et al.
Published: (2026)
Multi-Field Tool Retrieval
by: Tang, Yichen, et al.
Published: (2026)
by: Tang, Yichen, et al.
Published: (2026)
Wikiformer: Pre-training with Structured Information of Wikipedia for Ad-hoc Retrieval
by: Su, Weihang, et al.
Published: (2023)
by: Su, Weihang, et al.
Published: (2023)
CrossPT-EEG: A Benchmark for Cross-Participant and Cross-Time Generalization of EEG-based Visual Decoding
by: Zhu, Shuqi, et al.
Published: (2024)
by: Zhu, Shuqi, et al.
Published: (2024)
Investigating the Robustness of Counterfactual Learning to Rank Models: A Reproducibility Study
by: Niu, Zechun, et al.
Published: (2024)
by: Niu, Zechun, et al.
Published: (2024)
Foundations of GenIR
by: Ai, Qingyao, et al.
Published: (2025)
by: Ai, Qingyao, et al.
Published: (2025)
Towards Unification of Hallucination Detection and Fact Verification for Large Language Models
by: Su, Weihang, et al.
Published: (2025)
by: Su, Weihang, et al.
Published: (2025)
DRAGIN: Dynamic Retrieval Augmented Generation based on the Information Needs of Large Language Models
by: Su, Weihang, et al.
Published: (2024)
by: Su, Weihang, et al.
Published: (2024)
NextMem: Towards Latent Factual Memory for LLM-based Agents
by: Zhang, Zeyu, et al.
Published: (2026)
by: Zhang, Zeyu, et al.
Published: (2026)
RRCM: Ranking-Driven Retrieval over Collaborative and Meta Memories for LLM Recommendation
by: Li, Shijun, et al.
Published: (2026)
by: Li, Shijun, et al.
Published: (2026)
Mem-Rec: Memory Efficient Recommendation System using Alternative Representation
by: Jha, Gopi Krishna, et al.
Published: (2023)
by: Jha, Gopi Krishna, et al.
Published: (2023)
Learn to Memorize: Optimizing LLM-based Agents with Adaptive Memory Framework
by: Zhang, Zeyu, et al.
Published: (2025)
by: Zhang, Zeyu, et al.
Published: (2025)
Human-Inspired Memory Architecture for LLM Agents
by: Kerestecioglu, Doga, et al.
Published: (2026)
by: Kerestecioglu, Doga, et al.
Published: (2026)
Parametric Retrieval Augmented Generation
by: Su, Weihang, et al.
Published: (2025)
by: Su, Weihang, et al.
Published: (2025)
Long Context Modeling with Ranked Memory-Augmented Retrieval
by: Alselwi, Ghadir, et al.
Published: (2025)
by: Alselwi, Ghadir, et al.
Published: (2025)
Test-Time Training for Zero-Resource Dense Retrieval Reranking
by: Liu, Shiyan, et al.
Published: (2026)
by: Liu, Shiyan, et al.
Published: (2026)
AriadneMem: Threading the Maze of Lifelong Memory for LLM Agents
by: Zhu, Wenhui, et al.
Published: (2026)
by: Zhu, Wenhui, et al.
Published: (2026)
Improve Large Language Model Systems with User Logs
by: Wang, Changyue, et al.
Published: (2026)
by: Wang, Changyue, et al.
Published: (2026)
Advancing Vietnamese Information Retrieval with Learning Objective and Benchmark
by: Nguyen, Phu-Vinh, et al.
Published: (2025)
by: Nguyen, Phu-Vinh, et al.
Published: (2025)
AD-Bench: A Real-World, Trajectory-Aware Advertising Analytics Benchmark for LLM Agents
by: Hu, Lingxiang, et al.
Published: (2026)
by: Hu, Lingxiang, et al.
Published: (2026)
Graph Hopfield Networks: Energy-Based Node Classification with Associative Memory
by: Rao, Abinav, et al.
Published: (2026)
by: Rao, Abinav, et al.
Published: (2026)
Improved Adaboost Algorithm for Web Advertisement Click Prediction Based on Long Short-Term Memory Networks
by: Yu, Qixuan, et al.
Published: (2024)
by: Yu, Qixuan, et al.
Published: (2024)
Equity vs. Equality: Optimizing Ranking Fairness for Tailored Provider Needs
by: Tu, Yiteng, et al.
Published: (2026)
by: Tu, Yiteng, et al.
Published: (2026)
General Agentic Memory Via Deep Research
by: Yan, B. Y., et al.
Published: (2025)
by: Yan, B. Y., et al.
Published: (2025)
InterFormer: Effective Heterogeneous Interaction Learning for Click-Through Rate Prediction
by: Zeng, Zhichen, et al.
Published: (2024)
by: Zeng, Zhichen, et al.
Published: (2024)
CirrusBench: Evaluating LLM-based Agents Beyond Correctness in Real-World Cloud Service Environments
by: Yu, Yi, et al.
Published: (2026)
by: Yu, Yi, et al.
Published: (2026)
When LLM Reward Design Fails: Diagnostic-Driven Refinement for Sparse Structured RL
by: Wang, Youting, et al.
Published: (2026)
by: Wang, Youting, et al.
Published: (2026)
DiffGRM: Diffusion-based Generative Recommendation Model
by: Liu, Zhao, et al.
Published: (2025)
by: Liu, Zhao, et al.
Published: (2025)
LLM-assisted Vector Similarity Search
by: Riyadh, Md, et al.
Published: (2024)
by: Riyadh, Md, et al.
Published: (2024)
ECAT: A Entire space Continual and Adaptive Transfer Learning Framework for Cross-Domain Recommendation
by: Hou, Chaoqun, et al.
Published: (2024)
by: Hou, Chaoqun, et al.
Published: (2024)
Preference Discerning with LLM-Enhanced Generative Retrieval
by: Paischer, Fabian, et al.
Published: (2024)
by: Paischer, Fabian, et al.
Published: (2024)
Weightless Neural Networks for Continuously Trainable Personalized Recommendation Systems
by: Latif, Rafayel, et al.
Published: (2025)
by: Latif, Rafayel, et al.
Published: (2025)
Memory Assisted LLM for Personalized Recommendation System
by: Chen, Jiarui
Published: (2025)
by: Chen, Jiarui
Published: (2025)
R1-Ranker: Teaching LLM Rankers to Reason
by: Feng, Tao, et al.
Published: (2025)
by: Feng, Tao, et al.
Published: (2025)
Collaborative Filtering using Variational Quantum Hopfield Associative Memory
by: Kermanshahani, Amir, et al.
Published: (2025)
by: Kermanshahani, Amir, et al.
Published: (2025)
RL-MPCA: A Reinforcement Learning Based Multi-Phase Computation Allocation Approach for Recommender Systems
by: Zhou, Jiahong, et al.
Published: (2023)
by: Zhou, Jiahong, et al.
Published: (2023)
Similar Items
-
JuDGE: Benchmarking Judgment Document Generation for Chinese Legal System
by: Su, Weihang, et al.
Published: (2025) -
SurGE: A Benchmark and Evaluation Framework for Scientific Survey Generation
by: Su, Weihang, et al.
Published: (2025) -
Mitigating Entity-Level Hallucination in Large Language Models
by: Su, Weihang, et al.
Published: (2024) -
Enhancing Judgment Document Generation via Agentic Legal Information Collection and Rubric-Guided Optimization
by: Su, Weihang, et al.
Published: (2026) -
Analytical Search
by: Tu, Yiteng, et al.
Published: (2026)