MemoryArena: Benchmarking Agent Memory in Interdependent Multi-Session Agentic Tasks
Fuente:
arXiv
Saved in:
| Main Authors: | He, Zexue, Wang, Yu, Zhi, Churan, Hu, Yuanzhe, Chen, Tzu-Ping, Yin, Lang, Chen, Ze, Wu, Tong Arthur, Ouyang, Siru, Wang, Zihan, Pei, Jiaxin, McAuley, Julian, Choi, Yejin, Pentland, Alex |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Evaluating Memory in LLM Agents via Incremental Multi-Turn Interactions
by: Hu, Yuanzhe, et al.
Published: (2025)
by: Hu, Yuanzhe, et al.
Published: (2025)
M+: Extending MemoryLLM with Scalable Long-Term Memory
by: Wang, Yu, et al.
Published: (2025)
by: Wang, Yu, et al.
Published: (2025)
BiasFreeBench: a Benchmark for Mitigating Bias in Large Language Model Responses
by: Xu, Xin, et al.
Published: (2025)
by: Xu, Xin, et al.
Published: (2025)
CAMELoT: Towards Large Language Models with Training-Free Consolidated Associative Memory
by: He, Zexue, et al.
Published: (2024)
by: He, Zexue, et al.
Published: (2024)
LVCHAT: Facilitating Long Video Comprehension
by: Wang, Yu, et al.
Published: (2024)
by: Wang, Yu, et al.
Published: (2024)
Large Scale Knowledge Washing
by: Wang, Yu, et al.
Published: (2024)
by: Wang, Yu, et al.
Published: (2024)
Mem-α: Learning Memory Construction via Reinforcement Learning
by: Wang, Yu, et al.
Published: (2025)
by: Wang, Yu, et al.
Published: (2025)
On Generalization in Agentic Tool Calling: CoreThink Agentic Reasoner and MAVEN Dataset
by: Bhat, Vishvesh, et al.
Published: (2025)
by: Bhat, Vishvesh, et al.
Published: (2025)
Cognitive Bias in Decision-Making with LLMs
by: Echterhoff, Jessica, et al.
Published: (2024)
by: Echterhoff, Jessica, et al.
Published: (2024)
Multi-Behavior Generative Recommendation
by: Liu, Zihan, et al.
Published: (2024)
by: Liu, Zihan, et al.
Published: (2024)
Three central limit theorems for the unbounded excursion component of a Gaussian field
by: McAuley, Michael
Published: (2024)
by: McAuley, Michael
Published: (2024)
Children in Custody
by: McAuley, Mary
Published: (2022)
by: McAuley, Mary
Published: (2022)
Politics and the Soviet Union / Mary McAuley
by: McAuley, Mary
by: McAuley, Mary
Limit theorems for non-local functionals of smooth Gaussian fields via quasi-association
by: McAuley, Michael
Published: (2026)
by: McAuley, Michael
Published: (2026)
How Do AI Agents Spend Your Money? Analyzing and Predicting Token Consumption in Agentic Coding Tasks
by: Bai, Longju, et al.
Published: (2026)
by: Bai, Longju, et al.
Published: (2026)
InfoRank: Unbiased Learning-to-Rank via Conditional Mutual Information Minimization
by: Jin, Jiarui, et al.
Published: (2024)
by: Jin, Jiarui, et al.
Published: (2024)
Auto-Dreamer: Learning Offline Memory Consolidation for Language Agents
by: Ye, Chongrui, et al.
Published: (2026)
by: Ye, Chongrui, et al.
Published: (2026)
Momento: Evaluating Persistent Memory and Reasoning with Multi-Session Agentic Conversations
by: Merin, Adril Putra, et al.
Published: (2026)
by: Merin, Adril Putra, et al.
Published: (2026)
RoboMemArena: A Comprehensive and Challenging Robotic Memory Benchmark
by: Lei, Huashuo, et al.
Published: (2026)
by: Lei, Huashuo, et al.
Published: (2026)
Reddit2Deezer: A Scalable Dataset for Real-World Grounded Conversational Music Recommendation
by: Kim, Haven, et al.
Published: (2026)
by: Kim, Haven, et al.
Published: (2026)
Limit theorems for the number of sign and level-set clusters of the Gaussian free field
by: McAuley, Michael, et al.
Published: (2025)
by: McAuley, Michael, et al.
Published: (2025)
Educating Young Children: A Structural Approach. Routledge Library Editions: Early Years
by: McAuley, Helen, et al.
Published: (2022)
by: McAuley, Helen, et al.
Published: (2022)
ReCAP: Recursive Context-Aware Reasoning and Planning for Large Language Model Agents
by: Zhang, Zhenyu, et al.
Published: (2025)
by: Zhang, Zhenyu, et al.
Published: (2025)
Multi-Agent Collaborative Filtering: Orchestrating Users and Items for Agentic Recommendations
by: Xia, Yu, et al.
Published: (2025)
by: Xia, Yu, et al.
Published: (2025)
Episodic Memory in Agentic Frameworks: Suggesting Next Tasks
by: Fiorini, Sandro Rama, et al.
Published: (2025)
by: Fiorini, Sandro Rama, et al.
Published: (2025)
EvolMem: A Cognitive-Driven Benchmark for Multi-Session Dialogue Memory
by: Shen, Ye, et al.
Published: (2026)
by: Shen, Ye, et al.
Published: (2026)
GSPRec: Temporal-Aware Graph Spectral Filtering for Recommendation
by: Rabiah, Ahmad Bin, et al.
Published: (2025)
by: Rabiah, Ahmad Bin, et al.
Published: (2025)
TS-Memory: Plug-and-Play Memory for Time Series Foundation Models
by: Lyu, Sisuo, et al.
Published: (2026)
by: Lyu, Sisuo, et al.
Published: (2026)
AtomMem : Learnable Dynamic Agentic Memory with Atomic Memory Operation
by: Huo, Yupeng, et al.
Published: (2026)
by: Huo, Yupeng, et al.
Published: (2026)
AnnaAgent: Dynamic Evolution Agent System with Multi-Session Memory for Realistic Seeker Simulation
by: Wang, Ming, et al.
Published: (2025)
by: Wang, Ming, et al.
Published: (2025)
Mixed-Session Conversation with Egocentric Memory
by: Jang, Jihyoung, et al.
Published: (2024)
by: Jang, Jihyoung, et al.
Published: (2024)
PanguIR Technical Report for NTCIR-18 AEOLLM Task
by: Mei, Lang, et al.
Published: (2025)
by: Mei, Lang, et al.
Published: (2025)
Towards LifeSpan Cognitive Systems
by: Wang, Yu, et al.
Published: (2024)
by: Wang, Yu, et al.
Published: (2024)
Session-level Normalization and Click-through Data Enhancement for Session-based Evaluation
by: Chen, Haonan, et al.
Published: (2024)
by: Chen, Haonan, et al.
Published: (2024)
CoreThink: A Symbolic Reasoning Layer to reason over Long Horizon Tasks with LLMs
by: Vaghasiya, Jay, et al.
Published: (2025)
by: Vaghasiya, Jay, et al.
Published: (2025)
SAGE: A Self-Evolving Agentic Graph-Memory Engine for Structure-Aware Associative Memory
by: Wang, Juntong, et al.
Published: (2026)
by: Wang, Juntong, et al.
Published: (2026)
A covariance formula for the number of excursion set components of Gaussian fields and applications
by: Beliaev, Dmitry, et al.
Published: (2023)
by: Beliaev, Dmitry, et al.
Published: (2023)
A central limit theorem for the number of excursion set components of Gaussian fields
by: Beliaev, Dmitry, et al.
Published: (2022)
by: Beliaev, Dmitry, et al.
Published: (2022)
Extending Input Contexts of Language Models through Training on Segmented Sequences
by: Karypis, Petros, et al.
Published: (2023)
by: Karypis, Petros, et al.
Published: (2023)
FusID: Modality-Fused Semantic IDs for Generative Music Recommendation
by: Kim, Haven, et al.
Published: (2026)
by: Kim, Haven, et al.
Published: (2026)
Similar Items
-
Evaluating Memory in LLM Agents via Incremental Multi-Turn Interactions
by: Hu, Yuanzhe, et al.
Published: (2025) -
M+: Extending MemoryLLM with Scalable Long-Term Memory
by: Wang, Yu, et al.
Published: (2025) -
BiasFreeBench: a Benchmark for Mitigating Bias in Large Language Model Responses
by: Xu, Xin, et al.
Published: (2025) -
CAMELoT: Towards Large Language Models with Training-Free Consolidated Associative Memory
by: He, Zexue, et al.
Published: (2024) -
LVCHAT: Facilitating Long Video Comprehension
by: Wang, Yu, et al.
Published: (2024)