PERMA: Benchmarking Personalized Memory Agents via Event-Driven Preference and Realistic Task Environments
Fuente:
arXiv
Salvato in:
| Autori principali: | Liu, Shuochen, Zhu, Junyi, Shu, Long, Lin, Junda, Chen, Yuhao, Zhang, Haotian, Zhang, Chao, Xu, Derong, Li, Jia, Tang, Bo, Li, Zhiyu, Xiong, Feiyu, Chen, Enhong, Xu, Tong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
TeaRAG: A Token-Efficient Agentic Retrieval-Augmented Generation Framework
di: Zhang, Chao, et al.
Pubblicazione: (2025)
di: Zhang, Chao, et al.
Pubblicazione: (2025)
Look as You Think: Unifying Reasoning and Visual Evidence Attribution for Verifiable Document RAG via Reinforcement Learning
di: Liu, Shuochen, et al.
Pubblicazione: (2025)
di: Liu, Shuochen, et al.
Pubblicazione: (2025)
VIGIL: Defending LLM Agents Against Tool Stream Injection via Verify-Before-Commit
di: Lin, Junda, et al.
Pubblicazione: (2026)
di: Lin, Junda, et al.
Pubblicazione: (2026)
Learning How and What to Memorize: Cognition-Inspired Two-Stage Optimization for Evolving Memory
di: Xu, Derong, et al.
Pubblicazione: (2026)
di: Xu, Derong, et al.
Pubblicazione: (2026)
FastMem: Fast Memorization of Prompt Improves Context Awareness of Large Language Models
di: Zhu, Junyi, et al.
Pubblicazione: (2024)
di: Zhu, Junyi, et al.
Pubblicazione: (2024)
VehicleMemBench: An Executable Benchmark for Multi-User Long-Term Memory in In-Vehicle Agents
di: Chen, Yuhao, et al.
Pubblicazione: (2026)
di: Chen, Yuhao, et al.
Pubblicazione: (2026)
CRUD-RAG: A Comprehensive Chinese Benchmark for Retrieval-Augmented Generation of Large Language Models
di: Lyu, Yuanjie, et al.
Pubblicazione: (2024)
di: Lyu, Yuanjie, et al.
Pubblicazione: (2024)
Communication-Efficient Personalized Federated Learning for Speech-to-Text Tasks
di: Du, Yichao, et al.
Pubblicazione: (2024)
di: Du, Yichao, et al.
Pubblicazione: (2024)
MemPrivacy: Privacy-Preserving Personalized Memory Management for Edge-Cloud Agents
di: Chen, Yining, et al.
Pubblicazione: (2026)
di: Chen, Yining, et al.
Pubblicazione: (2026)
Inside Out: Evolving User-Centric Core Memory Trees for Long-Term Personalized Dialogue Systems
di: Zhao, Jihao, et al.
Pubblicazione: (2026)
di: Zhao, Jihao, et al.
Pubblicazione: (2026)
Xiangqi-R1: Enhancing Spatial Strategic Reasoning in LLMs for Chinese Chess via Reinforcement Learning
di: Chen, Yuhao, et al.
Pubblicazione: (2025)
di: Chen, Yuhao, et al.
Pubblicazione: (2025)
Evoking User Memory: Personalizing LLM via Recollection-Familiarity Adaptive Retrieval
di: Zhang, Yingyi, et al.
Pubblicazione: (2026)
di: Zhang, Yingyi, et al.
Pubblicazione: (2026)
More Edits, More Stable: Understanding the Lifelong Normalization in Sequential Model Editing
di: Ma, Xin, et al.
Pubblicazione: (2026)
di: Ma, Xin, et al.
Pubblicazione: (2026)
MemReader: From Passive to Active Extraction for Long-Term Agent Memory
di: Kang, Jingyi, et al.
Pubblicazione: (2026)
di: Kang, Jingyi, et al.
Pubblicazione: (2026)
Text2Mem: A Unified Memory Operation Language for Memory Operating System
di: Wang, Yi, et al.
Pubblicazione: (2025)
di: Wang, Yi, et al.
Pubblicazione: (2025)
MemFactory: Unified Inference & Training Framework for Agent Memory
di: Guo, Ziliang, et al.
Pubblicazione: (2026)
di: Guo, Ziliang, et al.
Pubblicazione: (2026)
How Does Personalized Memory Shape LLM Behavior? Benchmarking Rational Preference Utilization in Personalized Assistants
di: Feng, Xueyang, et al.
Pubblicazione: (2026)
di: Feng, Xueyang, et al.
Pubblicazione: (2026)
From Single to Multi-Granularity: Toward Long-Term Memory Association and Selection of Conversational Agents
di: Xu, Derong, et al.
Pubblicazione: (2025)
di: Xu, Derong, et al.
Pubblicazione: (2025)
MemReranker: Reasoning-Aware Reranking for Agent Memory Retrieval
di: Li, Chunyu, et al.
Pubblicazione: (2026)
di: Li, Chunyu, et al.
Pubblicazione: (2026)
Harnessing Large Language Models for Knowledge Graph Question Answering via Adaptive Multi-Aspect Retrieval-Augmentation
di: Xu, Derong, et al.
Pubblicazione: (2024)
di: Xu, Derong, et al.
Pubblicazione: (2024)
Multi-perspective Improvement of Knowledge Graph Completion with Large Language Models
di: Xu, Derong, et al.
Pubblicazione: (2024)
di: Xu, Derong, et al.
Pubblicazione: (2024)
Large Language Model based Long-tail Query Rewriting in Taobao Search
di: Peng, Wenjun, et al.
Pubblicazione: (2023)
di: Peng, Wenjun, et al.
Pubblicazione: (2023)
Streamlining the Collaborative Chain of Models into A Single Forward Pass in Generation-Based Tasks
di: Lyu, Yuanjie, et al.
Pubblicazione: (2025)
di: Lyu, Yuanjie, et al.
Pubblicazione: (2025)
Large Language Models for Generative Information Extraction: A Survey
di: Xu, Derong, et al.
Pubblicazione: (2023)
di: Xu, Derong, et al.
Pubblicazione: (2023)
HaluMem: Evaluating Hallucinations in Memory Systems of Agents
di: Chen, Ding, et al.
Pubblicazione: (2025)
di: Chen, Ding, et al.
Pubblicazione: (2025)
Benchmarking Large Language Models for Conversational Question Answering in Multi-instructional Documents
di: Wu, Shiwei, et al.
Pubblicazione: (2024)
di: Wu, Shiwei, et al.
Pubblicazione: (2024)
Adversarial Preference Learning for Robust LLM Alignment
di: Wang, Yuanfu, et al.
Pubblicazione: (2025)
di: Wang, Yuanfu, et al.
Pubblicazione: (2025)
MoM: Mixtures of Scenario-Aware Document Memories for Retrieval-Augmented Generation Systems
di: Zhao, Jihao, et al.
Pubblicazione: (2025)
di: Zhao, Jihao, et al.
Pubblicazione: (2025)
WebGym: Scaling Training Environments for Visual Web Agents with Realistic Tasks
di: Bai, Hao, et al.
Pubblicazione: (2026)
di: Bai, Hao, et al.
Pubblicazione: (2026)
MindBridge: Scalable and Cross-Model Knowledge Editing via Memory-Augmented Modality
di: Li, Shuaike, et al.
Pubblicazione: (2025)
di: Li, Shuaike, et al.
Pubblicazione: (2025)
Mitigating Hallucinations of Large Language Models in Medical Information Extraction via Contrastive Decoding
di: Xu, Derong, et al.
Pubblicazione: (2024)
di: Xu, Derong, et al.
Pubblicazione: (2024)
Beam Optics Ramping in Under-Constrained Lattice Design: Application to Electron-Ion Collider Hadron Storage Ring Cooling Section
di: Xu, Derong
Pubblicazione: (2025)
di: Xu, Derong
Pubblicazione: (2025)
MedMemoryBench: Benchmarking Agent Memory in Personalized Healthcare
di: Wang, Yihao, et al.
Pubblicazione: (2026)
di: Wang, Yihao, et al.
Pubblicazione: (2026)
EPBench: A Benchmark for Short-term Earthquake Prediction with Neural Networks
di: Xu, Zhiyu, et al.
Pubblicazione: (2025)
di: Xu, Zhiyu, et al.
Pubblicazione: (2025)
Generating Event-oriented Attribution for Movies via Two-Stage Prefix-Enhanced Multimodal LLM
di: Lyu, Yuanjie, et al.
Pubblicazione: (2024)
di: Lyu, Yuanjie, et al.
Pubblicazione: (2024)
Introducing EEG Analyses to Help Personal Music Preference Prediction
di: He, Zhiyu, et al.
Pubblicazione: (2024)
di: He, Zhiyu, et al.
Pubblicazione: (2024)
CodeFlowBench: A Multi-turn, Iterative Benchmark for Complex Code Generation
di: Wang, Sizhe, et al.
Pubblicazione: (2025)
di: Wang, Sizhe, et al.
Pubblicazione: (2025)
Align-GRAG: Anchor and Rationale Guided Dual Alignment for Graph Retrieval-Augmented Generation
di: Xu, Derong, et al.
Pubblicazione: (2025)
di: Xu, Derong, et al.
Pubblicazione: (2025)
Towards Realistic Personalization: Evaluating Long-Horizon Preference Following in Personalized User-LLM Interactions
di: Guo, Qianyun, et al.
Pubblicazione: (2026)
di: Guo, Qianyun, et al.
Pubblicazione: (2026)
EPO: Hierarchical LLM Agents with Environment Preference Optimization
di: Zhao, Qi, et al.
Pubblicazione: (2024)
di: Zhao, Qi, et al.
Pubblicazione: (2024)
Documenti analoghi
-
TeaRAG: A Token-Efficient Agentic Retrieval-Augmented Generation Framework
di: Zhang, Chao, et al.
Pubblicazione: (2025) -
Look as You Think: Unifying Reasoning and Visual Evidence Attribution for Verifiable Document RAG via Reinforcement Learning
di: Liu, Shuochen, et al.
Pubblicazione: (2025) -
VIGIL: Defending LLM Agents Against Tool Stream Injection via Verify-Before-Commit
di: Lin, Junda, et al.
Pubblicazione: (2026) -
Learning How and What to Memorize: Cognition-Inspired Two-Stage Optimization for Evolving Memory
di: Xu, Derong, et al.
Pubblicazione: (2026) -
FastMem: Fast Memorization of Prompt Improves Context Awareness of Large Language Models
di: Zhu, Junyi, et al.
Pubblicazione: (2024)