Salvato in:
| Autori principali: | Ren, Xiyu, Wang, Zhaowei, Du, Yiming, Xie, Zhongwei, Liu, Chi, Yang, Xinlin, Feng, Haoyue, Pan, Wenjun, Zheng, Tianshi, Xu, Baixuan, Li, Zhengnan, Song, Yangqiu, Wong, Ginny, See, Simon |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2605.14906 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
MMLongBench: Benchmarking Long-Context Vision-Language Models Effectively and Thoroughly
di: Wang, Zhaowei, et al.
Pubblicazione: (2025)
di: Wang, Zhaowei, et al.
Pubblicazione: (2025)
Can Retrieval Heads See Images? Multimodal Retrieval Heads in Long-Context Vision-Language Models
di: Li, Aaron Branson Cigres, et al.
Pubblicazione: (2026)
di: Li, Aaron Branson Cigres, et al.
Pubblicazione: (2026)
ComparisonQA: Evaluating Factuality Robustness of LLMs Through Knowledge Frequency Control and Uncertainty
di: Zong, Qing, et al.
Pubblicazione: (2024)
di: Zong, Qing, et al.
Pubblicazione: (2024)
The Curse of CoT: On the Limitations of Chain-of-Thought in In-Context Learning
di: Zheng, Tianshi, et al.
Pubblicazione: (2025)
di: Zheng, Tianshi, et al.
Pubblicazione: (2025)
The Cognitive Bandwidth Bottleneck: Shifting Long-Horizon Agent from Planning with Actions to Planning with Schemas
di: Xu, Baixuan, et al.
Pubblicazione: (2025)
di: Xu, Baixuan, et al.
Pubblicazione: (2025)
NewtonBench: Benchmarking Generalizable Scientific Law Discovery in LLM Agents
di: Zheng, Tianshi, et al.
Pubblicazione: (2025)
di: Zheng, Tianshi, et al.
Pubblicazione: (2025)
LogiDynamics: Unraveling the Dynamics of Inductive, Abductive and Deductive Logical Inferences in LLM Reasoning
di: Zheng, Tianshi, et al.
Pubblicazione: (2025)
di: Zheng, Tianshi, et al.
Pubblicazione: (2025)
Mem-Gallery: Benchmarking Multimodal Long-Term Conversational Memory for MLLM Agents
di: Bei, Yuanchen, et al.
Pubblicazione: (2026)
di: Bei, Yuanchen, et al.
Pubblicazione: (2026)
DixitWorld: Evaluating Multimodal Abductive Reasoning in Vision-Language Models with Multi-Agent Dixit Gameplay
di: Mo, Yunxiang, et al.
Pubblicazione: (2025)
di: Mo, Yunxiang, et al.
Pubblicazione: (2025)
CloneMem: Benchmarking Long-Term Memory for AI Clones
di: Hu, Sen, et al.
Pubblicazione: (2026)
di: Hu, Sen, et al.
Pubblicazione: (2026)
$\mathbb{R}^{2k}$ is Theoretically Large Enough for Embedding-based Top-$k$ Retrieval
di: Wang, Zihao, et al.
Pubblicazione: (2026)
di: Wang, Zihao, et al.
Pubblicazione: (2026)
AbsInstruct: Eliciting Abstraction Ability from LLMs through Explanation Tuning with Plausibility Estimation
di: Wang, Zhaowei, et al.
Pubblicazione: (2024)
di: Wang, Zhaowei, et al.
Pubblicazione: (2024)
LongMemEval: Benchmarking Chat Assistants on Long-Term Interactive Memory
di: Wu, Di, et al.
Pubblicazione: (2024)
di: Wu, Di, et al.
Pubblicazione: (2024)
TeleMem: Building Long-Term and Multimodal Memory for Agentic AI
di: Chen, Chunliang, et al.
Pubblicazione: (2025)
di: Chen, Chunliang, et al.
Pubblicazione: (2025)
CritiCal: Can Critique Help LLM Uncertainty or Confidence Calibration?
di: Zong, Qing, et al.
Pubblicazione: (2025)
di: Zong, Qing, et al.
Pubblicazione: (2025)
HyperMem: Hypergraph Memory for Long-Term Conversations
di: Yue, Juwei, et al.
Pubblicazione: (2026)
di: Yue, Juwei, et al.
Pubblicazione: (2026)
VehicleMemBench: An Executable Benchmark for Multi-User Long-Term Memory in In-Vehicle Agents
di: Chen, Yuhao, et al.
Pubblicazione: (2026)
di: Chen, Yuhao, et al.
Pubblicazione: (2026)
MemConflict: Evaluating Long-Term Memory Systems Under Memory Conflicts
di: Tao, Zhen, et al.
Pubblicazione: (2026)
di: Tao, Zhen, et al.
Pubblicazione: (2026)
HiMem: Hierarchical Long-Term Memory for LLM Long-Horizon Agents
di: Zhang, Ningning, et al.
Pubblicazione: (2026)
di: Zhang, Ningning, et al.
Pubblicazione: (2026)
AutoGraph-R1: End-to-End Reinforcement Learning for Knowledge Graph Construction
di: Tsang, Hong Ting, et al.
Pubblicazione: (2025)
di: Tsang, Hong Ting, et al.
Pubblicazione: (2025)
MemRouter: Memory-as-Embedding Routing for Long-Term Conversational Agents
di: Hu, Tianyu, et al.
Pubblicazione: (2026)
di: Hu, Tianyu, et al.
Pubblicazione: (2026)
DimMem: Dimensional Structuring for Efficient Long-Term Agent Memory
di: Qiu, Wentao, et al.
Pubblicazione: (2026)
di: Qiu, Wentao, et al.
Pubblicazione: (2026)
Towards Multi-Agent Reasoning Systems for Collaborative Expertise Delegation: An Exploratory Design Study
di: Xu, Baixuan, et al.
Pubblicazione: (2025)
di: Xu, Baixuan, et al.
Pubblicazione: (2025)
INFERENCEDYNAMICS: Efficient Routing Across LLMs through Structured Capability and Knowledge Profiling
di: Shi, Haochen, et al.
Pubblicazione: (2025)
di: Shi, Haochen, et al.
Pubblicazione: (2025)
MemReader: From Passive to Active Extraction for Long-Term Agent Memory
di: Kang, Jingyi, et al.
Pubblicazione: (2026)
di: Kang, Jingyi, et al.
Pubblicazione: (2026)
ES-Mem: Event Segmentation-Based Memory for Long-Term Dialogue Agents
di: Zou, Huhai, et al.
Pubblicazione: (2026)
di: Zou, Huhai, et al.
Pubblicazione: (2026)
MemGuard: Preventing Memory Contamination in Long-Term Memory-Augmented Large Language Models
di: Ha, Hyeonjeong, et al.
Pubblicazione: (2026)
di: Ha, Hyeonjeong, et al.
Pubblicazione: (2026)
Persona Knowledge-Aligned Prompt Tuning Method for Online Debate
di: Chan, Chunkit, et al.
Pubblicazione: (2024)
di: Chan, Chunkit, et al.
Pubblicazione: (2024)
Mem2ActBench: A Benchmark for Evaluating Long-Term Memory Utilization in Task-Oriented Autonomous Agents
di: Shen, Yiting, et al.
Pubblicazione: (2026)
di: Shen, Yiting, et al.
Pubblicazione: (2026)
PsyMem: Fine-grained psychological alignment and Explicit Memory Control for Advanced Role-Playing LLMs
di: Cheng, Xilong, et al.
Pubblicazione: (2025)
di: Cheng, Xilong, et al.
Pubblicazione: (2025)
MemX: A Local-First Long-Term Memory System for AI Assistants
di: Sun, Lizheng
Pubblicazione: (2026)
di: Sun, Lizheng
Pubblicazione: (2026)
EviMem: Evidence-Gap-Driven Iterative Retrieval for Long-Term Conversational Memory
di: Li, Yuyang, et al.
Pubblicazione: (2026)
di: Li, Yuyang, et al.
Pubblicazione: (2026)
Mem0: Building Production-Ready AI Agents with Scalable Long-Term Memory
di: Chhikara, Prateek, et al.
Pubblicazione: (2025)
di: Chhikara, Prateek, et al.
Pubblicazione: (2025)
LongMemEval-V2: Evaluating Long-Term Agent Memory Toward Experienced Colleagues
di: Wu, Di, et al.
Pubblicazione: (2026)
di: Wu, Di, et al.
Pubblicazione: (2026)
Legal Rule Induction: Towards Generalizable Principle Discovery from Analogous Judicial Precedents
di: Fan, Wei, et al.
Pubblicazione: (2025)
di: Fan, Wei, et al.
Pubblicazione: (2025)
ES-MemEval: Benchmarking Conversational Agents on Personalized Long-Term Emotional Support
di: Chen, Tiantian, et al.
Pubblicazione: (2026)
di: Chen, Tiantian, et al.
Pubblicazione: (2026)
MemGround: Long-Term Memory Evaluation Kit for Large Language Models in Gamified Scenarios
di: Ding, Yihang, et al.
Pubblicazione: (2026)
di: Ding, Yihang, et al.
Pubblicazione: (2026)
MemBuilder: Reinforcing LLMs for Long-Term Memory Construction via Attributed Dense Rewards
di: Shen, Zhiyu, et al.
Pubblicazione: (2026)
di: Shen, Zhiyu, et al.
Pubblicazione: (2026)
CarMem: Enhancing Long-Term Memory in LLM Voice Assistants through Category-Bounding
di: Kirmayr, Johannes, et al.
Pubblicazione: (2025)
di: Kirmayr, Johannes, et al.
Pubblicazione: (2025)
TA-Mem: Tool-Augmented Autonomous Memory Retrieval for LLM in Long-Term Conversational QA
di: Yuan, Mengwei, et al.
Pubblicazione: (2026)
di: Yuan, Mengwei, et al.
Pubblicazione: (2026)
Documenti analoghi
-
MMLongBench: Benchmarking Long-Context Vision-Language Models Effectively and Thoroughly
di: Wang, Zhaowei, et al.
Pubblicazione: (2025) -
Can Retrieval Heads See Images? Multimodal Retrieval Heads in Long-Context Vision-Language Models
di: Li, Aaron Branson Cigres, et al.
Pubblicazione: (2026) -
ComparisonQA: Evaluating Factuality Robustness of LLMs Through Knowledge Frequency Control and Uncertainty
di: Zong, Qing, et al.
Pubblicazione: (2024) -
The Curse of CoT: On the Limitations of Chain-of-Thought in In-Context Learning
di: Zheng, Tianshi, et al.
Pubblicazione: (2025) -
The Cognitive Bandwidth Bottleneck: Shifting Long-Horizon Agent from Planning with Actions to Planning with Schemas
di: Xu, Baixuan, et al.
Pubblicazione: (2025)