MemEye: A Visual-Centric Evaluation Framework for Multimodal Agent Memory
Fuente:
arXiv
Saved in:
| Main Authors: | Guo, Minghao, Jiao, Qingyue, Shi, Zeru, Quan, Yihao, Zhang, Boxuan, Li, Danrui, Che, Liwei, Xu, Wujiang, Liu, Shilong, Liu, Zirui, Kapadia, Mubbasir, Pavlovic, Vladimir, Liu, Jiang, Wang, Mengdi, Shi, Yiyu, Metaxas, Dimitris N., Tang, Ruixiang |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Improving Visual Reasoning with Iterative Evidence Refinement
by: Shi, Zeru, et al.
Published: (2026)
by: Shi, Zeru, et al.
Published: (2026)
TrajDiffuse: A Conditional Diffusion Model for Environment-Aware Trajectory Prediction
by: Qingze, et al.
Published: (2024)
by: Qingze, et al.
Published: (2024)
ArchSeek: Retrieving Architectural Case Studies Using Vision-Language Models
by: Li, Danrui, et al.
Published: (2025)
by: Li, Danrui, et al.
Published: (2025)
On the Equivalency, Substitutability, and Flexibility of Synthetic Data
by: Chang, Che-Jui, et al.
Published: (2024)
by: Chang, Che-Jui, et al.
Published: (2024)
Microscopic modeling of attention-based movement behaviors
by: Li, Danrui, et al.
Published: (2024)
by: Li, Danrui, et al.
Published: (2024)
AgentForesight: Online Auditing for Early Failure Prediction in Multi-Agent Systems
by: Zhang, Boxuan, et al.
Published: (2026)
by: Zhang, Boxuan, et al.
Published: (2026)
Counting Circuits: Mechanistic Interpretability of Visual Reasoning in Large Vision-Language Models
by: Che, Liwei, et al.
Published: (2026)
by: Che, Liwei, et al.
Published: (2026)
From Words to Worlds: Transforming One-line Prompt into Immersive Multi-modal Digital Stories with Communicative LLM Agent
by: Sohn, Samuel S., et al.
Published: (2024)
by: Sohn, Samuel S., et al.
Published: (2024)
CASIM: Composite Aware Semantic Injection for Text to Motion Generation
by: Chang, Che-Jui, et al.
Published: (2025)
by: Chang, Che-Jui, et al.
Published: (2025)
Enhancing Consistency Models for Multi-Agent Trajectory Prediction
by: Mrdovic, Alen, et al.
Published: (2026)
by: Mrdovic, Alen, et al.
Published: (2026)
Reinforcing Consistency in Video MLLMs with Structured Rewards
by: Quan, Yihao, et al.
Published: (2026)
by: Quan, Yihao, et al.
Published: (2026)
JACoP: Joint Alignment for Compliant Multi-Agent Prediction
by: Liu, Qingze, et al.
Published: (2026)
by: Liu, Qingze, et al.
Published: (2026)
Cardiverse: Harnessing LLMs for Novel Card Game Prototyping
by: Li, Danrui, et al.
Published: (2025)
by: Li, Danrui, et al.
Published: (2025)
M3Act: Learning from Synthetic Human Group Activities
by: Chang, Che-Jui, et al.
Published: (2023)
by: Chang, Che-Jui, et al.
Published: (2023)
Data Augmentation for High-Fidelity Generation of CAR-T/NK Immunological Synapse Images
by: Zhang, Xiang, et al.
Published: (2026)
by: Zhang, Xiang, et al.
Published: (2026)
Individual Turing Test: A Case Study of LLM-based Simulation Using Longitudinal Personal Data
by: Guo, Minghao, et al.
Published: (2026)
by: Guo, Minghao, et al.
Published: (2026)
Token-Controlled Re-ranking for Sequential Recommendation via LLMs
by: Dai, Wenxi, et al.
Published: (2025)
by: Dai, Wenxi, et al.
Published: (2025)
Pooling and Semantic Shift: The Fundamental Challenges in Long Text Embedding and Retrieval
by: Gao, Hang, et al.
Published: (2026)
by: Gao, Hang, et al.
Published: (2026)
Large Sign Language Models: Toward 3D American Sign Language Translation
by: Zhang, Sen, et al.
Published: (2025)
by: Zhang, Sen, et al.
Published: (2025)
Read the Scene, Not the Script: Outcome-Aware Safety for LLMs
by: Wu, Rui, et al.
Published: (2025)
by: Wu, Rui, et al.
Published: (2025)
Hallucinatory Image Tokens: A Training-free EAZY Approach on Detecting and Mitigating Object Hallucinations in LVLMs
by: Che, Liwei, et al.
Published: (2025)
by: Che, Liwei, et al.
Published: (2025)
AEL: Agent Evolving Learning for Open-Ended Environments
by: Xu, Wujiang, et al.
Published: (2026)
by: Xu, Wujiang, et al.
Published: (2026)
Trust or Abstain? A Self-Aware RAG Approach
by: Zhu, Xi, et al.
Published: (2026)
by: Zhu, Xi, et al.
Published: (2026)
A Single Layer to Explain Them All:Understanding Massive Activations in Large Language Models
by: Shi, Zeru, et al.
Published: (2026)
by: Shi, Zeru, et al.
Published: (2026)
Massive Values in Self-Attention Modules are the Key to Contextual Knowledge Understanding
by: Jin, Mingyu, et al.
Published: (2025)
by: Jin, Mingyu, et al.
Published: (2025)
From Commands to Prompts: LLM-based Semantic File System for AIOS
by: Shi, Zeru, et al.
Published: (2024)
by: Shi, Zeru, et al.
Published: (2024)
An Intrinsic Vector Heat Network
by: Gao, Alexander, et al.
Published: (2024)
by: Gao, Alexander, et al.
Published: (2024)
MediQ-GAN: Quantum-Inspired GAN for High Resolution Medical Image Generation
by: Jiao, Qingyue, et al.
Published: (2025)
by: Jiao, Qingyue, et al.
Published: (2025)
CARNet: Collaborative Adversarial Resilience for Robust Underwater Image Enhancement and Perception
by: Zhang, Zengxi, et al.
Published: (2023)
by: Zhang, Zengxi, et al.
Published: (2023)
A comparison on constrain encoding methods for quantum approximate optimization algorithm
by: Liu, Yiwen, et al.
Published: (2024)
by: Liu, Yiwen, et al.
Published: (2024)
MPDiT: Multi-Patch Global-to-Local Transformer Architecture For Efficient Flow Matching and Diffusion Model
by: Dao, Quan, et al.
Published: (2026)
by: Dao, Quan, et al.
Published: (2026)
StyleMotif: Multi-Modal Motion Stylization using Style-Content Cross Fusion
by: Guo, Ziyu, et al.
Published: (2025)
by: Guo, Ziyu, et al.
Published: (2025)
Double Helix of atomic displacements in Ferroelectric PbTiO$_3$
by: Hu, Yihao, et al.
Published: (2025)
by: Hu, Yihao, et al.
Published: (2025)
FCC: Fully Connected Correlation for One-Shot Segmentation
by: Moon, Seonghyeon, et al.
Published: (2024)
by: Moon, Seonghyeon, et al.
Published: (2024)
What Shapes a Creative Machine Mind? Comprehensively Benchmarking Creativity in Foundation Models
by: He, Zicong, et al.
Published: (2025)
by: He, Zicong, et al.
Published: (2025)
Micro-Defects Expose Macro-Fakes: Detecting AI-Generated Images via Local Distributional Shifts
by: Zhang, Boxuan, et al.
Published: (2026)
by: Zhang, Boxuan, et al.
Published: (2026)
AgentSelect: Benchmark for Narrative Query-to-Agent Recommendation
by: Shi, Yunxiao, et al.
Published: (2026)
by: Shi, Yunxiao, et al.
Published: (2026)
Human Resilience in the AI Era -- What Machines Can't Replace
by: Liu, Shaoshan, et al.
Published: (2025)
by: Liu, Shaoshan, et al.
Published: (2025)
Meaningless Tokens, Meaningful Gains: How Activation Shifts Enhance LLM Reasoning
by: Shi, Zeru, et al.
Published: (2025)
by: Shi, Zeru, et al.
Published: (2025)
Beyond Explicit Edges: Robust Reasoning over Noisy and Sparse Knowledge Graphs
by: Gao, Hang, et al.
Published: (2026)
by: Gao, Hang, et al.
Published: (2026)
Similar Items
-
Improving Visual Reasoning with Iterative Evidence Refinement
by: Shi, Zeru, et al.
Published: (2026) -
TrajDiffuse: A Conditional Diffusion Model for Environment-Aware Trajectory Prediction
by: Qingze, et al.
Published: (2024) -
ArchSeek: Retrieving Architectural Case Studies Using Vision-Language Models
by: Li, Danrui, et al.
Published: (2025) -
On the Equivalency, Substitutability, and Flexibility of Synthetic Data
by: Chang, Che-Jui, et al.
Published: (2024) -
Microscopic modeling of attention-based movement behaviors
by: Li, Danrui, et al.
Published: (2024)