MemAudit: Post-hoc Auditing of Poisoned Agent Memory via Causal Attribution and Structural Anomaly Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Tan, Zhewen, Yao, Yilun, Jin, Huiyan, Yu, Wenhan, Wang, Guoan, Fan, Mengyuan, lu, liang, Liu, Feng, Zhang, Xiangzheng, Ma, Duohe, Yang, Tong, Sun, Lin |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TriPlay-RL: Tri-Role Self-Play Reinforcement Learning for LLM Safety Alignment
by: Tan, Zhewen, et al.
Published: (2026)
by: Tan, Zhewen, et al.
Published: (2026)
Harness-Bench: Measuring Harness Effects across Models in Realistic Agent Workflows
by: Yao, Yilun, et al.
Published: (2026)
by: Yao, Yilun, et al.
Published: (2026)
Efficient Switchable Safety Control in LLMs via Magic-Token-Guided Co-Training
by: Si, Jianfeng, et al.
Published: (2025)
by: Si, Jianfeng, et al.
Published: (2025)
Auditable Agents
by: Nian, Yi, et al.
Published: (2026)
by: Nian, Yi, et al.
Published: (2026)
MedSkillAudit: A Domain-Specific Audit Framework for Medical Research Agent Skills
by: Hou, Yingyong, et al.
Published: (2026)
by: Hou, Yingyong, et al.
Published: (2026)
FedDyMem: Efficient Federated Learning with Dynamic Memory and Memory-Reduce for Unsupervised Image Anomaly Detection
by: Chen, Silin, et al.
Published: (2025)
by: Chen, Silin, et al.
Published: (2025)
MemReranker: Reasoning-Aware Reranking for Agent Memory Retrieval
by: Li, Chunyu, et al.
Published: (2026)
by: Li, Chunyu, et al.
Published: (2026)
Causality-Driven Audits of Model Robustness
by: Drenkow, Nathan, et al.
Published: (2024)
by: Drenkow, Nathan, et al.
Published: (2024)
ARC: Active and Reflection-driven Context Management for Long-Horizon Information Seeking Agents
by: Yao, Yilun, et al.
Published: (2026)
by: Yao, Yilun, et al.
Published: (2026)
RepoAudit: An Autonomous LLM-Agent for Repository-Level Code Auditing
by: Guo, Jinyao, et al.
Published: (2025)
by: Guo, Jinyao, et al.
Published: (2025)
Effects of Attribution and Auditors' Response Readability of Critical Audit Matter on Audit Quality Perceptions and Valuation Judgements
by: Li Huang, et al.
Published: (2026)
by: Li Huang, et al.
Published: (2026)
MemCog: From Memory-as-Tool to Memory-as-Cognition in Conversational Agents
by: Li, Zihan, et al.
Published: (2026)
by: Li, Zihan, et al.
Published: (2026)
Optimally Auditing Adversarial Agents
by: Das, Sanmay, et al.
Published: (2026)
by: Das, Sanmay, et al.
Published: (2026)
Auditing Agent Harness Safety
by: Liu, Chengzhi, et al.
Published: (2026)
by: Liu, Chengzhi, et al.
Published: (2026)
Auditable Homomorphic-based Decentralized Collaborative AI with Attribute-based Differential Privacy
by: Yeh, Lo-Yao, et al.
Published: (2024)
by: Yeh, Lo-Yao, et al.
Published: (2024)
RW-Post: Auditable Evidence-Grounded Multimodal Fact-Checking in the Wild
by: Xu, Danni, et al.
Published: (2025)
by: Xu, Danni, et al.
Published: (2025)
RW-Post: Auditable Evidence-Grounded Multimodal Fact-Checking in the Wild
by: Xu, Danni, et al.
Published: (2026)
by: Xu, Danni, et al.
Published: (2026)
WhyLab: A Causal Audit Framework for Stable Agent Self-Improvement
by: Anonymous
Published: (2026)
by: Anonymous
Published: (2026)
Project Ariadne: A Structural Causal Framework for Auditing Faithfulness in LLM Agents
by: Khanzadeh, Sourena
Published: (2026)
by: Khanzadeh, Sourena
Published: (2026)
Mem-T: Densifying Rewards for Long-Horizon Memory Agents
by: Yue, Yanwei, et al.
Published: (2026)
by: Yue, Yanwei, et al.
Published: (2026)
SimpleMem: Efficient Lifelong Memory for LLM Agents
by: Liu, Jiaqi, et al.
Published: (2026)
by: Liu, Jiaqi, et al.
Published: (2026)
DAIQ: Auditing Demographic Attribute Inference from Question in LLMs
by: Panda, Srikant, et al.
Published: (2025)
by: Panda, Srikant, et al.
Published: (2025)
Auditing for Bias in Ad Delivery Using Inferred Demographic Attributes
by: Imana, Basileal, et al.
Published: (2024)
by: Imana, Basileal, et al.
Published: (2024)
Fairness Auditing with Multi-Agent Collaboration
by: de Vos, Martijn, et al.
Published: (2024)
by: de Vos, Martijn, et al.
Published: (2024)
CSR‐Linked Compensation Contract and Audit Pricing
by: Yiqing Tan
Published: (2025)
by: Yiqing Tan
Published: (2025)
EvoMem: Improving Multi-Agent Planning with Dual-Evolving Memory
by: Fan, Wenzhe, et al.
Published: (2025)
by: Fan, Wenzhe, et al.
Published: (2025)
ParamMem: Augmenting Language Agents with Parametric Reflective Memory
by: Yao, Tianjun, et al.
Published: (2026)
by: Yao, Tianjun, et al.
Published: (2026)
MEMSAD: Gradient-Coupled Anomaly Detection for Memory Poisoning in Retrieval-Augmented Agents
by: Gowda, Ishrith
Published: (2026)
by: Gowda, Ishrith
Published: (2026)
TourMart: A Parametric Audit Instrument for Commission Steering in LLM Travel Agents
by: Liu, Yao
Published: (2026)
by: Liu, Yao
Published: (2026)
MemTrace: Tracing and Attributing Errors in Large Language Model Memory Systems
by: Deng, Xinle, et al.
Published: (2026)
by: Deng, Xinle, et al.
Published: (2026)
Geometric Auditing of AI - Geometric Auditing of AI
by: Wyneken, B.
Published: (2025)
by: Wyneken, B.
Published: (2025)
Reporting Connectivity, Audit Quality, and Audit Fees
by: Meiting Lu, et al.
Published: (2025)
by: Meiting Lu, et al.
Published: (2025)
MAVEN: Multi-Agent Verification-Elaboration Network with In-Step Epistemic Auditing
by: Yao, Yinsheng, et al.
Published: (2026)
by: Yao, Yinsheng, et al.
Published: (2026)
Automated Benchmark Auditing for AI Agents and Large Language Models
by: Wang, Junlin, et al.
Published: (2026)
by: Wang, Junlin, et al.
Published: (2026)
TiMem: Temporal-Hierarchical Memory Consolidation for Long-Horizon Conversational Agents
by: Li, Kai, et al.
Published: (2026)
by: Li, Kai, et al.
Published: (2026)
GRAFT: Auditing Graph Neural Networks via Global Feature Attribution
by: Sahoo, Rishi Raj, et al.
Published: (2026)
by: Sahoo, Rishi Raj, et al.
Published: (2026)
Towards Trustworthy LLMs for Code: A Data-Centric Synergistic Auditing Framework
by: Wang, Chong, et al.
Published: (2024)
by: Wang, Chong, et al.
Published: (2024)
Auditing and Controlling AI Agent Actions in Spreadsheets
by: Sabouri, Sadra, et al.
Published: (2026)
by: Sabouri, Sadra, et al.
Published: (2026)
Counterfactual Trace Auditing of LLM Agent Skills
by: Zhou, Xiaolin, et al.
Published: (2026)
by: Zhou, Xiaolin, et al.
Published: (2026)
Terminators: Terms of Service Parsing and Auditing Agents
by: Mridul, Maruf Ahmed, et al.
Published: (2025)
by: Mridul, Maruf Ahmed, et al.
Published: (2025)
Similar Items
-
TriPlay-RL: Tri-Role Self-Play Reinforcement Learning for LLM Safety Alignment
by: Tan, Zhewen, et al.
Published: (2026) -
Harness-Bench: Measuring Harness Effects across Models in Realistic Agent Workflows
by: Yao, Yilun, et al.
Published: (2026) -
Efficient Switchable Safety Control in LLMs via Magic-Token-Guided Co-Training
by: Si, Jianfeng, et al.
Published: (2025) -
Auditable Agents
by: Nian, Yi, et al.
Published: (2026) -
MedSkillAudit: A Domain-Specific Audit Framework for Medical Research Agent Skills
by: Hou, Yingyong, et al.
Published: (2026)