AD-LLM: Benchmarking Large Language Models for Anomaly Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Tiankai, Nian, Yi, Li, Shawn, Xu, Ruiyao, Li, Yuangang, Li, Jiaqi, Xiao, Zhuo, Hu, Xiyang, Rossi, Ryan, Ding, Kaize, Hu, Xia, Zhao, Yue |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
NLP-ADBench: NLP Anomaly Detection Benchmark
by: Li, Yuangang, et al.
Published: (2024)
by: Li, Yuangang, et al.
Published: (2024)
No Attacker Needed: Unintentional Cross-User Contamination in Shared-State LLM Agents
by: Yang, Tiankai, et al.
Published: (2026)
by: Yang, Tiankai, et al.
Published: (2026)
Large Language Models for Anomaly and Out-of-Distribution Detection: A Survey
by: Xu, Ruiyao, et al.
Published: (2024)
by: Xu, Ruiyao, et al.
Published: (2024)
Cat-DPO: Category-Adaptive Safety Alignment
by: Yang, Tiankai, et al.
Published: (2026)
by: Yang, Tiankai, et al.
Published: (2026)
CoAct: Co-Active LLM Preference Learning with Human-AI Synergy
by: Xu, Ruiyao, et al.
Published: (2026)
by: Xu, Ruiyao, et al.
Published: (2026)
Mitigating Hallucinations in Large Language Models via Causal Reasoning
by: Li, Yuangang, et al.
Published: (2025)
by: Li, Yuangang, et al.
Published: (2025)
AD-AGENT: A Multi-agent Framework for End-to-end Anomaly Detection
by: Yang, Tiankai, et al.
Published: (2025)
by: Yang, Tiankai, et al.
Published: (2025)
Towards More Accurate US Presidential Election via Multi-step Reasoning with Large Language Models
by: Yu, Chenxiao, et al.
Published: (2024)
by: Yu, Chenxiao, et al.
Published: (2024)
GNN-as-Judge: Unleashing the Power of LLMs for Graph Learning with GNN Feedback
by: Xu, Ruiyao, et al.
Published: (2026)
by: Xu, Ruiyao, et al.
Published: (2026)
A Large-Scale Simulation on Large Language Models for Decision-Making in Political Science
by: Yu, Chenxiao, et al.
Published: (2024)
by: Yu, Chenxiao, et al.
Published: (2024)
StealthRank: LLM Ranking Manipulation via Stealthy Prompt Optimization
by: Tang, Yiming, et al.
Published: (2025)
by: Tang, Yiming, et al.
Published: (2025)
PyOD 2: A Python Library for Outlier Detection with LLM-powered Model Selection
by: Chen, Sihan, et al.
Published: (2024)
by: Chen, Sihan, et al.
Published: (2024)
Counterfactual Trace Auditing of LLM Agent Skills
by: Zhou, Xiaolin, et al.
Published: (2026)
by: Zhou, Xiaolin, et al.
Published: (2026)
FORTIS: Benchmarking Over-Privilege in Agent Skills
by: Li, Shawn, et al.
Published: (2026)
by: Li, Shawn, et al.
Published: (2026)
CMOOD: Concept-based Multi-label OOD Detection
by: Liu, Zhendong, et al.
Published: (2024)
by: Liu, Zhendong, et al.
Published: (2024)
Dynamics of Adversarial Attacks on Large Language Model-Based Search Engines
by: Hu, Xiyang
Published: (2025)
by: Hu, Xiyang
Published: (2025)
Secure On-Device Video OOD Detection Without Backpropagation
by: Li, Shawn, et al.
Published: (2025)
by: Li, Shawn, et al.
Published: (2025)
Language Shapes Mental Health Evaluations in Large Language Models
by: Xu, Jiayi, et al.
Published: (2026)
by: Xu, Jiayi, et al.
Published: (2026)
Unifying Unsupervised Graph-Level Anomaly Detection and Out-of-Distribution Detection: A Benchmark
by: Wang, Yili, et al.
Published: (2024)
by: Wang, Yili, et al.
Published: (2024)
DPU: Dynamic Prototype Updating for Multimodal Out-of-Distribution Detection
by: Li, Shawn, et al.
Published: (2024)
by: Li, Shawn, et al.
Published: (2024)
Empowering Large Language Models for Textual Data Augmentation
by: Li, Yichuan, et al.
Published: (2024)
by: Li, Yichuan, et al.
Published: (2024)
AnomalyLLM: Few-shot Anomaly Edge Detection for Dynamic Graphs using Large Language Models
by: Liu, Shuo, et al.
Published: (2024)
by: Liu, Shuo, et al.
Published: (2024)
CXR-AD: Component X-ray Image Dataset for Industrial Anomaly Detection
by: Bai, Haoyu, et al.
Published: (2025)
by: Bai, Haoyu, et al.
Published: (2025)
Multitask Active Learning for Graph Anomaly Detection
by: Chang, Wenjing, et al.
Published: (2024)
by: Chang, Wenjing, et al.
Published: (2024)
"Someone Hid It": Query-Agnostic Black-Box Attacks on LLM-Based Retrieval
by: Li, Jiate, et al.
Published: (2026)
by: Li, Jiate, et al.
Published: (2026)
Let's Ask GNN: Empowering Large Language Model for Graph In-Context Learning
by: Hu, Zhengyu, et al.
Published: (2024)
by: Hu, Zhengyu, et al.
Published: (2024)
MetaGAD: Meta Representation Adaptation for Few-Shot Graph Anomaly Detection
by: Xu, Xiongxiao, et al.
Published: (2023)
by: Xu, Xiongxiao, et al.
Published: (2023)
GeoChemAD: Benchmarking Unsupervised Geochemical Anomaly Detection for Mineral Exploration
by: Ding, Yihao, et al.
Published: (2026)
by: Ding, Yihao, et al.
Published: (2026)
A Survey on Model Extraction Attacks and Defenses for Large Language Models
by: Zhao, Kaixiang, et al.
Published: (2025)
by: Zhao, Kaixiang, et al.
Published: (2025)
Beyond Output Correctness: Benchmarking and Evaluating Large Language Model Reasoning in Coding Tasks
by: Li, Yuangang, et al.
Published: (2026)
by: Li, Yuangang, et al.
Published: (2026)
Graph Synthetic Out-of-Distribution Exposure with Large Language Models
by: Xu, Haoyan, et al.
Published: (2025)
by: Xu, Haoyan, et al.
Published: (2025)
The Autonomy Tax: Defense Training Breaks LLM Agents
by: Li, Shawn, et al.
Published: (2026)
by: Li, Shawn, et al.
Published: (2026)
MMR-AD: A Large-Scale Multimodal Dataset for Benchmarking General Anomaly Detection with Multimodal Large Language Models
by: Yao, Xincheng, et al.
Published: (2026)
by: Yao, Xincheng, et al.
Published: (2026)
Value-Action Alignment in Large Language Models under Privacy-Prosocial Conflict
by: Chen, Guanyu, et al.
Published: (2026)
by: Chen, Guanyu, et al.
Published: (2026)
AD-Bench: A Real-World, Trajectory-Aware Advertising Analytics Benchmark for LLM Agents
by: Hu, Lingxiang, et al.
Published: (2026)
by: Hu, Lingxiang, et al.
Published: (2026)
JailDAM: Jailbreak Detection with Adaptive Memory for Vision-Language Model
by: Nian, Yi, et al.
Published: (2025)
by: Nian, Yi, et al.
Published: (2025)
AD4AD: Benchmarking Visual Anomaly Detection Models for Safer Autonomous Driving
by: Genilotti, Fabrizio, et al.
Published: (2026)
by: Genilotti, Fabrizio, et al.
Published: (2026)
Treble Counterfactual VLMs: A Causal Approach to Hallucination
by: Li, Shawn, et al.
Published: (2025)
by: Li, Shawn, et al.
Published: (2025)
HopRank: Self-Supervised LLM Preference-Tuning on Graphs for Few-Shot Node Classification
by: Wang, Ziqing, et al.
Published: (2026)
by: Wang, Ziqing, et al.
Published: (2026)
TrajAD: Trajectory Anomaly Detection for Trustworthy LLM Agents
by: Liu, Yibing, et al.
Published: (2026)
by: Liu, Yibing, et al.
Published: (2026)
Similar Items
-
NLP-ADBench: NLP Anomaly Detection Benchmark
by: Li, Yuangang, et al.
Published: (2024) -
No Attacker Needed: Unintentional Cross-User Contamination in Shared-State LLM Agents
by: Yang, Tiankai, et al.
Published: (2026) -
Large Language Models for Anomaly and Out-of-Distribution Detection: A Survey
by: Xu, Ruiyao, et al.
Published: (2024) -
Cat-DPO: Category-Adaptive Safety Alignment
by: Yang, Tiankai, et al.
Published: (2026) -
CoAct: Co-Active LLM Preference Learning with Human-AI Synergy
by: Xu, Ruiyao, et al.
Published: (2026)