Toward Reasoning on the Boundary: A Mixup-based Approach for Graph Anomaly Detection

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Kim, Hwan, Kim, Junghoon, Lim, Sungsu
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866912940594036736
author Kim, Hwan
Kim, Junghoon
Lim, Sungsu
author_facet Kim, Hwan
Kim, Junghoon
Lim, Sungsu
contents While GNN-based detection methods excel at identifying overt outliers, they often struggle with boundary anomalies -- subtly camouflaged nodes that are difficult to distinguish from normal instances. This limitation highlights a fundamental gap in the reasoning capabilities of existing methods. We attribute this issue to the reliance of standard Graph Contrastive Learning (GCL) on easy negatives, which fosters the learning of simplistic decision boundaries. To address this issue, we propose ANOMIX, a framework that synthesizes informative hard negatives by linearly interpolating representations of normal and abnormal subgraphs. This graph mixup strategy intentionally populates the decision boundary with hard-to-detect samples. Through targeted experimental analysis, we demonstrate that ANOMIX successfully separates these boundary anomalies where state-of-the-art baselines fail, as shown by a clear distinction in the score distributions for these challenging cases. These findings suggest that synthesizing hard negatives via mixup is a potent strategy for refining GNN representation space, which in turn enhances its reasoning capacity for more robust and reliable graph anomaly detection. Code is available at https://github.com/missinghwan/ANOMIX.
format Preprint
id arxiv_https___arxiv_org_abs_2410_20310
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Toward Reasoning on the Boundary: A Mixup-based Approach for Graph Anomaly Detection
Kim, Hwan
Kim, Junghoon
Lim, Sungsu
Machine Learning
Artificial Intelligence
While GNN-based detection methods excel at identifying overt outliers, they often struggle with boundary anomalies -- subtly camouflaged nodes that are difficult to distinguish from normal instances. This limitation highlights a fundamental gap in the reasoning capabilities of existing methods. We attribute this issue to the reliance of standard Graph Contrastive Learning (GCL) on easy negatives, which fosters the learning of simplistic decision boundaries. To address this issue, we propose ANOMIX, a framework that synthesizes informative hard negatives by linearly interpolating representations of normal and abnormal subgraphs. This graph mixup strategy intentionally populates the decision boundary with hard-to-detect samples. Through targeted experimental analysis, we demonstrate that ANOMIX successfully separates these boundary anomalies where state-of-the-art baselines fail, as shown by a clear distinction in the score distributions for these challenging cases. These findings suggest that synthesizing hard negatives via mixup is a potent strategy for refining GNN representation space, which in turn enhances its reasoning capacity for more robust and reliable graph anomaly detection. Code is available at https://github.com/missinghwan/ANOMIX.
title Toward Reasoning on the Boundary: A Mixup-based Approach for Graph Anomaly Detection
topic Machine Learning
Artificial Intelligence
url https://arxiv.org/abs/2410.20310