Multi-Level Anomaly Detection on Time-Varying Graph Data

Bridges, Robert A.; Collins, John; Ferragut, Erik M.; Laska, Jason; Sullivan, Blair D.

Computer Science > Social and Information Networks

arXiv:1410.4355 (cs)

[Submitted on 16 Oct 2014 (v1), last revised 20 Apr 2015 (this version, v4)]

Title:Multi-Level Anomaly Detection on Time-Varying Graph Data

Authors:Robert A. Bridges, John Collins, Erik M. Ferragut, Jason Laska, Blair D. Sullivan

View PDF

Abstract:This work presents a novel modeling and analysis framework for graph sequences which addresses the challenge of detecting and contextualizing anomalies in labelled, streaming graph data. We introduce a generalization of the BTER model of Seshadhri et al. by adding flexibility to community structure, and use this model to perform multi-scale graph anomaly detection. Specifically, probability models describing coarse subgraphs are built by aggregating probabilities at finer levels, and these closely related hierarchical models simultaneously detect deviations from expectation. This technique provides insight into a graph's structure and internal context that may shed light on a detected event. Additionally, this multi-scale analysis facilitates intuitive visualizations by allowing users to narrow focus from an anomalous graph to particular subgraphs or nodes causing the anomaly.
For evaluation, two hierarchical anomaly detectors are tested against a baseline Gaussian method on a series of sampled graphs. We demonstrate that our graph statistics-based approach outperforms both a distribution-based detector and the baseline in a labeled setting with community structure, and it accurately detects anomalies in synthetic and real-world datasets at the node, subgraph, and graph levels. To illustrate the accessibility of information made possible via this technique, the anomaly detector and an associated interactive visualization tool are tested on NCAA football data, where teams and conferences that moved within the league are identified with perfect recall, and precision greater than 0.786.

Comments:	8 pages. Updated paper to address reviewer comments
Subjects:	Social and Information Networks (cs.SI); Machine Learning (cs.LG); Machine Learning (stat.ML)
Cite as:	arXiv:1410.4355 [cs.SI]
	(or arXiv:1410.4355v4 [cs.SI] for this version)
	https://doi.org/10.48550/arXiv.1410.4355

Submission history

From: Erik Ferragut [view email]
[v1] Thu, 16 Oct 2014 09:57:20 UTC (737 KB)
[v2] Fri, 17 Oct 2014 19:08:37 UTC (737 KB)
[v3] Fri, 17 Apr 2015 16:58:08 UTC (695 KB)
[v4] Mon, 20 Apr 2015 11:55:53 UTC (695 KB)

Computer Science > Social and Information Networks

Title:Multi-Level Anomaly Detection on Time-Varying Graph Data

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Social and Information Networks

Title:Multi-Level Anomaly Detection on Time-Varying Graph Data

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators