MaNIACS : Approximate Mining of Frequent Subgraph Patterns through Sampling

Preti, Giulia; De Francisci Morales, Gianmarco; Riondato, Matteo

doi:10.1145/3587254

Citation Details

MaNIACS : Approximate Mining of Frequent Subgraph Patterns through Sampling

We present MaNIACS , a sampling-based randomized algorithm for computing high-quality approximations of the collection of the subgraph patterns that are frequent in a single, large, vertex-labeled graph, according to the Minimum Node Image-based (MNI) frequency measure. The output of MaNIACS comes with strong probabilistic guarantees, obtained by using the empirical Vapnik–Chervonenkis (VC) dimension, a key concept from statistical learning theory, together with strong probabilistic tail bounds on the difference between the frequency of a pattern in the sample and its exact frequency. MaNIACS leverages properties of the MNI-frequency to aggressively prune the pattern search space, and thus to reduce the time spent in exploring subspaces that contain no frequent patterns. In turn, this pruning leads to better bounds to the maximum frequency estimation error, which leads to increased pruning, resulting in a beneficial feedback effect. The results of our experimental evaluation of MaNIACS on real graphs show that it returns high-quality collections of frequent patterns in large graphs up to two orders of magnitude faster than the exact algorithm. more »

Award ID(s):: 2006765

PAR ID:: 10464552

Author(s) / Creator(s):: Preti, Giulia; De Francisci Morales, Gianmarco; Riondato, Matteo

Date Published:: 2023-06-30

Journal Name:: ACM Transactions on Intelligent Systems and Technology

Volume:: 14

Issue:: 3

ISSN:: 2157-6904

Page Range / eLocation ID:: 1 to 29

Format(s):: Medium: X

Sponsoring Org:: National Science Foundation

Free Publicly Accessible Full Text
Accepted Manuscript1.0
Journal Article:
https://doi.org/10.1145/3587254

More Like this