When Subgraph Isomorphism is Really Hard, and Why This Matters for Graph Databases-Reference-Cited by-同舟云学术

When Subgraph Isomorphism is Really Hard, and Why This Matters for Graph Databases

Published:2018-03-30 Issue: Volume:61 Page:723-759
ISSN:1076-9757
Container-title:Journal of Artificial Intelligence Research
language:
Short-container-title:jair

Author:

McCreesh Ciaran,Prosser Patrick,Solnon Christine,Trimble James

Abstract

The subgraph isomorphism problem involves deciding whether a copy of a pattern graph occurs inside a larger target graph. The non-induced version allows extra edges in the target, whilst the induced version does not. Although both variants are NP-complete, algorithms inspired by constraint programming can operate comfortably on many real-world problem instances with thousands of vertices. However, they cannot handle arbitrary instances of this size. We show how to generate "really hard" random instances for subgraph isomorphism problems, which are computationally challenging with a couple of hundred vertices in the target, and only twenty pattern vertices. For the non-induced version of the problem, these instances lie on a satisfiable / unsatisfiable phase transition, whose location we can predict; for the induced variant, much richer behaviour is observed, and constrainedness gives a better measure of difficulty than does proximity to a phase transition. These results have practical consequences: we explain why the widely researched "filter / verify" indexing technique used in graph databases is founded upon a misunderstanding of the empirical hardness of NP-complete problems, and cannot be beneficial when paired with any reasonable subgraph isomorphism algorithm.

Publisher

AI Access Foundation

Subject

Artificial Intelligence

Cited by 29 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. PathLAD+: Towards effective exact methods for subgraph isomorphism problem;Artificial Intelligence;2024-12

2. ArcMatch: high-performance subgraph matching for labeled graphs by exploiting edge domains;Data Mining and Knowledge Discovery;2024-08-07

3. l2Match: Optimization Techniques on Subgraph Matching Algorithm Using Label Pair, Neighboring Label Index, and Jump-Redo Method;2024 International Conference on Electronics, Information, and Communication (ICEIC);2024-01-28

4. Chemical Similarity and Substructure Searches;Reference Module in Life Sciences;2024

5. A Two-stage Approach for Tables Extraction in Invoices;2023 IEEE 35th International Conference on Tools with Artificial Intelligence (ICTAI);2023-11-06