Engineering Motif Search for Large Motifs

Bulletin of the Society of Sea Water Science, Japan Pub Date : 2018-01-01 DOI:10.4230/LIPIcs.SEA.2018.28

P. Kaski, Juho Lauri, Suhas Thejaswi

{"title":"Engineering Motif Search for Large Motifs","authors":"P. Kaski, Juho Lauri, Suhas Thejaswi","doi":"10.4230/LIPIcs.SEA.2018.28","DOIUrl":null,"url":null,"abstract":"Given a vertex-colored graph H and a multiset M of colors as input, the graph motif problem asks us to decide whether H has a connected induced subgraph whose multiset of colors agrees with M. The graph motif problem is NP-complete but known to admit randomized algorithms based on constrained multilinear sieving over GF(2^b) that run in time O(2^kk^2m {M({2^b})}) and with a false-negative probability of at most k/2^{b-1} for a connected m-edge input and a motif of size k. On modern CPU microarchitectures such algorithms have practical edge-linear scalability to inputs with billions of edges for small motif sizes, as demonstrated by Bjorklund, Kaski, Kowalik, and Lauri [ALENEX'15]. This scalability to large graphs prompts the dual question whether it is possible to scale to large motif sizes. We present a vertex-localized variant of the constrained multilinear sieve that enables us to obtain, in time O(2^kk^2m{M({2^b})}) and for every vertex simultaneously, whether the vertex participates in at least one match with the motif, with a per-vertex probability of at most k/2^{b-1} for a false negative. Furthermore, the algorithm is easily vector-parallelizable for up to 2^k threads, and parallelizable for up to 2^kn threads, where n is the number of vertices in H. Here {M({2^b})} is the time complexity to multiply in GF(2^b). We demonstrate with an open-source implementation that our variant of constrained multilinear sieving can be engineered for vector-parallel microarchitectures to yield hardware utilization that is bound by the available memory bandwidth. Our main engineering contributions are (a) a version of the recurrence for tightly labeled arborescences that can be executed as a sequence of memory-and-arithmetic coalescent parallel workloads on multiple GPUs, and (b) a bit-sliced low-level implementation for arithmetic in characteristic 2 to support (a).","PeriodicalId":9448,"journal":{"name":"Bulletin of the Society of Sea Water Science, Japan","volume":"70 1","pages":"28:1-28:19"},"PeriodicalIF":0.0000,"publicationDate":"2018-01-01","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"5","resultStr":null,"platform":"Semanticscholar","paperid":null,"PeriodicalName":"Bulletin of the Society of Sea Water Science, Japan","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.4230/LIPIcs.SEA.2018.28","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"","JCRName":"","Score":null,"Total":0}

引用次数: 5

Abstract

Given a vertex-colored graph H and a multiset M of colors as input, the graph motif problem asks us to decide whether H has a connected induced subgraph whose multiset of colors agrees with M. The graph motif problem is NP-complete but known to admit randomized algorithms based on constrained multilinear sieving over GF(2^b) that run in time O(2^kk^2m {M({2^b})}) and with a false-negative probability of at most k/2^{b-1} for a connected m-edge input and a motif of size k. On modern CPU microarchitectures such algorithms have practical edge-linear scalability to inputs with billions of edges for small motif sizes, as demonstrated by Bjorklund, Kaski, Kowalik, and Lauri [ALENEX'15]. This scalability to large graphs prompts the dual question whether it is possible to scale to large motif sizes. We present a vertex-localized variant of the constrained multilinear sieve that enables us to obtain, in time O(2^kk^2m{M({2^b})}) and for every vertex simultaneously, whether the vertex participates in at least one match with the motif, with a per-vertex probability of at most k/2^{b-1} for a false negative. Furthermore, the algorithm is easily vector-parallelizable for up to 2^k threads, and parallelizable for up to 2^kn threads, where n is the number of vertices in H. Here {M({2^b})} is the time complexity to multiply in GF(2^b). We demonstrate with an open-source implementation that our variant of constrained multilinear sieving can be engineered for vector-parallel microarchitectures to yield hardware utilization that is bound by the available memory bandwidth. Our main engineering contributions are (a) a version of the recurrence for tightly labeled arborescences that can be executed as a sequence of memory-and-arithmetic coalescent parallel workloads on multiple GPUs, and (b) a bit-sliced low-level implementation for arithmetic in characteristic 2 to support (a).

查看原文本刊更多论文

大型Motif的工程Motif搜索

给定一个顶点彩色图H和一个多色集M作为输入，H图图案的问题让我们决定是否有一个连接诱导子图,其多重集的颜色同意M图图案的问题是np完全但已知承认基于多重线性约束的随机算法筛选/ GF (2 ^ b)中运行时间O (2 ^ kk ^ 2 M M {} ({2 ^ b}))和假阴性的可能性最多k / 2 ^ {b}连接m-edge输入和图案的大小k。在现代处理器微体系结构这类算法实用edge-linear可伸缩性Bjorklund, Kaski, Kowalik和Lauri [ALENEX'15]证明，小图案尺寸的输入具有数十亿条边。这种对大型图形的可扩展性提出了两个问题，即是否有可能缩放到大型主题尺寸。我们提出了约束多线性筛的顶点局部化变体，使我们能够在时间O(2^kk^2m{M({2^b})})中同时获得每个顶点是否与基序至少有一次匹配，对于假阴性，每个顶点的概率最多为k/2^{b-1}。此外，该算法可以很容易地对多达2^k个线程进行向量并行化，并且可以并行化多达2^kn个线程，其中n是h中的顶点数，这里{M({2^b})}是在GF(2^b)中相乘的时间复杂度。我们通过一个开源实现证明，我们的约束多线性筛分变体可以设计用于矢量并行微架构，以产生受可用内存带宽约束的硬件利用率。我们的主要工程贡献是(a)一个紧密标记的树形序列的递归版本，可以在多个gpu上作为内存和算术合并并行工作负载的序列执行，以及(b)一个特征2中的算术的位切片低级实现来支持(a)。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

求助全文

约1分钟内获得全文求助全文

来源期刊

Bulletin of the Society of Sea Water Science, Japan

自引率

0.00%

发文量