GraVF-M-Reference-Cited by-同舟云学术

GraVF-M

Published:2019-11-27 Issue:4 Volume:12 Page:1-28
ISSN:1936-7406
Container-title:ACM Transactions on Reconfigurable Technology and Systems
language:en
Short-container-title:ACM Trans. Reconfigurable Technol. Syst.

Author:

Engelhardt Nina¹,So Hayden K.-H.¹

Affiliation:

1. University of Hong Kong, Pokfulam Road, Hong Kong

Abstract

Due to the irregular nature of connections in most graph datasets, partitioning graph analysis algorithms across multiple computational nodes that do not share a common memory inevitably leads to large amounts of interconnect traffic. Previous research has shown that FPGAs can outcompete software-based graph processing in shared memory contexts, but it remains an open question if this advantage can be maintained in distributed systems. In this work, we present GraVF-M, a framework designed to ease the implementation of FPGA-based graph processing accelerators for multi-FPGA platforms with distributed memory. Based on a lightweight description of the algorithm kernel, the framework automatically generates optimized RTL code for the whole multi-FPGA design. We exploit an aspect of the programming model to present a familiar message-passing paradigm to the user, while under the hood implementing a more efficient architecture that can reduce the necessary inter-FPGA network traffic by a factor equal to the average degree of the input graph. A performance model based on a theoretical analysis of the factors influencing performance serves to evaluate the efficiency of our implementation. With a throughput of up to 5.8GTEPS (billions of traversed edges per second) on a 4-FPGA system, the designs generated by GraVF-M compare favorably to state-of-the-art frameworks from the literature and reach 94% of the projected performance limit of the system.

Funder

Research Grants Council of Hong Kong

Croucher Foundation

Publisher

Association for Computing Machinery (ACM)

Subject

General Computer Science

Link

https://dl.acm.org/doi/pdf/10.1145/3357596

Reference36 articles.

1. A framework for FPGA acceleration of large graph problems: Graphlet counting case study

2. D. Chakrabarti Y. Zhan and C. Faloutsos. 2004. R-MAT: A Recursive Model for Graph Mining. 442--446. DOI:https://doi.org/10.1137/1.9781611972740.43 D. Chakrabarti Y. Zhan and C. Faloutsos. 2004. R-MAT: A Recursive Model for Graph Mining. 442--446. DOI:https://doi.org/10.1137/1.9781611972740.43

Cited by 8 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. HashGrid: An optimized architecture for accelerating graph computing on FPGAs;Future Generation Computer Systems;2025-01

2. Optimising Graph Representation for Hardware Implementation of Graph Convolutional Networks for Event-Based Vision;Lecture Notes in Computer Science;2024

3. Distributed large-scale graph processing on FPGAs;Journal of Big Data;2023-06-04

4. Rethinking Design Paradigm of Graph Processing System with a CXL-like Memory Semantic Fabric;2023 IEEE/ACM 23rd International Symposium on Cluster, Cloud and Internet Computing (CCGrid);2023-05

5. ThunderGP: Resource-Efficient Graph Processing Framework on FPGAs with HLS;ACM Transactions on Reconfigurable Technology and Systems;2022-12-09