NASCENT2: Generic Near-Storage Sort Accelerator for Data Analytics on SmartSSD-Reference-Cited by-同舟云学术

NASCENT2: Generic Near-Storage Sort Accelerator for Data Analytics on SmartSSD

Published:2022-01-28 Issue:2 Volume:15 Page:1-29
ISSN:1936-7406
Container-title:ACM Transactions on Reconfigurable Technology and Systems
language:en
Short-container-title:ACM Trans. Reconfigurable Technol. Syst.

Author:

Salamat Sahand¹^ORCID,Zhang Hui²,Ki Yang Seok²,Rosing Tajana¹

Affiliation:

1. UC San Diego, La Jolla, CA

2. Samsung Semiconductor Inc., San Jose, CA

Abstract

As the size of data generated every day grows dramatically, the computational bottleneck of computer systems has shifted toward storage devices. The interface between the storage and the computational platforms has become the main limitation due to its limited bandwidth, which does not scale when the number of storage devices increases. Interconnect networks do not provide simultaneous access to all storage devices and thus limit the performance of the system when executing independent operations on different storage devices. Offloading the computations to the storage devices eliminates the burden of data transfer from the interconnects. Near-storage computing offloads a portion of computations to the storage devices to accelerate big data applications. In this article, we propose a generic near-storage sort accelerator for data analytics, NASCENT2, which utilizes Samsung SmartSSD, an NVMe flash drive with an on-board FPGA chip that processes data in situ. NASCENT2 consists of dictionary decoder, sort, and shuffle FPGA-based accelerators to support sorting database tables based on a key column with any arbitrary data type. It exploits data partitioning applied by data processing management systems, such as SparkSQL, to breakdown the sort operations on colossal tables to multiple sort operations on smaller tables. NASCENT2 generic sort provides 2 × speedup and 15.2 × energy efficiency improvement as compared to the CPU baseline. It moreover considers the specifications of the SmartSSD (e.g., the FPGA resources, interconnect network, and solid-state drive bandwidth) to increase the scalability of computer systems as the number of storage devices increases. With 12 SmartSSDs, NASCENT2 is 9.9× (137.2 ×) faster and 7.3 × (119.2 ×) more energy efficient in sorting the largest tables of TPCC and TPCH benchmarks than the FPGA (CPU) baseline.

Funder

CRISP, one of six centers in JUMP, an SRC program sponsored by DARPA

SRC Global Research Collaboration (GRC) grant

NSF

Publisher

Association for Computing Machinery (ACM)

Subject

General Computer Science

Link

https://dl.acm.org/doi/pdf/10.1145/3472769

Reference66 articles.

1. Raghu Ramakrishnan, Johannes Gehrke, and Johannes Gehrke. 2003. Database Management Systems. Vol. 3. McGraw-Hill, New York, NY.

2. A Survey: Classification of Big Data

3. Hung-Wei Tseng, Yang Liu, Mark Gahagan, Jing Li, Yanqin Jing, and Steven J. Swanson. 2015. Gullfoss: Accelerating and Simplifying Data Movement Among Heterogeneous Computing and Storage Resources. Department of Computer Science and Engineering, University of California.

4. Analyzing and Modeling In-Storage Computing Workloads On EISC — An FPGA-Based System-Level Emulation Platform

5. Summarizer

Cited by 15 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Accelerating Ransomware Defenses with Computational Storage Drive-Based API Call Sequence Classification;Proceedings of the 17th Cyber Security Experimentation and Test Workshop;2024-08-13

2. Examining the Standardization of Solutions for the Integration of Implementation of Warehouse Management Systems;Canadian Journal of Business and Information Studies;2024-07-04

3. Empowering Data Centers with Computational Storage Drive-Based Deep Learning Inference Functionality to Combat Ransomware;2024 54th Annual IEEE/IFIP International Conference on Dependable Systems and Networks - Supplemental Volume (DSN-S);2024-06-24

4. FINESSD: Near-Storage Feature Selection with Mutual Information for Resource-Limited FPGAs;2024 IEEE 32nd Annual International Symposium on Field-Programmable Custom Computing Machines (FCCM);2024-05-05

5. Adaptive DRAM Cache Division for Computational Solid-state Drives;2024 Design, Automation & Test in Europe Conference & Exhibition (DATE);2024-03-25