Improving execution efficiency of just-in-time compilation based query processing on GPUs-Reference-Cited by-同舟云学术

Improving execution efficiency of just-in-time compilation based query processing on GPUs

Published:2020-10 Issue:2 Volume:14 Page:202-214
ISSN:2150-8097
Container-title:Proceedings of the VLDB Endowment
language:en
Short-container-title:Proc. VLDB Endow.

Author:

Paul Johns¹,He Bingsheng¹,Lu Shengliang¹,Lau Chiew Tong²

Affiliation:

1. National University of Singapore

2. Nanyang Technological University, Singapore

Abstract

In recent years, we have witnessed significant efforts to improve the performance of Online Analytical Processing (OLAP) on graphics processing units (GPUs). Most existing studies have focused on improving memory efficiency since memory stalls can play an essential role in query processing performance on GPUs. Motivated by the recent rise of just-in-time (JIT) compilation in query processing, we investigate whether and how we can further improve query processing performance on GPU. Specifically, we study the execution of state-of-the-art JIT compile-based query processing systems. We find that thanks to advanced techniques such as database compression and JIT compilation, memory stalls are no longer the most significant bottleneck. Instead, current JIT compile-based query processing encounters severe under-utilization of GPU hardware due to divergent execution and degraded parallelism arising from resource contention. To address these issues, we propose a JIT compile-based query engine named Pyper to improve GPU utilization during query execution. Specifically, Pyper has two new operators, Shuffle and Segment , for query plan transformation, which can be plugged into a physical query plan in order to reduce divergent execution and resolve resource contention, respectively. To determine the insertion points for these two operators, we present an analytical model that helps insert Shuffle and Segment operators into a query plan in a cost-based manner. Our experiments show that 1) the analytical analysis of divergent execution and resource contention helps to improve the accuracy of the cost model, 2) Pyper significantly outperforms other GPU query engines on TPC-H and SSB queries.

Publisher

VLDB Endowment

Subject

General Earth and Planetary Sciences,Water Science and Technology,Geography, Planning and Development

Link

https://dl.acm.org/doi/pdf/10.14778/3425879.3425890

Cited by 16 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. How Does Software Prefetching Work on GPU Query Processing?;Proceedings of the 20th International Workshop on Data Management on New Hardware;2024-06-09

2. Give a JIT on GPUs: NVRTC for Code-Generating Database Systems;2024 IEEE 40th International Conference on Data Engineering Workshops (ICDEW);2024-05-13

3. Accelerating Merkle Patricia Trie with GPU;Proceedings of the VLDB Endowment;2024-04

4. BladeDISC: Optimizing Dynamic Shape Machine Learning Workloads via Compiler Approach;Proceedings of the ACM on Management of Data;2023-11-13

5. GPU Database Systems Characterization and Optimization;Proceedings of the VLDB Endowment;2023-11