arrow
Return

Fine-Grained Multi-Query Stream Processing on Integrated Architectures

delete2021-09-01
delete13
PRE
AI
张峰 (Feng Zhang)
C
Chenyang Zhang
杨琳 cover
杨琳 (Lin Yang)
S
Shuhao Zhang
B
Bingsheng He
卢卫 cover
卢卫 (Wei Lü)
X
Xiaoyong Du *
DOI:10.1109/TPDS.2021.3066407delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Exploring the sharing opportunities among multiple stream queries is crucial for high-performance stream processing. Modern stream processing necessitates accelerating multiple queries by utilizing heterogeneous coprocessors, such as GPUs, and this has shown to be an effective method. Emerging CPU-GPU integrated architectures 6integrate CPU and GPU on the same chip and eliminate PCI-e bandwidth bottleneck. Such a novel architecture provides new opportunities for improving multi-query performance in stream processing but has not been fully explored by existing systems. We introduce a stream processing engine, called FineStream, for efficient multi-query window-based stream processing on CPU-GPU integrated architectures. FineStream's key contribution is a novel fine-grained workload scheduling mechanism between CPU and GPU to take advantage of both architectures. Particularly, FineStream is able to efficiently handle multiple queries in both static and dynamic streams. Our experimental results show that 1) on integrated architectures, FineStream achieves an average 52 percent throughput improvement and 36 percent lower latency over the state-of-the-art stream processing engine; 2) compared to the coarse-grained strategy of applying different devices for multiple queries, FineStream achieves 32 percent throughput improvement; 3) compared to the stream processing engine on the discrete architecture, FineStream on the integrated architecture achieves 10.4x price-throughput ratio, 1.8x energy efficiency, and can enjoy lower latency benefits.
Keywords:
Computer architecture
Graphics processing units
Structured Query Language
Performance evaluation
Throughput
Engines
Bandwidth
Fine-grained
multi-query
stream processing
CPU
GPU
integrated architectures
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

IEEE Transactions on Parallel and Distributed Systems cover
IEEE Transactions on Parallel and Distributed Systems
IF:
6
Papers:
5.2K
Citations:
1.1W

Organization

R
Renmin University of China
Scholars:
8.1K
Papers: 7.7K
Citations: 1.1W
S
singapore university of technology & design
Scholars:
2.8K
Papers: 3.6K
Citations: 5
N
National University of Singapore
Scholars:
7.5W
Papers: 6.4W
Citations: 11.4W
researcher View more organizations