arrow
Return

Re-thinking Memory-Bound Limitations in CGRAs

delete2025-09-01
delete0
delete
OA
AI
X
X.P. Liu *
Z
Zhe Jiang
A
Anzhen Zhu
X
Xiaomeng Han
M
Mingsong Lyu
Q
Qingxu Deng
N
Nan Guan
DOI:10.1145/3760386delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
Coarse-Grained Reconfigurable Arrays (CGRAs) are specialized accelerators commonly employed to boost performance in workloads with iterative structures. Existing research typically focuses on compiler or architecture optimizations aimed at improving CGRA performance, energy efficiency, flexibility, and area utilization, under the idealistic assumption that kernels can access all data from Scratchpad Memory (SPM). However, certain complex workloads-particularly in fields like graph analytics, irregular database operations, and specialized forms of high-performance computing (e.g., unstructured mesh simulations)-exhibit irregular memory access patterns that hinder CGRA utilization, sometimes dropping below 1.5%, making the CGRA memory-bound. To address this challenge, we conduct a thorough analysis of the underlying causes of performance degradation, then propose a redesigned memory subsystem and refine the memory model. With both microarchitectural and theoretical optimization, our solution can effectively manage irregular memory accesses through CGRAspecific runahead execution mechanism and cache reconfiguration techniques. Our results demonstrate that we can achieve performance comparable to the original SPM-only system while requiring only 1.27% of the storage size. The runahead execution mechanism achieves an average 3.04x speedup (up to 6.91x), with cache reconfiguration technique providing an additional 6.02% improvement, significantly enhancing CGRA performance for irregular memory access patterns.
Keywords:
Coarse-Grained Reconfigurable Array (CGRA)
irregular memory access
memory subsystem
runahead execution
cache reconfiguration
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

ACM Transactions on Embedded Computing Systems cover
ACM Transactions on Embedded Computing Systems
IF:
2.6
Papers:
227
Citations:
2.3K

Organization

H
hong kong polytechnic university
Scholars:
3.0W
Papers: 4.1W
Citations: 921
C
City University of Hong Kong
Scholars:
2.3W
Papers: 3.0W
Citations: 6.1W
N
northeastern university - china
Scholars:
3.1W
Papers: 2.7W
Citations: 37
S
southeast university - china
Scholars:
5.3W
Papers: 4.9W
Citations: 57
researcher View more organizations