返回
Exploring the parallel capabilities of GPU: Berlekamp-Massey algorithm case study
DOI:10.1007/s10586-019-02961-x.png)
摘要
En 中文
Graphics processors Unit (GPU) architectures are becoming increasingly programmable, offering the potential for dramatic speedups for a variety of general purpose applications compared to contemporary general- purpose processors (CPUs). However, there are several optimization techniques which are used to maximize the benefit of the GPU resources. This research exploits optimization techniques for CUDA enabled GPU architecture in order to achieve the best possible performance for Berlekamp-Massey Algorithm (BMA) as a case study. Berlekamp-Massey Algorithm (BMA) is one of the best solutions to find the shortest linear feedback shift register which is very important for several applications such as digital processing and cryptography. The experimental results show that the optimized BMA implementation is almost 160 x faster than non-bit CPU serial implementation, 7 x faster than bit serial implementation and 4 x faster than an initial parallel bit implementation.
Keyword:
Linear complexity
BerlekampMassey algorithm
Parallel computing
GPU
CUDA optimization techniques
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
C
IF:
4.1
论文数:
5.1K
被引数:
7.5K

