arrow
返回

Efficient high-precision integer multiplication on the GPU

delete2022-03-20
delete1
delete
OA
AI
A
Adrián Perez Diéguez *
M
Margarita Amor
R
Ramón Doallo
A
Akira Nukada
S
Satoshi Matsuoka
DOI:10.1177/10943420221077964delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
The multiplication of large integers, which has many applications in computer science, is an operation that can be expressed as a polynomial multiplication followed by a carry normalization. This work develops two approaches for efficient polynomial multiplication: one approach is based on tiling the classical convolution algorithm, but taking advantage of new CUDA architectures, a novelty approach to compute the multiplication using integers without accuracy lossless; the other one is based on the Strassen algorithm, an algorithm that multiplies large polynomials using the FFT operation, but adapting the fastest FFT libraries for current GPUs and working on the complex field. Previous studies reported that the Strassen algorithm is an effective implementation for large enough integers on GPUs. Additionally, most previous studies do not examine the implementation of the carry normalization, but this work describes a parallel implementation for this operation. Our results show the efficiency of our approaches for short, medium, and large sizes.
Keyword:
large integers
multiplication
FFT
GPU
CUDA

期刊

International Journal of High Performance Computing Applications 封面图
International Journal of High Performance Computing Applications
IF:
2.5
论文数:
1.1K
被引数:
1.3K

机构

U
Universidade da Coruna
学者数:
6.6K
论文数: 5.7K
被引数: 11
U
University of Tsukuba
学者数:
1.8W
论文数: 1.5W
被引数: 1.7W
R
riken
学者数:
2.2W
论文数: 1.9W
被引数: 24
学者 查看更多机构