arrow
Return

A High-Performance Scalable Shared-Memory SVD Processor Architecture Based on Jacobi Algorithm and Batcher's Sorting Network

delete2020-06-01
delete13
PRE
AI
S
Seyed Mohammad Reza Shahshahani
H
Hamid Reza Mahdiani *
DOI:10.1109/TCSI.2020.2973249delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Eigenvalue Decomposition (EVD) and Singular Value Decomposition (SVD) are two crucial transformations in many signal processing applications. The main drawback of these algorithms is their computationally intensive nature which prevents them to be efficiently exploited in high-performance, real-time and mobile applications. By extracting the inherent parallelism of the Jacobi SVD, a new parallel data distribution and access pattern for this algorithm is proposed first. Based on the proposed parallel data distribution, a novel shared-memory architecture is then proposed to support EVD/SVD computation in a high-performance and scalable manner. A new Multistage Interconnection Network based on Batcher's odd-even merge sorting network is developed and exploited in the architecture to preserve its performance and scalability by simultaneously connecting different numbers of processing elements to the system memory hierarchy in a parallel conflict-free manner. The proposed architecture can be configured to compute EVD/SVD of matrices of arbitrary size, with different numbers of processing elements achieving a linear speed-up. The synthesis results in a 90 nm technology show that the system with one, two, and four processing elements achieves a throughput of 1.81, 3.63, and 7.26 million EVD/SVD's per second, respectively with a frequency of 813 MHz for an 8x8 matrix.
Keywords:
Jacobian matrices
Computer architecture
Signal processing algorithms
Principal component analysis
Symmetric matrices
Parallel processing
Scalability
ASIC
Batcher's sorting network
brain-computer interface
independent component analysis
Jacobi EVD
SVD
motor imagery
multi-stage interconnection network
scalability
shared-memory architecture
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

IEEE Transactions on Circuits and Systems I-Regular Papers cover
IEEE Transactions on Circuits and Systems I-Regular Papers
IF:
5.2
Papers:
9.8K
Citations:
2.2W

Organization

S
Shahid Beheshti University
Scholars:
7.6K
Papers: 6.8K
Citations: 6.9K
Cited Papers

Cited Papers

Real-time ocular artifacts removal of EEG data using a hybrid ICA-ANC approach
err2017-01-01
err39
PREAI
errJafarifarmand, Aysa; Badamchizadeh, Mohammad-Ali; Khanmohammadi, Sohrab; Nazari, Mohammad Ali; Tazehkand, Behzad Mozaffari
errShare
errSave
Aircraft Aerodynamic Design
err
IF0
err2014-10-03
err0
PREAI
errAndrás Sóbester; Alexander I J Forrester
errShare
errSave
errShare
errSave
Photoelectrocatalytic activity PbO2 loaded highly oriented TiO2 nanotubes arrays
err2021-01-01
err0
PREAI
errZ.M. Alimirzaeva; A.B. Isaev; N.S. Shabanov; A.G. Magomedova; M.V. Kadiev; K. Kaviyarasu
errShare
errSave
Industrial Applications
err1984-01-01
err0
PREAI
errDavid S. Breslow
errShare
errSave
Deep Centres in ZnO
err2010-05-15
err0
PREAI
errA. Hoffmann; E. Malguth; B. K. Meyer
errShare
errSave
researcher View more