返回
Examining parallelization in kernel regression
DOI:10.1007/s00500-023-09285-4.png)
摘要
En 中文
For a few decades, parallelization in statistical computing has been an increasing trend, and researchers have put significant effort into converting or adjusting known statistical methods and algorithms in parallel. The main reasons for the transition to parallel processes are the rapid growth in the size and the volume of data and the accelerated hardware developments. Divide and (re)combine (DnR) is one of the parallelization methods that allows the existing data or method to be implemented by dividing it into smaller pieces. It is possible to use the DnR method in most regression methods to reveal the relationship between the data. Although several libraries have been created in existing programming languages for many regression methods, such an approach is not yet used for kernel regression. However, it should be kept in mind that the kernel regression calculation method takes a relatively long time. For this reason, parallelization would be a handy strategy to decrease the calculation time in kernel regression. In this study, we aim to demonstrate how time efficiency is achieved using DnR methods for kernel regression with the help of several parallelization strategies in R. The results indicate that the computation time can be reduced proportionally with a trade-off between time and accuracy.
Keyword:
Kernel regression
Parallel computing
Foreach
parLapply
Multi-core systems
R
期刊
IF:
2.5
论文数:
1.0W
被引数:
2.1W
机构
引用论文
The Role of Trust in the Strategic Management Process: A Case Study of Finnish Grocery Retail Company Kesko Ltd信任在战略管理过程中的作用:芬兰食品零售公司Kesko Ltd的案例研究
A joint introduction to Gaussian Processes and Relevance Vector Machines with connections to Kalman filtering and other kernel smoothers
INFORMATION FUSION
IF15.5
Dysfunctions of decision‐making and cognitive control as transdiagnostic mechanisms of mental disorders: advances, gaps, and needs in current research作为精神障碍的跨诊断机制的决策和认知控制功能障碍: 当前研究的进展,差距和需求

