科言猫
学术研究的AI总结
首页
文献互助
订阅
我的收藏
科研工具
选题分析
论文总结
专利管理
未登录
返回
期刊详情
P
Parallel Computing
IF
2.1
论文数
19
被引数
相关解读
0
订阅
期刊论文
19
相关解读
0
期刊论文
19
发表时间
发表时间
IF
被引数
Evaluating SYCL as a unified programming model for heterogeneous systems☆
评估SYCL作为异构系统统一编程模型的效果
Parallel Computing
IF
2.1
2026-06-01
0
PRE
AI
Marowka, Ami
分享
收藏
LSAF: A load-balancing SpGEMM acceleration framework with dynamic package and static partition for multi-core systolic arrays
LSAF:一种面向多核脉动阵列的负载均衡稀疏矩阵-稀疏矩阵乘法加速框架,包含动态打包和静态划分
Parallel Computing
IF
2.1
2026-02-01
0
PRE
AI
Cao, Yongxiang; Jiang, Hongxu; Zhao, Guocheng; Shi, Dongcheng; Zhang, Runhua; Wang, Wei
分享
收藏
Exploring metrics for analyzing dynamic behavior in MPI programs via a coupled-oscillator model
探索通过耦合振荡器模型分析MPI程序动态行为的指标
Parallel Computing
IF
2.1
2026-02-01
0
OA
AI
Afzal, Ayesha; Hager, Georg; Wellein, Gerhard
分享
收藏
Microarchitectural comparison, in-core modeling, and memory hierarchy analysis of state-of-the-art CPUs: Grace, Sapphire Rapids, and Genoa
先进CPU的微架构比较、核心建模及内存层次分析:Grace、Sapphire Rapids和Genoa
Parallel Computing
IF
2.1
2026-02-01
0
OA
AI
Laukemann, Jan; Hager, Georg; Wellein, Gerhard
分享
收藏
HRPF: A parallel programming framework for recursive algorithms on heterogeneous CPU-GPU systems
HRPF:异构CPU-GPU系统上递归算法的并行编程框架
Parallel Computing
IF
2.1
2026-02-01
0
PRE
AI
Wang, Yizhuo; Liu, Bowen; Shao, Senhao; Gao, Jianhua; Ji, Weixing; Xing, Hongbo
分享
收藏
A case study in hardware specialization for Monte Carlo cross-section lookup☆
针对蒙特卡洛截面查找的硬件专用化案例研究☆
Parallel Computing
IF
2.1
2026-01-01
0
PRE
AI
Yoshii, Kazutomo; Tramm, John R.; Allen, Bryce; Ueno, Tomohiro; Sano, Kentaro; Siegel, Andrew; Beckman, Pete
分享
收藏
Towards analysis and refinement of auto-tuning spaces
面向自动调优空间的分析与精化
Parallel Computing
IF
2.1
2026-01-01
0
OA
AI
Filipovic, Jiri; Gevorgyan, Suren Harutyunyan; Cesar, Eduardo; Sikora, Anna
分享
收藏
PROAD: Boosting Caffe Training via improving LevelDB I/O performance with Parallel Read, Out-of-Order Optimization, and Adaptive Design
PROAD:通过并行读取、乱序优化和自适应设计提升LevelDB I/O性能,从而加速Caffe训练
Parallel Computing
IF
2.1
2025-12-01
0
PRE
AI
Pan, Yubiao; Tian, Ailing; Zhang, Huizhen
分享
收藏
Analysis of the impact of NUMA node configuration on the performance of offloading computations to GPUs
NUMA节点配置对卸载计算至GPU性能影响的分析
Parallel Computing
IF
2.1
2025-12-01
0
PRE
AI
Malkovsky, Sergey; Sorokin, Aleksei; Korolev, Sergey
分享
收藏
Butterfly factorization for vision transformers on multi-IPU systems
蝴蝶分解在多IPU系统上的视觉Transformer应用
Parallel Computing
IF
2.1
2025-12-01
0
OA
AI
Shekofteh, S. -Kazem; Bogacz, Daniel; Alles, Christian; Froning, Holger
分享
收藏
Benchmark of classical disk array and software-defined storage on near-identical hardware
经典磁盘阵列与软件定义存储在近乎相同硬件上的基准测试
Parallel Computing
IF
2.1
2025-12-01
0
PRE
AI
Vondra, Tomas; Sebek, David
分享
收藏
Machine learning-driven fault-tolerant core mapping in Network-on-Chip architectures for advanced computing networks
机器学习驱动的NoC架构中先进计算网络的容错核心映射
Parallel Computing
IF
2.1
2025-12-01
1
PRE
AI
Yadav, Challa Muralikrishna; Reddy, B. Naresh Kumar
分享
收藏
Cache partitioning for sparse matrix-vector multiplication on the A64FX
A64FX上的稀疏矩阵-向量乘法的缓存分区
Parallel Computing
IF
2.1
2025-12-01
0
OA
AI
Breiter, Sergej; Trotter, James D.; Fuerlinger, Karl
分享
收藏
A sleek lock-free hash map in an ERA of safe memory reclamation methods
Parallel Computing
IF
2.1
2025-11-01
0
OA
AI
Moreno, Pedro; Areias, Miguel; Rocha, Ricardo
分享
收藏
LSHDP: Locally sharded heterogeneous data parallel for distributed deep learning
LSHDP:分布式深度学习中的本地分片异构数据并行
Parallel Computing
IF
2.1
2025-11-01
0
PRE
AI
Mirzaei, Motahhare; Ashtiani, Mehrdad; Pirhadi, Mohammad Javad; Eetemadi, Sauleh
分享
收藏
Detecting chaotic regions of recurrent equations in parallel environments
在并行环境中检测递归方程的混沌区域
Parallel Computing
IF
2.1
2025-11-01
0
OA
AI
Margaris, Athanasios; Souravlas, Stavros
分享
收藏
A dependency-aware task offloading in IoT-based edge computing system using an optimized deep learning approach
Parallel Computing
IF
2.1
2025-10-01
0
PRE
AI
Reddy, Shiva Shankar; Nrusimhadri, Silpa; Mahesh, Gadiraju; Rao, Veeranki Venkata Rama Maheswara
分享
收藏
GPU/CUDA-Accelerated gradient growth optimizer for efficient complex numerical global optimization
GPU/CUDA加速的梯度增长优化器,用于高效复杂数值全局优化
Parallel Computing
IF
2.1
2025-10-01
0
PRE
AI
Zhang, Qingke; Chen, Wenliang; Pang, Shuzhao; Tao, Sichen; Li, Conglin; Yin, Xin
分享
收藏
ALBBA: An efficient ALgebraic Bypass BFS Algorithm on long vector architectures
ALBBA:一种适用于长向量架构的高效代数绕过BFS算法
PARALLEL COMPUTING
IF
2.1
2025-09-01
0
OA
AI
Niu, Yuyao; Casas, Marc
分享
收藏