返回
Prediction-based GPU Sharing for Distributed Training
DOI:10.1016/j.future.2026.108413.png)
摘要
En 中文
• Formulate the inconsistent JCT problem using gSLA for the first time. • Design a new JCT increase prediction model and job scheduler for GPU sharing. • Achieve up to 47.3× better gSLA satisfaction and 50× lower gSLA excess ratio. • Improve JCT and GPU efficiency by ∼ 60% and ∼ 44% over existing methods. • Demonstrate TensorShare’s effectiveness in improving gSLA and JCT for unseen jobs.
Keyword:
Cloud computing
GPU sharing
Service level agreement
Performance prediction
GPU scheduling
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
F
IF:
0
论文数:
642
被引数:
0
机构
引用论文
Managing Performance Overhead of Virtual Machines in Cloud Computing: A Survey, State of the Art, and Future Directions
PROCEEDINGS OF THE IEEE
IF25.9

