跳到主要导航 跳到搜索 跳到主要内容

Prophet: Precise QoS prediction on non-preemptive accelerators to improve utilization in warehouse-scale computers

  • Quan Chen
  • , Hailong Yang
  • , Minyi Guo
  • , Ram Srivatsa Kannan
  • , Jason Mars
  • , Lingjia Tang
  • Shanghai Jiao Tong University
  • University of Michigan, Ann Arbor

科研成果: 书/报告/会议事项章节会议稿件同行评审

摘要

Guaranteeing Quality-of-Service (QoS) of latency-sensitive applications while improving server utilization through application co-location is important yet challenging in modern datacenters. The key challenge is that when applications are co-located on a server, performance interference due to resource contention can be detrimental to the application QoS. Although prior work has proposed techniques to identify "safe" co-locations where application QoS is satisfied by predicting the performance interference on multicores, no such prediction technique on accelerators such as GPUs. In this work, we present Prophet, an approach to precisely predict the performance degradation of latency-sensitive applications on accelerators due to application co-location. We analyzed the performance interference on accelerators through a real system investigation and found that unlike on multicores where the key contentious resources are shared caches and main memory bandwidth, the key contentious resources on accelerators are instead processing elements, accelerator memory bandwidth and PCIe bandwidth. Based on this observation, we designed interference models that enable the precise prediction for processing element, accelerator memory bandwidth and PCIe bandwidth contention on real hardware. By using a novel technique to forecast solorun execution traces of the co-located applications using interference models, Prophet can accurately predict the performance degradation of latency-sensitive applications on non-preemptive accelerators. Using Prophet, we can identify "safe" co-locations on accelerators to improve utilization without violating the QoS target. Our evaluation shows that Prophet can predict the performance degradation with an average prediction error 5.47% on real systems. Meanwhile, based on the prediction, Prophet achieves accelerator utilization improvements of 49.9% on average while maintaining the QoS target of latency-sensitive applications.

源语言英语
主期刊名ASPLOS 2017 - 22nd International Conference on Architectural Support for Programming Languages and Operating Systems
出版商Association for Computing Machinery
17-32
页数16
ISBN(电子版)9781450344654
DOI
出版状态已出版 - 4 4月 2017
活动22nd International Conference on Architectural Support for Programming Languages and Operating Systems, ASPLOS 2017 - Xi'an, 中国
期限: 8 4月 201712 4月 2017

丛书

姓名International Conference on Architectural Support for Programming Languages and Operating Systems - ASPLOS
Part F127193

会议

会议22nd International Conference on Architectural Support for Programming Languages and Operating Systems, ASPLOS 2017
国家/地区中国
Xi'an
时期8/04/1712/04/17

学术指纹

探究 'Prophet: Precise QoS prediction on non-preemptive accelerators to improve utilization in warehouse-scale computers' 的科研主题。它们共同构成独一无二的学术指纹。

引用此