Optimizing a shared virtual memory system for a heterogeneous CPU-accelerator platform

Abstract

The client computing platform is moving towards a heterogeneous architecture that combines scalar-oriented CPU cores and throughput-oriented accelerator cores. Recognizing that existing programming models for such heterogeneous platforms are still difficult for most programmers, we advocate a shared virtual memory programming model to improve programmability. In this paper, we focus on performance, and demonstrate that users need not sacrifice performance for programmability. We describe our approaches, experiences, and results in optimizing MYO on a heterogeneous platform consisting of a CPU and an Aubrey Isle accelerator. Our efforts involve the whole system software stack including the OS, runtime, and application.

DOI: 10.1145/1945023.1945035

Extracted Key Phrases

11 Figures and Tables

Cite this paper

@article{Yan2011OptimizingAS, title={Optimizing a shared virtual memory system for a heterogeneous CPU-accelerator platform}, author={Shoumeng Yan and Xiaocheng Zhou and Ying Gao and Hu Chen and Gansha Wu and Sai Luo and Bratin Saha}, journal={Operating Systems Review}, year={2011}, volume={45}, pages={92-100} }