Providing performance portable numerics for Intel GPUs
- Steinbuch Centre for Computing Karlsruhe Institute of Technology Karlsruhe Baden‐Württemberg Germany
- Steinbuch Centre for Computing Karlsruhe Institute of Technology Karlsruhe Baden‐Württemberg Germany, The Innovative Computing Laboratory University of Tennessee Knoxville Tennessee
Summary With discrete Intel GPUs entering the high‐performance computing landscape, there is an urgent need for production‐ready software stacks for these platforms. In this article, we report how we enable the Ginkgo math library to execute on Intel GPUs by developing a kernel backed based on the DPC++ programming environment. We discuss conceptual differences between the CUDA and DPC++ programming models and describe workflows for simplified code conversion. We evaluate the performance of basic and advanced sparse linear algebra routines available in Ginkgo's DPC++ backend in the hardware‐specific performance bounds and compare against routines providing the same functionality that ship with Intel's oneMKL vendor library.
- Research Organization:
- University of Tennessee, Knoxville, TN (United States)
- Sponsoring Organization:
- USDOE; USDOE National Nuclear Security Administration (NNSA); USDOE Office of Science (SC)
- OSTI ID:
- 1894739
- Journal Information:
- Concurrency and Computation. Practice and Experience, Journal Name: Concurrency and Computation. Practice and Experience Journal Issue: 20 Vol. 35; ISSN 1532-0626
- Publisher:
- Wiley Blackwell (John Wiley & Sons)Copyright Statement
- Country of Publication:
- United Kingdom
- Language:
- English
Similar Records
Evaluating CUDA Portability with HIPCL and DPCT
Portability for GPU-accelerated molecular docking applications for cloud and HPC: can portable compiler directives provide performance across all platforms?