Optimization and performance evaluation of the IDR iterative Krylov solver on GPUs

Anzt H, Kreutzer M, Ponce E, Peterson GD, Wellein G, Dongarra J (2016)


Publication Type: Journal article

Publication year: 2016

Journal

Publisher: SAGE Publications (UK and US)

URI: http://hpc.sagepub.com/content/early/2016/05/05/1094342016646844.abstract

DOI: 10.1177/1094342016646844

Abstract

In this paper, we present an optimized GPU implementation for the induced dimension reduction algorithm. We improve data locality, combine it with an efficient sparse matrix vector kernel, and investigate the potential of overlapping computation with communication as well as the possibility of concurrent kernel execution. A comprehensive performance evaluation is conducted using a suitable performance model. The analysis reveals efficiency of up to 90%, which indicates that the implementation achieves performance close to the theoretically attainable bound.

Authors with CRIS profile

Involved external institutions

How to cite

APA:

Anzt, H., Kreutzer, M., Ponce, E., Peterson, G.D., Wellein, G., & Dongarra, J. (2016). Optimization and performance evaluation of the IDR iterative Krylov solver on GPUs. International Journal of High Performance Computing Applications. https://doi.org/10.1177/1094342016646844

MLA:

Anzt, Hartwig, et al. "Optimization and performance evaluation of the IDR iterative Krylov solver on GPUs." International Journal of High Performance Computing Applications (2016).

BibTeX: Download