Portable high-order finite element kernels I: Streaming Operations

Noel Chalmers, Tim Warburton
AMD Research, Advanced Micro Devices Inc.
arXiv:2009.10917 [cs.MS], (23 Sep 2020)


   title={Portable high-order finite element kernels I: Streaming Operations},

   author={Noel Chalmers and Tim Warburton},






Download Download (PDF)   View View   Source Source   



This paper is devoted to the development of highly efficient kernels performing vector operations relevant in linear system solvers. In particular, we focus on the low arithmetic intensity operations (i.e., streaming operations) performed within the conjugate gradient iterative method, using the parameters specified in the CEED benchmark problems for high-order hexahedral finite elements. We propose a suite of new Benchmark Streaming tests to focus on the distinct streaming operations which must be performed. We implemented these new tests using the OCCA abstraction framework to demonstrate portability of these streaming operations on different GPU architectures, and propose a simple performance model for such kernels which can accurately capture data movement rates as well as kernel launch costs.
No votes yet.
Please wait...

* * *

* * *

HGPU group © 2010-2024 hgpu.org

All rights belong to the respective authors

Contact us: