https://hgpu.org/?p=6929
Linear Algebra Algorithms for Hybrid Architectures with XKaapi