https://hgpu.org/?p=19202
Compiler-Driven Performance on Heterogeneous Computing Platforms