https://hgpu.org/?p=13586
Model-driven optimisation of memory hierarchy and multithreading on GPUs