https://hgpu.org/?p=15863
Improving GPU Performance: Reducing Memory Conflicts and Latency