https://hgpu.org/?p=8861
GPU Enhanced Stream-Based Matrix Multiplication