https://hgpu.org/?p=3663
Automatically Tuning Sparse Matrix-Vector Multiplication for GPU Architectures