https://hgpu.org/?p=13061
Code Optimization on Kepler GPUs and Xeon Phi