https://hgpu.org/?p=8361
Exploiting Data Parallelism in GPUs