https://hgpu.org/?p=9061
Efficient GPU implementation of the integral histogram