https://hgpu.org/?p=8028
CuBA - a CUDA implementation of BAMPS