https://hgpu.org/?p=8788
Three Dimensional Fast Fourier Transform CUDA Implementation