https://hgpu.org/?p=13508
Stochastic Gradient Descent on GPUs