https://hgpu.org/?p=9020
Performance Traps in OpenCL for CPUs