https://hgpu.org/?p=18436
Optimizing Communication for Clusters of GPUs