https://hgpu.org/?p=14129
Autotuning Tensor Contraction Computations on GPUs