https://hgpu.org/?p=16872
Language Modeling with Gated Convolutional Networks