High-Throughput Parallel Viterbi Decoder on GPU Tensor Cores
Department of Electrical Engineering, Sharif University of Technology, Tehran, Iran
arXiv:2011.13579 [cs.DC], (27 Nov 2020)
@misc{mohammadidoost2020highthroughput,
title={High-Throughput Parallel Viterbi Decoder on GPU Tensor Cores},
author={Alireza Mohammadidoost and Matin Hashemi},
year={2020},
eprint={2011.13579},
archivePrefix={arXiv},
primaryClass={cs.DC}
}
Many research works have been performed on implementation of Vitrerbi decoding algorithm on GPU instead of FPGA because this platform provides considerable flexibility in addition to great performance. Recently, the recently-introduced Tensor cores in modern GPU architectures provide incredible computing capability. This paper proposes a novel parallel implementation of Viterbi decoding algorithm based on Tensor cores in modern GPU architectures. The proposed parallel algorithm is optimized to efficiently utilize the computing power of Tensor cores. Experiments show considerable throughput improvements in comparison with previous works.
December 6, 2020 by hgpu