English

High-Throughput Parallel Viterbi Decoder on GPU Tensor Cores

Distributed, Parallel, and Cluster Computing 2020-11-30 v1 Signal Processing

Abstract

Many research works have been performed on implementation of Vitrerbi decoding algorithm on GPU instead of FPGA because this platform provides considerable flexibility in addition to great performance. Recently, the recently-introduced Tensor cores in modern GPU architectures provide incredible computing capability. This paper proposes a novel parallel implementation of Viterbi decoding algorithm based on Tensor cores in modern GPU architectures. The proposed parallel algorithm is optimized to efficiently utilize the computing power of Tensor cores. Experiments show considerable throughput improvements in comparison with previous works.

Keywords

Cite

@article{arxiv.2011.13579,
  title  = {High-Throughput Parallel Viterbi Decoder on GPU Tensor Cores},
  author = {Alireza Mohammadidoost and Matin Hashemi},
  journal= {arXiv preprint arXiv:2011.13579},
  year   = {2020}
}

Comments

arXiv admin note: substantial text overlap with arXiv:2011.09337