High-Throughput Parallel Viterbi Decoder on GPU Tensor Cores
Many research works have been performed on implementation of Vitrerbi decoding algorithm on GPU instead of FPGA because this platform provides considerable flexibility in addition to great performance. Recently, the recently-introduced Tensor cores in modern GPU architectures provide incredible comp...
Gespeichert in:
Hauptverfasser: | , |
---|---|
Format: | Artikel |
Sprache: | eng |
Schlagworte: | |
Online-Zugang: | Volltext bestellen |
Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Zusammenfassung: | Many research works have been performed on implementation of Vitrerbi
decoding algorithm on GPU instead of FPGA because this platform provides
considerable flexibility in addition to great performance. Recently, the
recently-introduced Tensor cores in modern GPU architectures provide incredible
computing capability. This paper proposes a novel parallel implementation of
Viterbi decoding algorithm based on Tensor cores in modern GPU architectures.
The proposed parallel algorithm is optimized to efficiently utilize the
computing power of Tensor cores. Experiments show considerable throughput
improvements in comparison with previous works. |
---|---|
DOI: | 10.48550/arxiv.2011.13579 |