Enter your keyword

2-s2.0-85059931413

[vc_empty_space][vc_empty_space]

Online Speech Decoding Optimization Strategy with Viterbi Algorithm on GPU

Arsadjaja A.R.a, Imam Kistijantoro A.a

a School of Electrical Engineering and Informatics, Institut Teknologi Bandung, Bandung, Indonesia

[vc_row][vc_column][vc_row_inner][vc_column_inner][vc_separator css=”.vc_custom_1624529070653{padding-top: 30px !important;padding-bottom: 30px !important;}”][/vc_column_inner][/vc_row_inner][vc_row_inner layout=”boxed”][vc_column_inner width=”3/4″ css=”.vc_custom_1624695412187{border-right-width: 1px !important;border-right-color: #dddddd !important;border-right-style: solid !important;border-radius: 1px !important;}”][vc_empty_space][megatron_heading title=”Abstract” size=”size-sm” text_align=”text-left”][vc_column_text]© 2018 IEEE.Automatic Speech Recognition (ASR) has been popular recently. But the current algorithm for speech recognition is slow and needed the way to recognize faster. One way to achieve it is with GPU, which provides parallel computation; but ASR is hard to parallelize directly. This paper describes how to build parallel ASR system, which requires several steps. First, we must convert the data structure to make it compatible with GPU, then we have to make several kernels that equivalent to the serial algorithm in CPU.We will describe several optimization strategies for make ASR run much faster after we got the correct GPU program. Those strategies are based on profiling result and analysis of the GPU program execution flow. Best implementation that we had have a speedup around 5.59-6.18 times from the serial CPU implementation.[/vc_column_text][vc_empty_space][vc_separator css=”.vc_custom_1624528584150{padding-top: 25px !important;padding-bottom: 25px !important;}”][vc_empty_space][megatron_heading title=”Author keywords” size=”size-sm” text_align=”text-left”][vc_column_text]Automatic speech recognition,GPU programs,Online speech,Optimization strategy,parallel,Parallel Computation,Serial algorithms[/vc_column_text][vc_empty_space][vc_separator css=”.vc_custom_1624528584150{padding-top: 25px !important;padding-bottom: 25px !important;}”][vc_empty_space][megatron_heading title=”Indexed keywords” size=”size-sm” text_align=”text-left”][vc_column_text]ASR,decoding,GPU,optimization strategy,parallel[/vc_column_text][vc_empty_space][vc_separator css=”.vc_custom_1624528584150{padding-top: 25px !important;padding-bottom: 25px !important;}”][vc_empty_space][megatron_heading title=”Funding details” size=”size-sm” text_align=”text-left”][vc_column_text][/vc_column_text][vc_empty_space][vc_separator css=”.vc_custom_1624528584150{padding-top: 25px !important;padding-bottom: 25px !important;}”][vc_empty_space][megatron_heading title=”DOI” size=”size-sm” text_align=”text-left”][vc_column_text]https://doi.org/10.1109/ICAICTA.2018.8541343[/vc_column_text][/vc_column_inner][vc_column_inner width=”1/4″][vc_column_text]Widget Plumx[/vc_column_text][/vc_column_inner][/vc_row_inner][/vc_column][/vc_row][vc_row][vc_column][vc_separator css=”.vc_custom_1624528584150{padding-top: 25px !important;padding-bottom: 25px !important;}”][/vc_column][/vc_row]