International Journal of Artificial Intelligence Research
Vol 10, No 1 (2026): June

Minimal Gated Recurrent Unit with Temporal Convolution Based Acoustic Modeling on Speech Recognition System for Evaluating the Recitation of Quran

Isjhar Kautsar (Institut Teknologi Bandung)
Dessi Puji Lestari (Institut Teknologi Bandung)
Aulia Rahmawati (Institut Teknologi dan Bisnis Kalla)



Article Info

Publish Date
01 Jul 2026

Abstract

The use of future context in acoustic modeling seems to give an impact on system performance such us Bidirectional Long Short-Term Memory (BLSTM). It has been used as an acoustic model on Speech Recognition System for Quran recitation and show better result than Hidden Markov Model - Gaussian mixture model (HMM-GMM) with average Word Error Rate (WER) value 4.6%. but, the architectural complexity of BLSTM make the latency during decoding process is high.  To reduce the latency, Minimal Gated Recurrent Unit with Temporal Convolution (mGRUIPTC) acoustic model was used. Text data such as transcription, lexicon, and corpus used in training are represented at phone level to handle phone level detection. The transcription is generated using modified QScript to handle reciting rules in detail. In the test, the system can reduce decode process latency by up to 11 seconds with Phone Error Rate (PER) difference of up to 1.46% compared to BLSTM. However, our model still needs to be trained with more data to detect error recitation better.

Copyrights © 2026






Journal Info

Abbrev

IJAIR

Publisher

Subject

Computer Science & IT Electrical & Electronics Engineering

Description

International Journal Of Artificial Intelligence Research (IJAIR) is a peer-reviewed open-access journal. The journal invites scientists and engineers throughout the world to exchange and disseminate theoretical and practice-oriented topics of Artificial intelligent Research which covers four (4) ...