Racial Bias and English Intelligibility on Automatic Speech Recognition among Indonesian Speakers

Open

Tanzir Masykar, Febri Nurrahmi, Roni Agusmaniza, Hery Wiharja, Ary Firnanda, Haimi Ardiansyah

2026 Asian Journal of Human Services Vol. 30 Issue 3 Article Cited by 0 SDG 4SDG 16SDG 17 Quartile

Abstract

Speech recognition software is widely adopted due to its efficiency and convenience, yet its accuracy remains inconsistent across speaker groups. While prior studies show strong performance with native English speakers, limited research has examined its effectiveness with indigenized English accents. This study investigates the intelligibility of English spoken by Indonesian learners using two automatic speech recognition (ASR) platforms: Google Voice Typing and Macintosh Dictation. A total of 27 Indonesian EFL speakers read 27 English words embedded in carrier sentences, including both monophthongs and diphthongs. The same word list was also recorded by native American English speakers for comparison. Each utterance was transcribed by both ASR platforms. Recognition performance was evaluated using recognition accuracy, as the study employed an Isolated Word Recognition (IWR) approach. Recognition accuracy between ASR platforms was analyzed using the Wilcoxon Signed Rank test, while gender differences were examined using the Mann–Whitney U test. The results showed that Google Voice Typing consistently outperformed Macintosh Dictation in transcribing Indonesian EFL speech for both overall vowels and each vowel type. Both systems recognized diphthongs more accurately than monophthongs. Native English speech was transcribed more accurately than Indonesian speech on both platforms. Additionally, both male and female participants achieved significantly higher recognition accuracy with Google, although no significant gender-based differences were observed within each ASR system. These findings suggest that Indonesian-accented English presents intelligibility challenges for current ASR technologies, highlighting the need for more inclusive speech recognition systems that can support greater linguistic diversity. © 2026 Tanzir MASYKAR, Febri NURRAHMI, Roni AGUSMANIZA, Hery WIHARJA, Ary FIRNANDA & Haimi ARDIANSYAH.

Affiliations

Akademi Komunitas Negeri Aceh Barat, Indonesia; Universitas Syiah Kuala, Indonesia; Universitas Teuku Umar, Indonesia

Research at a Glance

Premium content — register to unlock

Research at a Glance

Register to unlock

Topics & SDG Alignment

Premium content — register to unlock

Topics & SDG Alignment

Register to unlock

Collaboration

Premium content — register to unlock

Collaboration

Register to unlock

Author Profile (Selected)

Premium content — register to unlock

Author Profile (Selected)

Register to unlock

References Overview

Premium content — register to unlock

References Overview

Register to unlock

Journal & Source

Premium content — register to unlock

Journal & Source

Register to unlock

Metadata & Integrity

Premium content — register to unlock

Metadata & Integrity

Register to unlock