XLS-R Deep Learning Model For Multilingual ASR On Low- Resource Languages: Indonesian, Javanese, And Sundanese
2024 Β· Panji Arisaputra, Alif Tri Handoyo, Amalia Zahra
Abstract
This research paper focuses on the development and evaluation of Automatic Speech Recognition (ASR) technology using the XLS-R 300m model. The study aims to improve ASR performance in converting spoken language into written text, specifically for Indonesian, Javanese, and Sundanese languages. The paper discusses the testing procedures, datasets used, and methodology employed in training and evaluating the ASR systems. The results show that the XLS-R 300m model achieves competitive Word Error Rate (WER) measurements, with a slight compromise in performance for Javanese and Sundanese languages. The integration of a 5-gram KenLM language model significantly reduces WER and enhances ASR accuracy. The research contributes to the advancement of ASR technology by addressing linguistic diversity and improving performance across various languages. The findings provide insights into optimizing ASR accuracy and applicability for diverse linguistic contexts.
Authors
(none)
Tags
Stats
Related papers
- Enhancing Indonesian Automatic Speech Recognition: Evaluating Multilingual Models With Diverse Speech Variabilities (2024)4.52
- Towards Building Speech Large Language Models For Multitask Understanding In Low-resource Languages (2025)0.00
- Whisper-lm: Improving ASR Models With Language Models For Low-resource Languages (2025)3.29
- Evaluating Standard And Dialectal Frisian ASR: Multilingual Fine-tuning And Language Identification For Improved Low-resource Performance (2025)0.00
- Semi-supervised Development Of ASR Systems For Multilingual Code-switched Speech In Under-resourced Languages (2020)0.00
- Transsion Tsup's Speech Recognition System For ASRU 2023 MADASR Challenge (2023)0.00
- Dyn-asr: Compact, Multilingual Speech Recognition Via Spoken Language And Accent Identification (2021)5.24
- Building Robust And Scalable Multilingual ASR For Indian Languages (2025)0.00