RSIS Repository Open-access research from RSIS International journals

International Journal of Research and Innovation in Applied Science (IJRIAS)

Indigenous People’s Language Identification Using Machine Learning for Linguistic Preservation

byAdrales, Lorelyn F; Garingo, Joshua Razzi B; Geraldez, Jan Anthony Q; Taboada, Vene Lucille T; Taladtad, Jelan Roy L.

Published November 8, 2025  •  Vol. 10, Issue 10, pp. 973–984Open Access
DOI: 10.51584/IJRIAS.2025.1010000080

Abstract

Language is a fundamental aspect of human identity, deeply connected to geographical origins, cultural heritage, and social belonging. However, many indigenous languages across the world are gradually declining due to modernization, migration, and the growing influence of technology and global languages. The loss of these languages often leads to the disappearance of cultural values, oral traditions, and historical knowledge. This study explores the integration of machine learning techniques such as Long Short-Term Memory (LSTM), Yoon Kim’s Convolutional Neural Network model, and TextConvoNet in developing a mobile text-to-text identification and translation application for Blaan dialects spoken in General Santos City, Polomolok, and Sarangani. The goal of the application is to aid in the preservation and revitalization of the Blaan language while providing an accessible platform for both native speakers and learners to understand, translate, and communicate in their local dialects.
To evaluate the usability and effectiveness of the application, User Acceptance Testing (UAT) was conducted among selected users. Data were collected through structured interviews, document analysis, and standardized evaluation tools to ensure comprehensive assessment and validation. Experimental results showed that the TextConvoNet model achieved the highest accuracy rate of 74.00 percent, surpassing the performance of both LSTM and CNN-based models. This demonstrates the model’s efficiency in identifying and classifying Blaan dialects, highlighting its potential in the field of Natural Language Processing (NLP).
Future research should focus on expanding the dataset by collecting transcriptions from diverse age groups, locations, and communication contexts to improve model generalization and accuracy. Further refinement of the model’s architecture and parameter tuning is also recommended to enhance dialect classification and translation capabilities. Moreover, integrating speech-to-text and text-to-speech functionalities could facilitate real-time translation, pronunciation learning, and accessibility for non-literate speakers, ensuring the continued preservation and appreciation of indigenous languages.

Keywords: Natural Language Processing (NLP), TextConvoNet, Yoon Kim, LSTM

JournalInternational Journal of Research and Innovation in Applied Science (IJRIAS)
ISSN2454-6194
Volume / IssueVolume 10, Issue 10
Pages973–984
Publication dateNovember 8, 2025
DOI10.51584/IJRIAS.2025.1010000080
PublisherRSIS International
LicenseOpen Access

How to cite this article

Adrales, Lorelyn F, Garingo, Joshua Razzi B, Geraldez, Jan Anthony Q, Taboada, Vene Lucille T, & Taladtad, Jelan Roy L. (2025). Indigenous People’s Language Identification Using Machine Learning for Linguistic Preservation. International Journal of Research and Innovation in Applied Science (IJRIAS), 10(10), 973-984. https://doi.org/10.51584/IJRIAS.2025.1010000080

BibTeX

@article{Adrales2025,
  title   = {Indigenous People’s Language Identification Using Machine Learning for Linguistic Preservation},
  author  = {Adrales, Lorelyn F and Garingo, Joshua Razzi B and Geraldez, Jan Anthony Q and Taboada, Vene Lucille T and Taladtad, Jelan Roy L.},
  journal = {International Journal of Research and Innovation in Applied Science (IJRIAS)},
  volume  = {10},
  number  = {10},
  pages   = {973--984},
  year    = {2025},
  doi     = {10.51584/IJRIAS.2025.1010000080},
  publisher = {RSIS International}
}