ASR Error Correction in Low-Resource Burmese with Alignment-Enhanced Transformers using Phonetic Features

Lin, Ye Bhone, Aung, Thura, Thu, Ye Kyaw, Oo, Thazin Myint

Nov-27-2025–arXiv.org Artificial Intelligence

Abstract--This paper investigates sequence-to-sequence T ransformer models for automatic speech recognition (ASR) error correction in low-resource Burmese, focusing on different feature integration strategies including IP A and alignment information. T o our knowledge, this is the first study addressing ASR error correction specifically for Burmese. W e evaluate five ASR backbones and show that our ASR Error Correction (AEC) approaches consistently improve word-and character-level accuracy over baseline outputs. The proposed AEC model, combining IP A and alignment features, reduced the average WER of ASR models from 51.56 to 39.82 before augmentation (and 51.56 to 43.59 after augmentation) and improving chrF++ scores from 0.5864 to 0.627, demonstrating consistent gains over the baseline ASR outputs without AEC. Our results highlight the robustness of AEC and the importance of feature design for improving ASR outputs in low-resource settings.

data quality, error correction, machine learning, (17 more...)

arXiv.org Artificial Intelligence

Nov-27-2025

arXiv.org PDF

Add feedback

Country:
- Europe (0.46)
- Asia
  - Myanmar (0.16)
  - Thailand (0.14)

Genre:
- Research Report > New Finding (1.00)

Technology:
- Information Technology
  - Data Science > Data Quality
    - Data Cleaning (1.00)
  - Artificial Intelligence
    - Speech > Speech Recognition (1.00)
    - Machine Learning > Neural Networks
      - Deep Learning (0.46)

Duplicate Docs Excel Report

Title
None found

Similar Docs Excel Report more

Title	Similarity	Source
None found