PulseAugur
EN
LIVE 08:27:44

New Burmese Medical Speech Corpus and Fine-Tuned Whisper Model Developed

Researchers have developed myMediWhisper, a new framework for recognizing Burmese medical speech, addressing limitations in existing models like Whisper. This framework is built upon a 28-hour corpus of medical dialogues recorded and validated by native Burmese speakers. The study fine-tuned Whisper models using both full fine-tuning and parameter-efficient methods, incorporating data augmentation techniques to enhance robustness in noisy environments. The resulting myMediWhisper-Medium model achieved a state-of-the-art Word Error Rate of 23.44% on Burmese clinical dialogues, outperforming larger, general-purpose models. AI

IMPACT Establishes a new benchmark for Burmese medical speech recognition, potentially improving healthcare accessibility in the region.

RANK_REASON The cluster describes a research paper detailing the creation of a specialized corpus and the fine-tuning of an existing model for a specific domain and language. [lever_c_demoted from research: ic=1 ai=1.0]

Read on arXiv cs.CL →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

New Burmese Medical Speech Corpus and Fine-Tuned Whisper Model Developed

COVERAGE [1]

  1. arXiv cs.CL TIER_1 English(EN) · Ye Kyaw Thu, Ye Bhone Lin, Thura Aung, Htet Arkar, Myat Oo Swe, Thet Htet San, Min Thiha Tun, Thazin Myint Oo, Thepchai Supnithi ·

    myMediWhisper: Construction of Burmese Medical Speech Corpus and Whisper Fine-Tuning for Clinical Dialogue ASR

    arXiv:2608.11036v1 Announce Type: new Abstract: Although Whisper models benefit from large-scale multilingual pre-training, their performance on Burmese medical speech remains limited. This work presents a Burmese medical speech recognition framework built on a high-quality 28-ho…