Researchers have developed a new method for annotating mistakes in Quran memorization transcripts generated by automatic speech recognition (ASR). This annotation process distinguishes between actual errors, repetitions, and accepted spelling variations. The developed evaluator scores these labels and their positions, achieving a label-aware F1 score of 0.525 and a localization F1 score of 0.826 with a plain diff. Preliminary tests with six coding agents and eight models showed varied performance, with most outperforming baseline methods, highlighting the importance of convention and normalization in ASR evaluation. AI
IMPACT This research could lead to more accurate evaluation of ASR systems for specialized domains like religious text memorization.
RANK_REASON The cluster contains an academic paper detailing a new annotation method for ASR transcripts. [lever_c_demoted from research: ic=1 ai=1.0]
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →