Researchers have introduced WASIL, a new dataset designed to improve Arabic spoken interactions with Large Language Models (LLMs). The dataset includes over 8,500 turns of in-the-wild spoken interactions, complete with audio, ASR hypotheses, assistant responses, and user feedback, with 14.2% of interactions marked as disliked. WASIL also features a 2,000-turn test set covering Modern Standard Arabic and four major dialects, along with annotations for answerability to distinguish between ASR errors and genuine unanswerability. This resource aims to facilitate better evaluation of LLM voice assistants by separating speech recognition issues from the model's inherent capabilities. AI
IMPACT Enables more accurate evaluation and development of Arabic-speaking AI assistants.
RANK_REASON The item is a research paper detailing a new dataset for Arabic spoken interactions with LLMs. [lever_c_demoted from research: ic=1 ai=1.0]
- Arabic
- arXiv
- DagsHub
- Gotit.pub
- Hugging Face
- Large Language Models
- Modern Standard Arabic
- ScienceCast
- WASIL
- Zien Sheikh Ali
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →