Researchers have introduced Nuha-Speech, a project aimed at developing general-purpose Arabic Speech Large Language Models (speech-LLMs). This initiative addresses the underrepresentation of Arabic in multilingual speech-LLMs by creating a large-scale Arabic Speech Question-Answering (SQA) corpus with over 1.5 million training samples. The corpus was used to fine-tune Qwen Omni model variants, and a comprehensive evaluation framework was designed to establish foundational infrastructures for Arabic speech-LLMs despite limited resources. AI
IMPACT This work aims to improve the representation and capabilities of Arabic language models in speech-based AI applications.
RANK_REASON The cluster describes a research paper detailing the creation of a new dataset and the fine-tuning of existing models for a specific language and modality. [lever_c_demoted from research: ic=1 ai=1.0]
- Arabic
- Arabic Speech Question-Answering
- arXiv
- DagsHub
- Hugging Face
- Nuha-Speech
- Qwen Omni
- Speech Large Language Models
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →