The audio.cpp project has released version 0.4, introducing support for several new high-quality text-to-speech (TTS) and automatic speech recognition (ASR) models, including Higgs Audio v3 TTS 4B and Fish Audio S2 Pro. This release also fully integrates the GGUF format across all supported model families, offering significant speed and VRAM improvements, particularly with Q8 quantization. The project now supports 35 model families and features a new community models area for maturing ports. AI
IMPACT Enhances local LLM capabilities for audio processing, offering faster and more efficient TTS and ASR.
RANK_REASON This is a software release for a specific tool, not a frontier model release or significant industry event.
- audio.cpp
- CPP
- Fish Audio S2 Pro
- GGML
- GGUF
- Higgs Audio v3 TTS 4B
- OuteTTS TTS
- VieNeu-TTS-v3
- Voxtral Realtime ASR
- Qwen3-TTS
- RTX 5090
AI-generated summary · Google Gemini · from 2 sources. How we write summaries →