PulseAugur
EN
LIVE 16:12:37

audio.cpp v0.4 adds Higgs Audio TTS, Fish Audio S2 Pro, and GGUF support

The audio.cpp project has released version 0.4, introducing support for new high-quality text-to-speech (TTS) models including Higgs Audio v3 TTS 4B and Fish Audio S2 Pro. This update also makes GGUF format a first-class citizen across the project, with all supported model families now compatible. Performance improvements are noted, with Q8 quantization showing speed and VRAM gains, and specific models like Higgs Audio TTS running significantly faster than real-time. AI

IMPACT Enhances performance and model compatibility for audio processing tools, potentially improving user experience and efficiency for AI-driven audio applications.

RANK_REASON This is a software release for a specific tool (audio.cpp) that integrates various AI models, rather than a release from a frontier AI lab.

Read on r/StableDiffusion →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

audio.cpp v0.4 adds Higgs Audio TTS, Fish Audio S2 Pro, and GGUF support

COVERAGE [1]

  1. r/StableDiffusion TIER_2 English(EN) · /u/Acceptable-Cycle4645 ·

    [audio.cpp] Release 0.4: Higgs Audio v3 TTS 4B (10x real time)+ Fish Audio S2 Pro in C++/GGML, full GGUF loading, Q8 speed and VRAM gains

    <table> <tr><td> <a href="https://www.reddit.com/r/StableDiffusion/comments/1v4wj6z/audiocpp_release_04_higgs_audio_v3_tts_4b_10x/"> <img alt="[audio.cpp] Release 0.4: Higgs Audio v3 TTS 4B (10x real time)+ Fish Audio S2 Pro in C++/GGML, full GGUF loading, Q8 speed and VRAM gains…