PulseAugur
EN
LIVE 09:02:40
Русский(RU) GigaChat Audio научился распознавать эмоции и находить нужные моменты в трёхчасовых записях: gigachat open source

Sber releases open-source GigaChat Audio with emotion detection and timestamping

Sber has updated its GigaChat Audio model, enhancing its ability to process audio files directly without full transcription and to detect emotional tones in speech. The updated model can now identify speakers, summarize content, and provide timestamps for specific moments within recordings up to three hours long. Additionally, Sber has released an open-source version, GigaChat3.1-Audio-10B, and the GigaAM Multilingual speech recognition family, supporting multiple languages. AI

IMPACT This release enhances audio processing capabilities for developers, potentially improving applications like meeting summarization and call analysis.

RANK_REASON The cluster describes a new model release and open-source availability from a major AI lab (Sber). [lever_c_demoted from frontier_release: ic=1 ai=1.0]

Read on dev.to — LLM tag →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

Sber releases open-source GigaChat Audio with emotion detection and timestamping

COVERAGE [1]

  1. dev.to — LLM tag TIER_1 Русский(RU) · Promptra Team ·

    GigaChat Audio learned to recognize emotions and find relevant moments in three-hour recordings: gigachat open source

    <p>14 июля 2026 года Сбер обновил аудиомодель в своём ИИ-помощнике и выложил облегчённую версию для разработчиков. Главное, что изменилось на практике: модель работает с голосом и аудиофайлами напрямую, без обязательной предварительной расшифровки всей записи в текст, ловит позит…