PulseAugur
EN
LIVE 14:12:29

Nemotron Labs unveils unified audio-text LLM, Audex

Nemotron Labs has introduced Audex (Nemotron-Labs-Audex-30B-A3B), a unified audio-text large language model. Built upon the Nemotron-Cascade-2-30B-A3B text-only model, Audex uses a single Transformer decoder to process and generate both audio and text seamlessly. The model was trained on a large dataset of 157.4 billion audio tokens and 320.5 billion text tokens, incorporating supervised training and reinforcement learning techniques. Audex achieves state-of-the-art performance in various audio tasks, including understanding, speech recognition, translation, and generation, while retaining the strong reasoning and knowledge capabilities of its text-based predecessor. AI

IMPACT This unified model architecture could advance multimodal AI capabilities, enabling more seamless integration of audio and text processing in future applications.

RANK_REASON The cluster describes a research paper detailing a new model release.

Read on Hugging Face Daily Papers →

AI-generated summary · Google Gemini · from 5 sources. How we write summaries →

Nemotron Labs unveils unified audio-text LLM, Audex

How we ranked this

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Research
The cluster describes a research paper detailing a new model release.
Source corroboration
5 independent sources
Strong cross-source corroboration — multiple independent publishers covered this within the clustering window.
Topics
model release, paper
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
94 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.
Coverage growth since scoring
+1 source(s) since last score
New sources have picked up this story since our last re-score. Score will update on the next scoring pass.

Full methodology in our editorial standards.

COVERAGE [5]

  1. arXiv cs.AI TIER_1 English(EN) · Zhifeng Kong, Sang-gil Lee, Jaehyeon Kim, Boxin Wang, Zihan Liu, Sungwon Kim, Yang Chen, Arushi Goel, Rajarshi Roy, Wenliang Dai, Zhuolin Yang, Yangyi Chen, Dongfu Jiang, Sreyan Ghosh, Tuomas Rintamaki, Andrew Tao, Jonathan Raiman, Mohammad Shoeybi, Brya… ·

    Unified Audio Intelligence Without Regressing on Text Intelligence

    arXiv:2607.05196v1 Announce Type: cross Abstract: Audio intelligence involves understanding, reasoning about, and generating both audio and speech. In this work, we introduce Nemotron-Labs-Audex-30B-A3B (Audex), a unified audio-text LLM built on Nemotron-Cascade-2-30B-A3B, a stro…

  2. arXiv cs.AI TIER_1 English(EN) · Wei Ping ·

    Unified Audio Intelligence Without Regressing on Text Intelligence

    Audio intelligence involves understanding, reasoning about, and generating both audio and speech. In this work, we introduce Nemotron-Labs-Audex-30B-A3B (Audex), a unified audio-text LLM built on Nemotron-Cascade-2-30B-A3B, a strong text-only MoE LLM. Audex adopts a simple unifie…

  3. Hugging Face Daily Papers TIER_1 English(EN) ·

    Unified Audio Intelligence Without Regressing on Text Intelligence

    Audio intelligence involves understanding, reasoning about, and generating both audio and speech. In this work, we introduce Nemotron-Labs-Audex-30B-A3B (Audex), a unified audio-text LLM built on Nemotron-Cascade-2-30B-A3B, a strong text-only MoE LLM. Audex adopts a simple unifie…

  4. arXiv cs.CL TIER_1 English(EN) · Wei Ping ·

    Unified Audio Intelligence Without Regressing on Text Intelligence

    Audio intelligence involves understanding, reasoning about, and generating both audio and speech. In this work, we introduce Nemotron-Labs-Audex-30B-A3B (Audex), a unified audio-text LLM built on Nemotron-Cascade-2-30B-A3B, a strong text-only MoE LLM. Audex adopts a simple unifie…

  5. Hugging Face Daily Papers TIER_1 English(EN) ·

    Unified Audio Intelligence Without Regressing on Text Intelligence

    A unified audio-text large language model is presented that integrates audio and text processing through a shared transformer decoder, achieving superior performance across multiple audio and speech tasks while maintaining strong text reasoning capabilities.