PulseAugur
EN
LIVE 15:23:14

NVIDIA launches Nemotron 3.5 Lightning for efficient agentic AI

NVIDIA has launched Nemotron 3.5 Lightning, a 30-billion-parameter mixture-of-experts model designed for efficient agentic AI workloads. This model offers up to 4x faster output speed and 30% faster task completion compared to similar-sized models, with a 1 million token context window. Alongside the model, NVIDIA released NeMo Switchyard, an open-source library for intelligent routing within agent tools, enabling requests to be directed to the most suitable model. Nemotron 3.5 Lightning is available on platforms like Together AI and is customizable for specialized tasks across various industries, including cybersecurity, legal services, and software development. AI

IMPACT Accelerates development of specialized AI agents by providing an efficient, customizable open model and intelligent routing capabilities.

RANK_REASON NVIDIA's official announcement of a new model family member with performance metrics and system integration details.

Read on The Decoder →

AI-generated summary · Google Gemini · from 26 sources. How we write summaries →

NVIDIA launches Nemotron 3.5 Lightning for efficient agentic AI

How we ranked this

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Frontier Release
NVIDIA's official announcement of a new model family member with performance metrics and system integration details.
Source corroboration
26 independent sources
Strong cross-source corroboration — multiple independent publishers covered this within the clustering window.
Topics
model release, product, infra
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
46 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.

Full methodology in our editorial standards.

COVERAGE [26]

  1. NVIDIA Blog TIER_1 English(EN) · Kari Briski ·

    NVIDIA Nemotron 3.5 Lightning and NeMo Switchyard Deliver Faster, Smarter, More Efficient Agentic AI

    As AI shifts from chatbots to autonomous agents, open models are serving market demands for full control over where AI runs and how it’s deployed and evolves. Today, NVIDIA is expanding its Nemotron 3 model family with Nemotron 3.5 Lightning, the highest-efficiency model in its c…

  2. X — Together (inference / OSS) TIER_1 English(EN) · togethercompute ·

    Nemotron 3.5 Lightning is available on Together AI through Dedicated Model Inference, giving teams reserved capacity and predictable performance for high-volume

    Nemotron 3.5 Lightning is available on Together AI through Dedicated Model Inference, giving teams reserved capacity and predictable performance for high-volume agent workloads. Start building: https://t.co/4U09N6nL56

  3. X — Together (inference / OSS) TIER_1 English(EN) · togethercompute ·

    What Nemotron 3.5 Lightning brings:

    What Nemotron 3.5 Lightning brings: → Fastest open model in its class → Up to 4x higher throughput → Up to 30% faster task completion → 30B hybrid MoE (3B active) → 1M context and DFlash speculative decoding → Open and customizable for post-training

  4. X — Together (inference / OSS) TIER_1 English(EN) · togethercompute ·

    NVIDIA Nemotron 3.5 Lightning is now live on Together AI.

    NVIDIA Nemotron 3.5 Lightning is now live on Together AI. The fastest open model in its class is built for always-on agents that need to complete high-volume, specialized work quickly. https://t.co/ScHmJGOIGk

  5. The Decoder TIER_1 English(EN) · Matthias Bastian ·

    Nvidia's open-weight Nemotron 3.5 Lightning prioritizes speed over maximum intelligence

    <p><img alt="" class="attachment-full size-full wp-post-image" height="1152" src="https://the-decoder.com/wp-content/uploads/2026/08/nvidia_logo.png" style="height: auto; margin-bottom: 10px;" width="2048" /></p> <p> Nvidia's Nemotron 3.5 Lightning is an open-weights model with j…

  6. MarkTechPost TIER_1 English(EN) · Asif Razzaq ·

    NVIDIA AI Releases Nemotron 3.5 Lightning: A 30B Open MoE with 3B Active Parameters, and NeMo Switchyard Model Router

    <p>NVIDIA's open 30B MoE targets the agent execution layer, with Switchyard routing each step to the cheapest capable model.</p> <p>The post <a href="https://www.marktechpost.com/2026/08/11/nvidia-ai-releases-nemotron-3-5-lightning-and-nemo-switchyard/">NVIDIA AI Releases Nemotro…

  7. Towards AI TIER_1 English(EN) · allglenn ·

    Nemotron 3.5: NVIDIA is Open-Sourcing AI to Protect its GPU Monopoly

    <div class="medium-feed-item"><p class="medium-feed-image"><a href="https://pub.towardsai.net/nemotron-3-5-nvidia-is-open-sourcing-ai-to-protect-its-gpu-monopoly-fb9f66e65056?source=rss----98111c9905da---4"><img src="https://cdn-images-1.medium.com/max/1672/1*8SYoQAyVmnBeREUzX_Dx…

  8. Mastodon — sigmoid.social TIER_1 English(EN) · [email protected] ·

    NVIDIA’s Nemotron 3.5 Lightning proves that sparse compute is the new standard for enterprise deployment. By routing queries through a 30B Mixture-of-Experts ar

    NVIDIA’s Nemotron 3.5 Lightning proves that sparse compute is the new standard for enterprise deployment. By routing queries through a 30B Mixture-of-Experts architecture that only activates 3B parameters per token, it drastically lowers inference latency. # LLMs # AI

  9. Mastodon — sigmoid.social TIER_1 English(EN) · [email protected] ·

    Nvidia is reportedly building a trillion-parameter Nemotron 4 AI model, expanding its focus beyond chips into models and software. Source: Tech Wire Asia https:

    Nvidia is reportedly building a trillion-parameter Nemotron 4 AI model, expanding its focus beyond chips into models and software. Source: Tech Wire Asia https:// techwireasia.com/2026/08/nvidi a-nemotron-4-trillion-parameter-ai-model/ # AI # Nvidia

  10. Mastodon — sigmoid.social TIER_1 English(EN) · [email protected] ·

    https://www. europesays.com/3187764/ NVIDIA Agentic AI Models Power Nemotron 3.5 Lightning # agentic # AgenticAI # AgenticArtificialIntelligence # AI # Artifici

    https://www. europesays.com/3187764/ NVIDIA Agentic AI Models Power Nemotron 3.5 Lightning # agentic # AgenticAI # AgenticArtificialIntelligence # AI # ArtificialIntelligence # models # Nvidia

  11. dev.to — LLM tag TIER_1 English(EN) · Prabhakar Chaudhary ·

    Nemotron 3.5 Lightning and NeMo Switchyard: How NVIDIA Is Solving the Execution Layer Problem in Agentic AI

    <h1> Nemotron 3.5 Lightning and NeMo Switchyard: How NVIDIA Is Solving the Execution Layer Problem in Agentic AI </h1> <p>When people talk about AI agents, the conversation usually centers on the planning model — the frontier system that reasons, decomposes tasks, and decides wha…

  12. dev.to — LLM tag TIER_1 English(EN) · 武乐丹 ·

    Nemotron 3.5 Lightning: Nvidia's Open-Weight Agent Executor, the Switchyard Routing Debate, and Qwen's 27B Coming in 48 Hours

    <p>Nvidia dropped Nemotron 3.5 Lightning on Aug 11 and the open-source community spent two days arguing about it — not about the model itself, but about the routing system Nvidia shipped alongside it. Meanwhile, Qwen's 3.8-27B hits Hugging Face in two days. Here's the signal insi…

  13. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    NVIDIA has released Nemotron 3.5 Lightning, a 30B open AI model built for agentic workflows. Paired with NeMo Switchyard, it routes each step to the most capabl

    NVIDIA has released Nemotron 3.5 Lightning, a 30B open AI model built for agentic workflows. Paired with NeMo Switchyard, it routes each step to the most capable model, delivering up to 4x faster output. Ready for single-GPU deployment. https:// marktechpost.com/2026/08/11/nv idi…

  14. r/LocalLLaMA TIER_1 English(EN) · /u/curiousily_ ·

    Tested Nemotron 3.5 Lightning locally on coding, Hermes Agent and agentic work

    <!-- SC_OFF --><div class="md"><p>Ran the model with quants (Q5) and MTP by <a href="https://huggingface.co/bartowski/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-GGUF">bartowski</a> with llama.cpp server.</p> <p>It takes ~24GB ram running on M5 Pro with 48GB at about 65t/s. On some tas…

  15. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    Nvidia is reportedly building Nemotron 4, a 1 trillion parameter model, to compete with open AI models Source: Channel News Asia Technology https://www. channel

    Nvidia is reportedly building Nemotron 4, a 1 trillion parameter model, to compete with open AI models Source: Channel News Asia Technology https://www. channelnewsasia.com/business/n vidia-building-1-trillion-parameter-nemotron-4-rival-open-ai-models-information-reports-6312521 …

  16. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    NVIDIA releases Nemotron 3.5 Lightning and NeMo Switchyard to speed up agentic AI workflows with faster output and task completion Source: NVIDIA Blog https://

    NVIDIA releases Nemotron 3.5 Lightning and NeMo Switchyard to speed up agentic AI workflows with faster output and task completion Source: NVIDIA Blog https:// blogs.nvidia.com/blog/nemotron -lightning-switchyard-rtx-dgx/ # AI # Nvidia

  17. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    🚀💡 Ah, the Nemotron 3.5 Lightning—a name so cumbersome, it takes longer to read than the # AI takes to become self-aware. 😏 NVIDIA's new toy promises to control

    🚀💡 Ah, the Nemotron 3.5 Lightning—a name so cumbersome, it takes longer to read than the # AI takes to become self-aware. 😏 NVIDIA's new toy promises to control everything from toasters to data centers, but we all know it's just a fancy way to burn your budget and bandwidth. 🙄 ht…

  18. Mastodon — fosstodon.org TIER_1 Deutsch(DE) · [email protected] ·

    RT @MiaAI_lab: Just released: @NVIDIAAI Nemotron 3.5 Lightning 30B A3B, specifically for high-volume execution layer of long-running AI agents entwic

    RT @MiaAI_lab: Soeben veröffentlicht: @NVIDIAAI Nemotron 3.5 Lightning 30B A3B, speziell für die hochvolumige Ausführungsschicht langlaufender KI-Agenten entwickelt ✨ Tool-Aufrufe, Validierung, Formatierung und Subagent-Arbeit sollten sehr schnell und dabei präzise ablaufen. Ich …

  19. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    Ubuntu users can now run NVIDIA’s Nemotron 3.5 Lightning model through Canonical’s newly released inference Snap. https:// linuxiac.com/nvidia-nemotron-3 -5-lig

    Ubuntu users can now run NVIDIA’s Nemotron 3.5 Lightning model through Canonical’s newly released inference Snap. https:// linuxiac.com/nvidia-nemotron-3 -5-lightning-lands-on-ubuntu-via-snap/ # linux # ubuntu # nvidia # ai # opensource

  20. Mastodon — mastodon.social TIER_1 Deutsch(DE) · [email protected] ·

    RT @MiaAI_lab: Just released: @NVIDIAAI Nemotron 3.5 Lightning 30B A3B, specifically for high-volume execution layer of long-running AI agents

    RT @MiaAI_lab: Soeben veröffentlicht: @NVIDIAAI Nemotron 3.5 Lightning 30B A3B, speziell für die Hochvolumen-Ausführungsschicht langlaufender KI-Agenten entwickelt ✨ Tool-Aufrufe, Validierung, Formatierung und Subagenten-Arbeit sollten sehr schnell ablaufen, dabei aber präzise bl…

  21. Mastodon — mastodon.social TIER_1 日本語(JA) · [email protected] ·

    NVIDIA releases free high-speed model "Nemotron 3.5 Lightning" for agent AI https:// pc.watch.impress.co.jp/docs/ne ws/2132324.html # impress # market # AI # other

    NVIDIA、エージェントAI向け高速モデル「Nemotron 3.5 Lightning」無償公開 https:// pc.watch.impress.co.jp/docs/ne ws/2132324.html # impress # 市場 # AI # その他

  22. Mastodon — mastodon.social TIER_1 Italiano(IT) · AI_BEAR_NEWS ·

    🚀 Nvidia Launches Nemotron 3.5 Lightning: Open AI for Always-On Agents. An open 30B MoE model with 3B active parameters, designed for always-on AI agents. 4x faster

    🚀 Nvidia lancia Nemotron 3.5 Lightning: AI open per agenti Un modello open 30B MoE con 3B parametri attivi, progettato per agenti AI sempre online. 4x più veloce, funziona su singola GPU. Perfetto per code review, sicurezza e task specializzati. Fonte: NVIDIA Blog Segui 👇 # Nemot…

  23. Mastodon — mastodon.social TIER_1 English(EN) · [email protected] ·

    Nvidia Nemotron 3.5 lightning and Nemo Switchyard https://blogs.nvidia.com/blog/nemotron-lightning-switchyard-rtx-dgx/ # HackerNews # Tech # AI

    Nvidia Nemotron 3.5 lightning and Nemo Switchyard https://blogs.nvidia.com/blog/nemotron-lightning-switchyard-rtx-dgx/ # HackerNews # Tech # AI

  24. Mastodon — mastodon.social TIER_1 Deutsch(DE) · aisyndicate ·

    Nvidia releases Nemotron 3.5 Lightning: Open-weights model with only 3.6 billion active parameters reaches the level of OpenAI's four in the Intelligence Index

    Nvidia veröffentlicht Nemotron 3.5 Lightning: Open-Weights-Modell mit nur 3,6 Mrd. aktiven Parametern erreicht im Intelligence Index das Niveau von OpenAIs viermal größerem gpt-oss-120b. Mit knapp 670 Tokens/Sekunde ist es zugleich das schnellste Modell im Vergleich – relevant fü…

  25. Mastodon — mastodon.social TIER_1 Polski(PL) · aisight ·

    Nvidia unveiled Nemotron 3.5 Lightning – a compact open-weights model that offers record inference speeds, performing tasks up to 30% faster

    Nvidia zaprezentowała Nemotron 3.5 Lightning – kompaktowy model typu open-weights, który oferuje rekordową prędkość wnioskowania, wykonując zadania nawet o 30% szybciej niż konkurencja. # si # ai # sztucznainteligencja # wiadomości # informacje # technologia https:// aisight.pl/a…

  26. Mastodon — mastodon.social TIER_1 English(EN) · sipirtu ·

    Nvidia launches Nemotron 3.5 Lightning and NeMo Switchyard to help enterprises choose AI models by purpose rather than raw power Source: SiliconANGLE https:// s

    Nvidia launches Nemotron 3.5 Lightning and NeMo Switchyard to help enterprises choose AI models by purpose rather than raw power Source: SiliconANGLE https:// siliconangle.com/2026/08/11/nv idia-releases-nemotron-3-5-lightning-nemo-switchyard-give-enterprise-ai-capability-options…