PulseAugur
EN
LIVE 10:36:39

Thinking Machines releases Inkling-Small, outperforming larger predecessor

Thinking Machines Lab has launched Inkling-Small, a new open-weights multimodal model that prioritizes efficiency over sheer size. Despite being significantly smaller than its predecessor, Inkling, Inkling-Small demonstrates superior performance on various coding and reasoning benchmarks. The model is designed for easier deployment, capable of running on a single GPU, and is available under the Apache 2.0 license. AI

IMPACT This release signals a potential shift towards more efficient, deployable models that can still achieve state-of-the-art performance, lowering barriers for adoption.

RANK_REASON Frontier-lab model release with system card and benchmark results.

Read on The Decoder →

AI-generated summary · Google Gemini · from 4 sources. How we write summaries →

Thinking Machines releases Inkling-Small, outperforming larger predecessor

COVERAGE [4]

  1. The Decoder TIER_1 English(EN) · Matthias Bastian ·

    Thinking Machines bets on efficiency over size with its second model, Inkling Small

    <p><img alt="" class="attachment-full size-full wp-post-image" height="801" src="https://the-decoder.com/wp-content/uploads/2026/07/thinking_machines_inkling_logo.png" style="height: auto; margin-bottom: 10px;" width="1256" /></p> <p> Thinking Machines, the AI lab from former Ope…

  2. MarkTechPost TIER_1 English(EN) · Asif Razzaq ·

    Thinking Machines Lab Releases Inkling-Small: A 276B Total, 12B Active Open Weights Multimodal MoE Model

    <p>Inkling-Small matches Inkling at a quarter the size, and its NVFP4 checkpoint runs on one NVIDIA B300 GPU</p> <p>The post <a href="https://www.marktechpost.com/2026/08/02/thinking-machines-lab-releases-inkling-small-276b-open-weights-multimodal-moe-model/">Thinking Machines La…

  3. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    Thinking Machines Lab has released Inkling-Small, an open-weights multimodal MoE model with 276B total and 12B active parameters. The model beats its larger sib

    Thinking Machines Lab has released Inkling-Small, an open-weights multimodal MoE model with 276B total and 12B active parameters. The model beats its larger sibling on coding and reasoning benchmarks while running on a single NVIDIA B300 GPU. Available under Apache 2.0 on Hugging…

  4. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    Thinking Machines releases Inkling Small, a smaller open-weights model that outperforms its predecessor on coding and reasoning tasks. Efficiency gains over siz

    Thinking Machines releases Inkling Small, a smaller open-weights model that outperforms its predecessor on coding and reasoning tasks. Efficiency gains over size may signal a practical shift in model development. Source: The Decoder AI https:// the-decoder.com/thinking-machi nes-…