PulseAugur
EN
LIVE 19:00:48

Cognition's SWE-2 coding model debuts with high benchmark scores

Cognition has released its new coding model, SWE-2, which boasts a massive 2.8 trillion parameters with 104 billion active per token using a Mixture of Experts (MoE) architecture. The model reportedly achieves a 92.8 score on the Terminal-Bench 2.1 benchmark, and while it shows strong performance on tasks like FrontierCode, it lags behind competitors like Claude Fable 5.1 and GPT-6 Astra on longer-horizon agentic tasks. SWE-2 is integrated into Cognition's Devin coding assistant products but is not available as a standalone API or for local deployment. AI

IMPACT Sets new SOTA on Terminal-Bench 2.1, but highlights remaining gaps in long-horizon agentic tasks.

RANK_REASON Frontier-lab model release with system card and benchmark data.

Read on Mastodon — mastodon.social →

AI-generated summary · Google Gemini · from 7 sources. How we write summaries →

Cognition's SWE-2 coding model debuts with high benchmark scores

How we ranked this

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Frontier Release
Frontier-lab model release with system card and benchmark data.
Source corroboration
7 independent sources
Strong cross-source corroboration — multiple independent publishers covered this within the clustering window.
Topics
model release, product
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
10 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.
Coverage growth since scoring
+3 source(s) since last score
New sources have picked up this story since our last re-score. Score will update on the next scoring pass.

Full methodology in our editorial standards.

COVERAGE [7]

  1. Hacker News — AI stories ≥50 points TIER_1 English(EN) · cdnsteve ·

    Cognition's SWE-2 achieves 92.8 on Terminal-Bench 2.1

  2. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    Cognition's SWE-2 achieves 92.8 on Terminal-Bench 2.1 https:// tokenstead.ai/models/swe-2 # ai

    Cognition's SWE-2 achieves 92.8 on Terminal-Bench 2.1 https:// tokenstead.ai/models/swe-2 # ai

  3. Mastodon — mastodon.social TIER_1 Polski(PL) · aisight ·

    Cognition's new SWE-2 model aims to match leading coding systems at a significantly lower cost, thanks to unique RL training. The tool remains je

    Nowy model SWE-2 od Cognition ma dorównywać czołowym systemom kodującym przy znacząco niższych kosztach, dzięki unikalnemu treningowi RL. Narzędzie pozostaje jednak produktem zamkniętym, dostępnym wyłącznie dla użytkowników agenta Devina. # si # ai # sztucznainteligencja # wiadom…

  4. Mastodon — mastodon.social TIER_1 English(EN) · [email protected] ·

    Cognition, the company behind the Devin coding agent, has released SWE-2, its most capable coding model to date. SWE-2 is post-trained with reinforcement learni

    Cognition, the company behind the Devin coding agent, has released SWE-2, its most capable coding model to date. SWE-2 is post-trained with reinforcement learning from Kimi K3, Moonshot AI’s 2.8T-parameter open model. The model scores 50.0% on FrontierCode 1.1 Main, within 1 poin…

  5. Mastodon — mastodon.social TIER_1 English(EN) · [email protected] ·

    Cognition releases SWE-2 coding model, lowering development costs by up to 64% through single-run reinforcement learning # AI # AINews

    Cognition releases SWE-2 coding model, lowering development costs by up to 64% through single-run reinforcement learning # AI # AINews

  6. Mastodon — mastodon.social TIER_1 English(EN) · [email protected] ·

    Cognition releases SWE-2 coding model, achieving high accuracy scores and supporting long-running tasks # AI # AINews

    Cognition releases SWE-2 coding model, achieving high accuracy scores and supporting long-running tasks # AI # AINews

  7. Mastodon — mastodon.social TIER_1 English(EN) · [email protected] ·

    Cognition's SWE-2 achieves 92.8 on Terminal-Bench 2.1 Article URL: https:// tokenstead.ai/models/swe-2 Comments URL: https:// news.ycombinator.com/item?id=4 964

    Cognition's SWE-2 achieves 92.8 on Terminal-Bench 2.1 Article URL: https:// tokenstead.ai/models/swe-2 Comments URL: https:// news.ycombinator.com/item?id=4 9646778 Points: 3 # Comments: 0 https:// tokenstead.ai/models/swe-2 # Tech # Technology # TechNews # AI # Gadgets # Softwar…