PulseAugur
EN
LIVE 16:54:20
Čeština(CS) Společnost DeepSeek zveřejnila 31. července 2026 model V4 Flash 0731 ve třech úrovních „reasoning“ nastavení (Low, High, Max), a výsledky na benchmarku ARC Priz

DeepSeek V4 Flash 0731 achieves high scores on ARC-AGI benchmark

DeepSeek has released its V4 Flash 0731 model, featuring three reasoning settings: Low, High, and Max. The model achieved impressive scores on the ARC Prize benchmark, a test for abstract reasoning in AI systems. Notably, in its Max setting, V4 Flash 0731 scored 89.0% on ARC-AGI-1 at a cost of $0.02 per task and 61.4% on the more challenging ARC-AGI-2 at $0.04 per task. This release highlights DeepSeek's continued focus on providing competitive performance at a fraction of the cost of Western models, with the model and its associated paper available on Hugging Face and arXiv. AI

IMPACT Sets new SOTA on abstract reasoning benchmarks, offering a cost-effective alternative to Western models.

RANK_REASON Frontier lab model release with benchmark results. [lever_c_demoted from frontier_release: ic=1 ai=1.0]

Read on Mastodon — fosstodon.org →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

DeepSeek V4 Flash 0731 achieves high scores on ARC-AGI benchmark

How we ranked this

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Significant
Frontier lab model release with benchmark results. [lever_c_demoted from frontier_release: ic=1 ai=1.0]
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
model release, paper, product
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
47 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.

Full methodology in our editorial standards.

COVERAGE [1]

  1. Mastodon — fosstodon.org TIER_1 Čeština(CS) · [email protected] ·

    DeepSeek released the V4 Flash 0731 model on July 31, 2026, in three 'reasoning' settings (Low, High, Max), and results on the ARC Priz benchmark

    Společnost DeepSeek zveřejnila 31. července 2026 model V4 Flash 0731 ve třech úrovních „reasoning“ nastavení (Low, High, Max), a výsledky na benchmarku ARC Prize, který testuje schopnost obecné abstraktní usuzování u AI systémů, jsou působivé. Klíčová čísla: Na verzi ARC-AGI-1 do…