PulseAugur
EN
LIVE 09:48:33
Čeština(CS) Společnost DeepSeek zveřejnila 31. července 2026 model V4 Flash 0731 ve třech úrovních „reasoning“ nastavení (Low, High, Max), a výsledky na benchmarku ARC Priz

DeepSeek V4 Flash 0731 achieves high scores on ARC-AGI benchmark

DeepSeek has released its V4 Flash 0731 model, featuring three reasoning settings: Low, High, and Max. The model achieved impressive scores on the ARC Prize benchmark, a test for abstract reasoning in AI systems. Notably, in its Max setting, V4 Flash 0731 scored 89.0% on ARC-AGI-1 at a cost of $0.02 per task and 61.4% on the more challenging ARC-AGI-2 at $0.04 per task. This release highlights DeepSeek's continued focus on providing competitive performance at a fraction of the cost of Western models, with the model and its associated paper available on Hugging Face and arXiv. AI

IMPACT Sets new SOTA on abstract reasoning benchmarks, offering a cost-effective alternative to Western models.

RANK_REASON Frontier lab model release with benchmark results. [lever_c_demoted from frontier_release: ic=1 ai=1.0]

Read on Mastodon — fosstodon.org →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

DeepSeek V4 Flash 0731 achieves high scores on ARC-AGI benchmark

COVERAGE [1]

  1. Mastodon — fosstodon.org TIER_1 Čeština(CS) · [email protected] ·

    DeepSeek released the V4 Flash 0731 model on July 31, 2026, in three 'reasoning' settings (Low, High, Max), and results on the ARC Priz benchmark

    Společnost DeepSeek zveřejnila 31. července 2026 model V4 Flash 0731 ve třech úrovních „reasoning“ nastavení (Low, High, Max), a výsledky na benchmarku ARC Prize, který testuje schopnost obecné abstraktní usuzování u AI systémů, jsou působivé. Klíčová čísla: Na verzi ARC-AGI-1 do…