DeepSeek V4 Flash 0420 (Max) has achieved notable scores on several benchmarks, including 89.4% on GPQA and 45.3% on SciCode. The model also demonstrated proficiency in long-context reasoning with a score of 74.3%. These results position DeepSeek V4 Flash 0420 as a strong performer in the open-source AI landscape, with 144 intelligence points per dollar. AI
IMPACT Sets new performance benchmarks for open-source models, potentially influencing future development and adoption.
RANK_REASON The item reports on benchmark scores for an AI model, which falls under research. [lever_c_demoted from research: ic=1 ai=1.0]
Read on Mastodon — mastodon.social →
- DeepSeek V4 Flash 0420
- GPQA: A Graduate-Level Google-Proof Q&A Benchmark
- Humanity's Last Exam
- long-context reasoning
- SciCode
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →