DeepSeek-V4 Flash 0731 has achieved a score of 90.8% on the GPQA benchmark and 38.6% on HLE. The model also demonstrated a speed of 233.8 tokens per second. Notably, it offers a high intelligence-to-cost ratio, providing 52.3 intelligence points per dollar, positioning it as an efficient option in the LLM landscape. AI
IMPACT Demonstrates competitive performance and efficiency in LLM benchmarks.
RANK_REASON Model benchmark results published. [lever_c_demoted from research: ic=1 ai=1.0]
Read on Mastodon — fosstodon.org →
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →