DeepSeek's R1 model, released in January 2025, has achieved notable scores on MMLU-Pro and GPQA benchmarks, reaching 84.4% and 70.8% respectively. However, the model's cost-effectiveness is highlighted, with an "intelligence points per dollar" metric of 4.6, suggesting it may be a significant factor for users. AI
IMPACT This model's performance and cost metrics provide valuable data for AI developers and researchers evaluating model efficiency.
RANK_REASON The item reports on benchmark scores for a specific AI model, which falls under research. [lever_c_demoted from research: ic=1 ai=1.0]
Read on Mastodon — mastodon.social →
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →