A new real-time dashboard created by QuantID visualizes all 48 official Hugging Face benchmark leaderboards in a single view. The tool aggregates data directly from Hugging Face's public API, showing participation patterns and top rankings. Notably, the Korean AI startup VIDRAFT, operating under the FINAL-Bench organization, currently leads in 10 of the 48 benchmarks, surpassing all other organizations. AI
IMPACT Provides a consolidated view of AI model performance across numerous benchmarks, aiding in comparative analysis.
RANK_REASON The item describes a new visualization tool for existing leaderboards, not a novel model release or research breakthrough.
- Alibaba Qwen
- Artificial Intelligence In Medical Epidemiology
- DeepSeek
- FINAL-Bench
- GPQA: A Graduate-Level Google-Proof Q&A Benchmark
- Hugging Face
- MMLU-Pro
- Moonshot AI
- SWE-bench
- VIDRAFT
- Zhipu AI
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →