The Granite 4.0 1B model has demonstrated performance metrics across several benchmarks, including GPQA at 28.1%, MMLU-Pro at 32.5%, HLE at 4.8%, and Long Context at 6%. These results were independently measured rather than self-reported, offering a transparent view of the model's capabilities for its size. AI
IMPACT Provides independent performance data for a 1B parameter model, aiding in comparative analysis for developers.
RANK_REASON The cluster reports benchmark results for an open-source model, which falls under research. [lever_c_demoted from research: ic=1 ai=1.0]
Read on Mastodon — mastodon.social →
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →