PulseAugur
EN
LIVE 12:35:37

GLM-4.7-Flash model shows performance metrics across benchmarks

The GLM-4.7-Flash model has demonstrated specific performance metrics across several benchmarks, including GPQA, Humanity's Last Exam, Long Context Reasoning, and SciCode. The model achieved 45.2% on GPQA and 25.5% on SciCode. Additionally, it processed 179.6 tokens per second and achieved 101.3 intelligence points per dollar, indicating its efficiency. AI

IMPACT Provides specific performance data for the GLM-4.7-Flash model across key benchmarks, useful for comparative analysis.

RANK_REASON The item reports benchmark results for an AI model, which falls under research. [lever_c_demoted from research: ic=1 ai=1.0]

Read on Mastodon — mastodon.social →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

GLM-4.7-Flash model shows performance metrics across benchmarks

COVERAGE [1]

  1. Mastodon — mastodon.social TIER_1 English(EN) · opensourceaitech ·

    📊 GLM-4.7-Flash (Non-reasoning) — the actual numbers GPQA: 45.2% Humanity's Last Exam: 4.9% Long Context Reasoning: 14.7% SciCode: 25.5% ⚡ 179.6 tokens/sec 💰 10

    📊 GLM-4.7-Flash (Non-reasoning) — the actual numbers GPQA: 45.2% Humanity's Last Exam: 4.9% Long Context Reasoning: 14.7% SciCode: 25.5% ⚡ 179.6 tokens/sec 💰 101.3 intelligence points per dollar Measured independently, not self-reported → https:// opensourceai.tech/leaderboard. h…