PulseAugur
EN
LIVE 00:42:02

Zaya1-8B model beats GPT-5-High on math test without NVIDIA GPUs

A new language model named Zaya1-8B, featuring 760 million active parameters in a Mixture-of-Experts architecture, has demonstrated impressive performance on the HMMT '25 math competition. Notably, this model achieved its results without any training on NVIDIA GPUs, a significant departure from typical high-performance AI training. Zaya1-8B surpassed the performance of GPT-5-High on this specific math benchmark, scoring 89.6%. AI

IMPACT Demonstrates novel training approaches can yield competitive results, potentially reducing reliance on expensive GPU infrastructure.

RANK_REASON The cluster reports on a new model's performance on a specific benchmark, which falls under research. [lever_c_demoted from research: ic=1 ai=1.0]

Read on Towards AI →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

Zaya1-8B model beats GPT-5-High on math test without NVIDIA GPUs

How we ranked this

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Tool
The cluster reports on a new model's performance on a specific benchmark, which falls under research. [lever_c_demoted from research: ic=1 ai=1.0]
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
model release, paper
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
130 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.

Full methodology in our editorial standards.

COVERAGE [1]

  1. Towards AI TIER_1 English(EN) · Chew Loong Nian - AI ENGINEER ·

    I Tested ZAYA1-8B — Trained on Zero NVIDIA GPUs, Its 760M Active Params Cheated GPT-5-High on Math

    <div class="medium-feed-item"><p class="medium-feed-snippet">A 760-million-active-parameter MoE that never touched a single NVIDIA H100 in training scored 89.6% on HMMT &#x2019;25 math &#x2014; 1.3 points higher&#x2026;</p><p class="medium-feed-link"><a href="https://pub.towardsa…