PulseAugur
EN
LIVE 13:25:59

Stepfun AI releases 198B parameter multimodal MoE model

Stepfun AI has released Step 3.7 Flash, a 198-billion parameter sparse Mixture-of-Experts (MoE) vision-language model. This model is optimized for agentic workflows, coding, and multimodal tasks, activating approximately 11 billion parameters per token for high throughput. It supports a 256k context window and offers selectable reasoning levels to balance speed and depth, with benchmarks showing competitive performance against models like DeepSeek V4 Flash and Gemini 3.5 Flash on coding and search tasks. AI

IMPACT Accelerates development of agentic workflows and multimodal applications with a high-performance, locally runnable model.

RANK_REASON Model release from a frontier lab (Stepfun AI) with detailed technical specifications and benchmark comparisons.

Read on Hugging Face Trending Models →

AI-generated summary · Google Gemini · from 5 sources. How we write summaries →

Stepfun AI releases 198B parameter multimodal MoE model

COVERAGE [5]

  1. Hugging Face Trending Models TIER_1 English(EN) · stepfun-ai ·

    stepfun-ai/Step-3.7-Flash-GGUF

    image-text-to-text · 476 downloads · 49 likes

  2. Hugging Face Trending Models TIER_1 Română(RO) · stepfun-ai ·

    stepfun-ai/Step-3.7-Flash

    image-text-to-text · 1,397 downloads · 52 likes

  3. 36氪 (36Kr) TIER_1 中文(ZH) ·

    Jieyue Release Step 3.7 Flash

    36氪获悉,据阶跃星辰消息,阶跃正式发布并开源Step 3.7 Flash。据介绍,Step 3.7 Flash是面向Agent生产化阶段推出的新一代Flash模型,围绕Agent、Coding、Search与多模态工作流进行系统优化。

  4. r/LocalLLaMA TIER_1 English(EN) · /u/-dysangel- ·

    Stepfun 3.7 Flash is very good

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1tss9nq/stepfun_37_flash_is_very_good/"> <img alt="Stepfun 3.7 Flash is very good" src="https://preview.redd.it/k37ol07vfg4h1.gif?width=640&amp;crop=smart&amp;s=b7802a3b13d13d1fee73f539427b4d2ff1c29628" title=…

  5. r/LocalLLaMA TIER_1 Deutsch(DE) · /u/Everlier ·

    StepFun 3.7 Flash

    <!-- SC_OFF --><div class="md"><p>StepFun dropped Step 3.7 Flash, 196B total / 11B active MoE, runs locally on 128GB RAM</p> <p>It's a multimodal MoE (196B total params, only 11B active) with a built-in 1.8B ViT for vision.</p> <p>Benchmark highlights vs. other flash-tier models:…