Stepfun AI has released Step 3.7 Flash, a 198-billion parameter sparse Mixture-of-Experts (MoE) vision-language model. This model is optimized for agentic workflows, coding, and multimodal tasks, activating approximately 11 billion parameters per token for high throughput. It supports a 256k context window and offers selectable reasoning levels to balance speed and depth, with benchmarks showing competitive performance against models like DeepSeek V4 Flash and Gemini 3.5 Flash on coding and search tasks. AI
IMPACT Accelerates development of agentic workflows and multimodal applications with a high-performance, locally runnable model.
RANK_REASON Model release from a frontier lab (Stepfun AI) with detailed technical specifications and benchmark comparisons.
Read on Hugging Face Trending Models →
- DeepSeek V4 Flash
- Gemini 3.5 Flash
- GPT-4.5
- Mixture-of-Experts
- NVIDIA NIM
- OpenRouter
- Step 3.7 Flash
- Hugging Face
- llama.cpp
- llama-cpp-python
- SGLang
- Transformers
- Unsloth Studio
- vLLM
AI-generated summary · Google Gemini · from 5 sources. How we write summaries →