StepFun's new model, Step 3.7 Flash, has achieved top rankings on the Artificial Analysis (AA) benchmark, excelling in speed, cost-efficiency, and end-to-end performance. The model demonstrates impressive output speeds of up to 416 tokens/s and significantly reduced costs, reportedly about one-ninth of Claude Opus 4.6's cost for similar programming capabilities. This efficiency focus aligns with the industry's shift towards practical applications in enterprise agents, where high-frequency, cost-effective model interactions are crucial for complex task completion. AI
IMPACT Sets new SOTA on speed and cost-efficiency benchmarks, pressuring competitors and accelerating enterprise agent adoption.
RANK_REASON New model release from a frontier lab with benchmark performance claims. [lever_c_demoted from frontier_release: ic=2 ai=1.0]
- Anthropic
- Blender
- Claude Opus 4.6
- DeepSeek
- GPT-5.3
- OpenRouter
- Step 3.7 Flash
- Step (Jieyue)
- Artificial Analysis (AA) benchmark
- OpenAI
- StepFun
AI-generated summary · Google Gemini · from 2 sources. How we write summaries →