A comparison of three new large language models—Google's Gemini 3.8 Flash, Anthropic's Claude Fable 5.1, and OpenAI's GPT-5.6 Sol—reveals significant price differences with comparable performance on an independent benchmark. Gemini 3.8 Flash, priced at a fraction of the cost of the other two, achieved an equal score on the Artificial Analysis Intelligence Index. However, Gemini 3.8 Flash consumes more tokens per task, potentially increasing overall costs despite its lower per-token price. The models also differ in their maximum token output, with Gemini 3.8 Flash having half the capacity of the others, which could impact performance on tasks involving long documents. AI
IMPACT Gemini 3.8 Flash offers a compelling cost-performance ratio, potentially influencing developer API choices and driving competition in the LLM market.
RANK_REASON Comparison of multiple LLM models on benchmarks and pricing.
- Anthropic
- Artificial Analysis
- Claude Fable 5.1
- Claude Opus 5
- Claude Sonnet 5
- Gemini 3.7 Flash
- Gemini 3.8 Flash
- GPT-5.6 Sol
- GPT-5.6 Terra
- Grok 4.6
- Muse Spark 1.2
- OpenAI
AI-generated summary · Google Gemini · from 2 sources. How we write summaries →