Several leading AI models have seen significant price reductions in their batch inference costs over the past month, with some dropping by as much as 67%. Models like DeepSeek V4.1 Flash, GPT-5.6 Sol Pro, Mistral Large 3 2512, and GLM 5.3 Flash are now considerably cheaper to use, with some also offering open weights. These price drops suggest a rapid decrease in the cost of running large language models, potentially driven by advancements in inference efficiency and hardware. AI
IMPACT Accelerates adoption of advanced LLMs by reducing operational costs and potentially enabling new use cases.
RANK_REASON Multiple AI models from different providers have seen significant price drops in batch inference costs, indicating a broader industry trend.
Read on Mastodon — mastodon.social →
AI-generated summary · Google Gemini · from 9 sources. How we write summaries →