OpenAI has released GPT-4o mini, a new, cost-effective LLM designed to significantly reduce the price of production applications. This model offers a substantial cost reduction compared to its predecessor, GPT-4o, with input tokens priced at $0.15 per million and output tokens at $0.60 per million. Despite its lower cost, GPT-4o mini demonstrates impressive performance, outperforming competitors like Gemini-1.5-Flash and Claude (Haiku) on benchmarks such as MMLU. It is well-suited for a wide range of tasks including classification, extraction, summarization, and chatbots, offering a compelling balance of capability and affordability for developers. AI
IMPACT Significantly lowers inference costs for AI applications, enabling wider adoption of advanced LLM capabilities in production environments.
RANK_REASON Frontier-lab model release with system card. [lever_c_demoted from frontier_release: ic=1 ai=1.0]
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →