Google has updated its Gemini 3.8 Flash model, introducing three distinct thinking levels: low, medium, and high. These levels control the amount of internal reasoning the model performs before responding, impacting latency, token usage, and cost. Unlike its predecessor, Gemini 3.7 Flash, the 3.8 Flash model defaults to the medium setting and no longer supports the 'minimal' thinking level, which has been mapped to 'low' for migration purposes. Developers are advised to explicitly set the thinking level for each use case to manage costs and performance effectively. AI
IMPACT New thinking level controls in Gemini 3.8 Flash offer developers more granular management of cost and performance for different tasks.
RANK_REASON Model release from a frontier lab (Google). [lever_c_demoted from frontier_release: ic=1 ai=1.0]
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →