OpenAI has introduced a new service tier called "Ultrafast" for its GPT-6 Astra model, offering significantly faster token generation at a six-fold price increase. This premium speed tier is available to all API users, with higher limits for enterprise clients. While the model itself remains the same, the Ultrafast tier aims to reduce latency for time-sensitive applications, though it is not recommended for batch processing. OpenAI also previewed Ultrafast for GPT-5.6 Sol and plans to support GPT-6.1 Sol, with pricing details for these models yet to be released. AI
IMPACT This premium speed tier may accelerate adoption for latency-sensitive AI applications, but its high cost requires careful cost-benefit analysis.
RANK_REASON This is a new service tier for an existing model, not a new model release.
AI-generated summary · Google Gemini · from 3 sources. How we write summaries →