PulseAugur
EN
LIVE 18:55:20

OpenAI's Codex model context cap tied to cache costs, not billing

OpenAI has stated that the 272,000 token context window limit for its Codex model is due to cache-read costs. This limit is significantly lower than the model's published specification of 1,050,000 tokens. The API also reprices after the 272,000 token threshold. AI

IMPACT This information is relevant for developers using OpenAI's Codex model, highlighting potential cost considerations for extended context usage.

RANK_REASON The item discusses a specific model's context window limitation and its pricing implications, which falls under AI tooling rather than a frontier release.

Read on Mastodon — fosstodon.org →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

OpenAI's Codex model context cap tied to cache costs, not billing

COVERAGE [1]

  1. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    🤖 OpenAI's stated reason for Codex's 272k context cap is cache-read cost, not the 2x billing line at the same number codex ships with a model catalog, and its g

    🤖 OpenAI's stated reason for Codex's 272k context cap is cache-read cost, not the 2x billing line at the same number codex ships with a model catalog, and its gpt-5.6 entry lists the context window as 272,000 tokens. the published spec for the model is 1,050,000. 272,000 is also …