Users of OpenAI's Responses API are encountering unexpectedly high costs due to the accumulation of context from previous steps in long-running workflows. While the model's generated output might be short, the input tokens, which include past tool outputs and context, can grow significantly over multiple steps. This leads to variable and often higher token usage for successful runs, prompting users to seek strategies for managing context and reducing costs without sacrificing necessary information. AI
IMPACT Highlights potential cost inefficiencies in complex API workflows, prompting developers to optimize context management.
RANK_REASON User discussion about API cost management for a specific product feature.
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →