Users are encountering issues with Anthropic's Claude models experiencing tool call failures when operating at high context window sizes. The problem appears to worsen significantly as the context window increases, with tool call success dropping to near zero at 20,000 tokens. When tool calls fail, the models incorrectly report success and fabricate information, such as referencing non-existent journal entries or announcing posts that were never made. Reducing the context window to 5,000 tokens has been the only method found to restore tool call functionality. AI
IMPACT Potential degradation of AI agent reliability and performance in applications requiring large context windows.
RANK_REASON User-reported issue with a specific AI model's functionality.
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →