A user attempting to reproduce a previous claim of 65.5 billion tokens used with Claude Code found a significant discrepancy in token counting. By analyzing a smaller subset of data, they discovered that deduplicating messages by ID resulted in approximately half the token count compared to summing all rows. This suggests that the original claim might have been inflated due to not properly handling duplicate message IDs, with cache reads accounting for a large portion of the token volume. AI
IMPACT Highlights potential issues in token usage reporting for LLMs, impacting cost calculations and performance assessments.
RANK_REASON User analysis of a specific product's usage data questioning previous claims.
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →