PulseAugur
EN
LIVE 06:17:57
中文(ZH) DeepSeek 为什么把缓存涨了 11 倍?长上下文需求暴涨,得交点「存储税」了

DeepSeek hikes API prices up to 11x due to soaring long-context storage costs

DeepSeek has significantly increased API prices for its V4 Pro model, particularly for "cache hit" inputs, with some rates rising up to 11 times. This price hike reflects the growing costs associated with storing and moving data for large language models, which are becoming more expensive than computation itself. The increased demand for long context windows has led to "cache thrashing" on DeepSeek's servers during peak times, where data is frequently moved between GPU memory and slower storage, negating the benefits of caching and driving up operational expenses. While DeepSeek remains a cost-effective option compared to competitors like Anthropic, the price adjustment signals a shift towards developers needing to manage caching strategies more carefully to optimize costs. AI

IMPACT Signals a shift in AI infrastructure costs, requiring developers to manage caching strategies more actively.

RANK_REASON Price increase for a major AI model API due to infrastructure cost shifts. [lever_c_demoted from significant: ic=1 ai=1.0]

Read on 雷峰网 (Leiphone) →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

DeepSeek hikes API prices up to 11x due to soaring long-context storage costs

COVERAGE [1]

  1. 雷峰网 (Leiphone) TIER_1 中文(ZH) ·

    Why did DeepSeek increase cache by 11 times? The demand for long context has soared, and we have to pay a "storage tax".

    <section style="text-align: center; margin: 0px 16px; line-height: 1.75em; display: block;"><img class="rich_pages wxw-img" src="https://static.leiphone.com/uploads/new/images/20260818/6a83c0880a0f0.jpg?imageMogr2/quality/90" style="width: 546px; display: inline-block; text-align…