ENTITY
KV state
KV state
PulseAugur coverage of KV state — every cluster mentioning KV state across labs, papers, and developer communities, ranked by signal.
Total · 30d
1
2 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
0
1 over 90d
TIER MIX · 90D
TOPICS
SENTIMENT · 30D
1 day(s) with sentiment data
RECENT · PAGE 1/1 · 2 TOTAL
-
Prompt caching, not model routing, drives LLM cost savings
The primary driver of cost savings in large language models (LLMs) is prompt caching, rather than automatic model routing. Prompt caching stores key-value (KV) state specific to each model, meaning switching models resu…
-
New research questions reliability of language model credit estimation methods
A new paper from arXiv explores the reliability of counterfactual token-credit estimation in language models. The research highlights that re-feeding the transcript prefix as a fresh prompt, a common method, can introdu…