ENTITY
TriAttention
TriAttention
PulseAugur coverage of TriAttention — every cluster mentioning TriAttention across labs, papers, and developer communities, ranked by signal.
Total · 30d
0
1 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
0
1 over 90d
TIER MIX · 90D
TOPICS
RECENT · PAGE 1/1 · 2 TOTAL
-
TriAttention method optimized for LLM KV cache reduction
A new paper introduces an optimization for the TriAttention method, designed to reduce the computational cost of shrinking the KV cache in long-context Large Language Models. The proposed method simplifies the scoring o…
-
New KV-cache compression method alpha outperforms existing techniques
Researchers have developed a new KV-cache compression method called alpha, which uses a diversity-penalty survivor approach. This method was found to outperform seven other mechanisms in a design-space study on mathemat…