ENTITY
HyQuant
HyQuant
PulseAugur coverage of HyQuant — every cluster mentioning HyQuant across labs, papers, and developer communities, ranked by signal.
Total · 30d
2
2 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
2
2 over 90d
TIER MIX · 90D
TOPICS
SENTIMENT · 30D
2 day(s) with sentiment data
RECENT · PAGE 1/1 · 2 TOTAL
-
HyQuant framework optimizes LLM attention with hybrid-precision quantization
Researchers have developed HyQuant, a novel hybrid-precision quantization framework designed to improve the efficiency of Large Language Model (LLM) attention mechanisms. This method quantizes most attention states to l…
-
New methods drastically shrink LLM size and boost inference speed
Researchers have developed two novel methods to significantly reduce the size and computational cost of large language models (LLMs) without substantial performance loss. Squeeze10-LLM employs a staged mixed-precision q…