ENTITY
MMVQ
MMVQ
PulseAugur coverage of MMVQ — every cluster mentioning MMVQ across labs, papers, and developer communities, ranked by signal.
Total · 30d
2
2 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
0
0 over 90d
TIER MIX · 90D
TOPICS
SENTIMENT · 30D
1 day(s) with sentiment data
RECENT · PAGE 1/1 · 2 TOTAL
-
llama.cpp adds MMVQ optimization for MoE models like Qwen 35B
A pull request to the llama.cpp project introduces optimizations for Mixture of Experts (MoE) models, specifically targeting architectures like Qwen 35B A3B. This enhancement, named MMVQ, aims to improve the speed of th…
-
llama.cpp update optimizes CUDA performance on DGX Spark
The llama.cpp project has released an update, b10481, which includes optimizations for CUDA and dense models running on DGX Spark. The release introduces changes related to MMVQ (Multi-Query Vector Quantization) with sp…