PulseAugur
EN
LIVE 01:28:23
ENTITY LLM-Compressor

LLM-Compressor

PulseAugur coverage of LLM-Compressor — every cluster mentioning LLM-Compressor across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
2
3 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
0
1 over 90d
TIER MIX · 90D
TOPICS
SENTIMENT · 30D

2 day(s) with sentiment data

RECENT · PAGE 1/1 · 3 TOTAL
  1. TOOL · CL_176600 ·

    Tiny Kimi-K3 Model Released for Testing on Hugging Face

    A smaller, 0.4 billion parameter version of the moonshotai/Kimi-K3 model has been released on Hugging Face for testing and development purposes. This tiny model, named inference-optimization/Kimi-K3-0.40B, retains key a…

  2. TOOL · CL_142327 ·

    MedGemma-1.5-4B quantized to INT4 using llm-compressor

    A technical guide details the process of quantizing Google's MedGemma-1.5-4B medical vision-language model to INT4 (W4A16) using the llm-compressor library. The author encountered and resolved several issues, including …

  3. RESEARCH · CL_09107 ·

    Stateful Transformers boost streaming inference; Intel releases AutoRound quantization toolkit

    A new paper introduces a stateful transformer inference engine that significantly speeds up processing for streaming data by maintaining a persistent KV cache. This approach allows for query latency that is independent …