PulseAugur
EN
LIVE 07:19:29
ENTITY BITNET

BITNET

PulseAugur coverage of BITNET — every cluster mentioning BITNET across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
9
18 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
1
5 over 90d
TIER MIX · 90D
TOPICS
TIMELINE
  1. 2026-08-12 product_launch Microsoft open-sourced the BitNet inference framework for 1-bit LLMs. source
SENTIMENT · 30D

7 day(s) with sentiment data

RECENT · PAGE 1/1 · 18 TOTAL
  1. TOOL · CL_214737 ·

    BitNet explores ternary weights to cut LLM memory by 10x

    Researchers are exploring methods to reduce the computational and memory demands of large language models, moving beyond simply increasing model size. One promising approach, BitNet, investigates training models where w…

  2. TOOL · CL_208349 ·

    New low-bit and ternary AI models released with performance updates · 1 source tracked

    Several new low-bit and ternary models have been released and are being tracked, including Bonsai's 1-bit and 1.58-bit (ternary) versions, with a 27B parameter model now running on mainline backends. Updates to llama.cp…

  3. TOOL · CL_202946 ·

    Ternary LLMs see resurgence with new models from smaller labs

    Recent developments suggest a resurgence in ternary (1.58-bit) large language models, with several new models released by smaller labs. These include prismML's 27B ternary model, Deepgrove's 20B Maple model, and Doses A…

  4. TOOL · CL_196524 ·

    Microsoft open-sources BitNet for 1-bit LLMs on single CPUs

    Microsoft has open-sourced BitNet, an inference framework designed for 1-bit Large Language Models (LLMs). This framework allows for the execution of models with up to 100 billion parameters on a single CPU, eliminating…

  5. TOOL · CL_189876 ·

    Developer builds zero-dependency C inference engine for BitNet models

    A developer has created a C inference engine designed for BitNet models, focusing on zero dependencies and CPU performance. The engine achieves 36.25 tokens per second on a BitNet b1.58-2B-4T model running on an Intel X…

  6. COMMENTARY · CL_186836 ·

    Author's early coding journey parallels current AI displacement fears

    The author recounts their early experiences with programming, starting with AppleScript on Mac OS 7 at age seven and progressing to developing their own compiled programs by age eight. They detail their personal history…

  7. TOOL · CL_174617 ·

    Ternary Quantization Shrinks Super-Resolution Transformer to 668 KB for Browser Use

    Researchers have successfully applied BitNet-style ternary quantization to a super-resolution transformer model, resulting in a significantly smaller model size. The quantized model, which uses weights of -1, 0, or +1, …

  8. SIGNIFICANT · CL_160855 ·

    Microsoft releases VibeVoice-ASR-BitNet for real-time CPU inference

    Microsoft has released VibeVoice-ASR-BitNet, a highly compressed automatic speech recognition model designed for real-time inference on edge CPUs without requiring a GPU. This model achieves significant speedups over ex…

  9. TOOL · CL_157942 ·

    New C99 inference engine Project Zero offers 1.8x speedup for BitNet models

    A new, from-scratch C99 inference engine called Project Zero has been developed, offering a 1.8x speedup over bitnet.cpp for BitNet models on Xeon processors. This engine boasts zero external dependencies, running solel…

  10. TOOL · CL_149071 ·

    Negative-Bit Quantization Frees VRAM by Inverting Tensor Embeddings

    A researcher has developed a novel technique called Negative-Bit Quantization (NBQ) that claims to achieve stable inference with "negative-bit" configurations, effectively freeing up VRAM. This method, termed Phase-Inve…

  11. SIGNIFICANT · CL_143192 ·

    PrismML releases Bonsai 27B, enabling Qwen3.6-27B on laptops and phones

    PrismML has released Bonsai 27B, a highly compressed version of Qwen3.6-27B, available in 1-bit and ternary variants. These models are designed to run on consumer hardware like laptops and phones, with the 1-bit version…

  12. RESEARCH · CL_109464 ·

    New BITEMBED framework drastically cuts LLM embedding costs

    Researchers have developed BITEMBED, a novel framework designed to create efficient text embeddings for large language models. This approach converts LLM backbones into low-bit encoders using ternary weights and quantiz…

  13. RESEARCH · CL_105121 ·

    New Quantization Methods Boost LLM Efficiency and Speed

    Researchers have developed CAT-Q, a novel post-training quantization method that significantly compresses and accelerates Large Language Models (LLMs) without requiring extensive retraining. This technique, which uses l…

  14. TOOL · CL_105045 ·

    New memristive synapse design promises highly efficient on-chip neural networks

    Researchers have developed a physics-based design for an on-chip neural network utilizing multi-level memristive synapses capable of supporting a wide range of conductance states. This design, rooted in ionic transport …

  15. COMMENTARY · CL_78658 ·

    Ternary LLMs stall at 2B parameters, frontier labs bypass approach

    Ternary LLMs, which use a three-value system for weights, showed early promise but have not seen significant development. The largest available ternary model is only 2 billion parameters, and major AI labs have not adop…

  16. TOOL · CL_71783 ·

    Rust engine achieves 150+ TPS for 1-bit LLMs on edge CPUs

    A developer has created a novel inference engine for 1-bit quantized Large Language Models (LLMs) entirely in Rust, bypassing traditional frameworks like PyTorch and CUDA. This engine achieves impressive performance, de…

  17. TOOL · CL_60288 ·

    AWS trains Azerbaijani models, Tether simplifies LLM fine-tuning

    AWS is developing language models for the Azerbaijani language on its SageMaker platform, aiming to increase AI accessibility for underrepresented languages. Tether has released Bitnet, a framework to simplify the fine-…

  18. RESEARCH · CL_17600 ·

    1-Bit AI Infrastructure enables faster, lossless LLM inference on CPUs

    Researchers have developed a software stack called 'this http URL' to enable fast and lossless inference of 1-bit Large Language Models (LLMs) like BitNet b1.58 on CPUs. This new infrastructure achieves significant spee…