Qwen3.5-0.8B
PulseAugur coverage of Qwen3.5-0.8B — every cluster mentioning Qwen3.5-0.8B across labs, papers, and developer communities, ranked by signal.
- 2026-06-27 research_milestone Spectral Labs developed a new quantization method, SpectralQuant, which significantly improves the performance of the Qwen3.5-0.8B model. source
-
OvisOCR2: A new 0.8B OCR model converts documents to structured Markdown
OvisOCR2 is a new 0.8B parameter end-to-end OCR model that converts full document pages into structured Markdown. Based on Qwen3.5-0.8B, it reportedly achieves high scores on document understanding benchmarks like OmniD…
-
Hugging Face releases OvisOCR2 for end-to-end document parsing
Hugging Face has released OvisOCR2, a new 0.8B parameter model designed for end-to-end document parsing. This model can take an image of a document page and output a Markdown representation that includes text, formulas,…
-
Voodoo Quant technique shows 95% KLD improvement over Unsloth Dynamic
A new quantization technique called Voodoo Quant has demonstrated significant improvements in model optimization, outperforming Unsloth Dynamic 2.0 KLD by 95% on Qwen3.5 models. Voodoo Quant optimizes each tensor indivi…
-
VultronRetriever models debut on Hugging Face, topping leaderboards
The VultronRetriever family of models has been released on Hugging Face, offering top performance in their respective classes on the MTEB Leaderboard. VultronRetrieverPrime-8B leads the pack, boasting a significantly sm…
-
Regolo.ai's Brick LLM router optimizes costs by selecting the best model
Regolo.ai has developed Brick, an LLM routing system designed to optimize costs by intelligently selecting the most appropriate model for a given prompt. Unlike traditional cascade systems that involve retries, Brick an…
-
Gepard 1.0 open-sourced for real-time streaming TTS
A new open-source text-to-speech (TTS) model named Gepard 1.0 has been released, designed for real-time conversational dialogue. Built with a focus on streaming, it begins generating audio as text arrives, achieving a t…
-
Modular's MAX models now run on Apple silicon GPUs
Modular has announced that its MAX models can now run on Apple silicon GPUs, including M1 through M5 chips. This update allows for the execution of various text, vision, and image diffusion models directly on Mac device…
-
Liquid AI ships tiny LFM2.5-230M for on-device agent tasks
Liquid AI has released LFM2.5-230M, its smallest model to date, designed for on-device inference on edge hardware like phones and robots. This 230-million-parameter model excels at data extraction and tool use, outperfo…
-
SpectralQuant method recovers 96.5% of BF16 performance gap in Qwen3.5 model
Spectral Labs has developed a new quantization method called SpectralQuant, which aims to improve the performance of smaller model footprints. Their initial release, a Qwen3.5 0.8B model quantized to Q4_K_M, reportedly …
-
Developer launches free RAG API for local LLMs to access medical facts
A developer has created a free Retrieval-Augmented Generation (RAG) API that provides local large language models (LLMs) with access to medical facts from Wikipedia. The API, accessible at hyfl.uk, aims for sub-second r…
-
Laptop GPU runs Qwen3.6 model with surprising speculative decoding boost
A user detailed their experience running the Qwen3.6-35B-A3B model on a laptop with an 8GB RTX 4060 GPU. They found that disabling memory mapping (`--no-mmap`), ensuring sufficient VRAM headroom, and closing CPU-intensi…
-
OpenBMB releases MiniCPM5-1B, a 1B parameter model outperforming larger rivals
OpenBMB has released MiniCPM5-1B, a small language model with one billion parameters that demonstrates performance comparable to larger models. This model is designed to run locally, accelerating the practical applicati…
-
Researchers explore optimal LoRA placement in hybrid language models
A new paper explores the optimal placement of LoRA adapters in hybrid language models, which combine attention and recurrent components. The research demonstrates that adapting the attention pathway is more effective th…