Qwen 3.5
PulseAugur coverage of Qwen 3.5 — every cluster mentioning Qwen 3.5 across labs, papers, and developer communities, ranked by signal.
- instance of large-language models 95%
- instance of Qwen3.6 90%
- used by Nex-N2 Pro 90%
- used by Nex AGI 90%
- competes with Gemma 4 70%
- competes with Qwen3.6 70%
- used by vLLM 70%
- competes with Qwen-3.6 70%
- competes with Kimi k3 70%
- developed by Qwen-3.6 70%
- developed Qwen-3.6 70%
- competes with Claude Fable-5 70%
14 day(s) with sentiment data
-
Intel releases OpenVINO 2026.4 with expanded model support and performance upgrades
Intel has released OpenVINO 2026.4, an updated toolkit for optimizing and deploying AI inference. This release introduces support for a wide array of new models, including Gemma-3n, Qwen3-VL-4B, and Granite 4.0 H Micro,…
-
Open-Source vs. Proprietary LLMs: A Strategic Decision Framework · 3 sources tracked
The debate between open-source and proprietary Large Language Models (LLMs) is evolving, with open-source models increasingly closing the capability gap with their proprietary counterparts. While proprietary models like…
-
GLM 5.2 pricing surges 106% as Qwen 3.5 drops 48%
Pricing strategies for AI models have diverged significantly since August. GLM 5.2 has seen a 106% price increase, while Qwen 3.5 has experienced a 48% price decrease. These contrasting approaches highlight different ma…
-
llama.cpp bug causes non-deterministic results for M-RoPE embedding batches
A bug in the llama.cpp library causes incorrect results when processing embedding batches for M-RoPE models like Qwen3.5 and Qwen2.5-VL. The issue stems from a heap buffer overflow where the library reads past the alloc…
-
Voodoo Dynamic Quant method released under MIT license
A new dynamic quantization method called Voodoo Quant has been released under an MIT license, making it available to the open-source community. This method utilizes gradient descent to optimize the per-tensor quant layo…
-
New methods improve harmful meme detection in vision-language models
Researchers have developed new methods to improve the detection of harmful memes by vision-language models. One approach, "Decodable but Misrouted," uses sparse autoencoders and causal interventions to identify whether …
-
DFlash diffusion model fails to speed up Gemma LLM in tests
A new technique called DFlash aims to accelerate LLM generation by using a diffusion model, typically used for image generation, to predict multiple tokens simultaneously. Unlike other methods that focus on specific mod…
-
Together AI expands fine-tuning with new models and live tracking
Together AI has enhanced its fine-tuning service by incorporating a wider array of open-weight models, including advanced options like GLM 5.3 and Kimi K2.7, alongside cost-effective choices such as Qwen 3.8-27B and Gem…
-
New RAG method improves retrieval sufficiency, outperforming LLM judges
Researchers have developed a new method called Distribution-Shape QPP to improve retrieval sufficiency in Retrieval-Augmented Generation (RAG) pipelines. This approach aims to prevent hallucinations by providing a relia…
-
AI's next frontier: Mamba, JEPA, and Diffusion Models poised to replace transformers
The AI landscape is experiencing a cyclical shift, with transformers, dominant since 2017, potentially being replaced by newer architectures like state space models (Mamba) and Joint Embedding Predictive Architectures (…
-
Qwen 3.5 0.8B model optimized for CPU with new quantization format
A developer has created a custom C++ engine and a new 4-bit quantization format, H128/Q4-G32-DOT4, for the Qwen 3.5 0.8B model. This new format results in a smaller model size of 425 MB, which is 71 MB less than Unsloth…
-
Fields Medalist's startup bridges AI models, slashing costs and boosting performance
A startup named Mostik, founded by a team including a Fields Medal winner, has developed a novel method to improve AI model collaboration. Their approach bypasses traditional text-based communication between models, ins…
-
LiteLLM adds image signing, DeepSeek cuts prices, Qwen 3.5 launches new models
The AI news brief highlights several developments in the field. LiteLLM has introduced image signing capabilities, while DeepSeek has announced price reductions for its services. Additionally, Qwen 3.5 has been released…
-
New method analyzes AI model bias per conversation
Researchers have developed a new method called Counterfactual Resampling to analyze model behavior, specifically focusing on Value Leakage. This technique allows for the measurement of bias on a per-conversation basis, …
-
AI Engineering Tackles Unanswerable Questions: Batching, vLLM, and Context
This week's AI newsletter delves into complex engineering decisions that lack single correct answers, focusing on trade-offs in model performance, context, retrieval, and infrastructure. It highlights techniques like co…
-
OpenAI's GPT-6 and Mostik's tech signal shift away from token-based AI
OpenAI is reportedly developing a new architecture called "recurrent depth" for its upcoming GPT-6 model, which aims to improve reasoning by allowing the model to internally loop and refine its thoughts rather than gene…
-
Mostik enables AI models to communicate via internal weights, bypassing text
A startup named Mostik has developed a novel method for AI models to communicate by directly interacting with their internal mathematical weights, bypassing traditional text-based outputs. This technique allows smaller …
-
Qwen3.6 and Qwen3.5 show similar inference speeds, with gains in agentic tasks
A recent benchmark comparison of Qwen3.6 and Qwen3.5 models revealed that their inference speeds on a GeForce RTX 4070 were nearly identical, contrary to initial findings that suggested a significant slowdown. This disc…
-
New AI systems tackle scientific figure generation and editing · 2 sources tracked
Two new research papers introduce novel approaches to generating and editing scientific figures. The first, "Figures as Programs," proposes a multi-agent system called FigTree that recursively constructs figures as SVG …
-
LLM scale's impact on ontology learning studied across Qwen and GPT models
A new study published on arXiv investigates the impact of Large Language Model (LLM) scale on ontology learning performance. Researchers evaluated 13 models, including variants from the Qwen3.5 and Qwen3.6 lineages, usi…