Qwen3.5:9b
PulseAugur coverage of Qwen3.5:9b — every cluster mentioning Qwen3.5:9b across labs, papers, and developer communities, ranked by signal.
13 day(s) with sentiment data
-
Visual Contrastive Self-Distillation Improves Qwen VL Models
Researchers have developed Visual Contrastive Self-Distillation (VCSD), a novel method for improving Vision-Language Models (VLMs) without requiring external teachers or privileged information. VCSD works by comparing a…
-
Nanbeige/Nanbeige4.2-3B model offers strong agentic capabilities at 3B scale
The Nanbeige/Nanbeige4.2-3B model is a 3 billion parameter language model designed for agentic tasks, boasting strong reasoning and alignment capabilities. It utilizes a Looped Transformer architecture to enhance capaci…
-
Research paper reveals FFNs actively steer long-context retrieval
A new research paper explores the role of Feed-Forward Networks (FFNs) in long-context retrieval tasks, moving beyond their traditional view as parametric memories. The study demonstrates that FFNs actively influence th…
-
AI systems advance legal question answering and translation capabilities · 4 sources tracked
Researchers have developed new AI systems to tackle complex legal tasks, including question answering and machine translation. One system, AILQA, is designed for the Indian legal system and uses retrieval-augmented gene…
-
New LLM safety research focuses on geometric constraints and trajectory-based patching
Two new research papers explore methods for enhancing Large Language Model (LLM) safety. The first paper, "Geometry-Guided Constraint Learning for LLM Safety Classification," introduces a technique that uses sparse auto…
-
Bonsai 27B models tested on Terminal-Bench 2.0, 1-bit version unusable
A user tested the Ternary-Bonsai-27B (2-bit) and Bonsai-27B (1-bit) models on the Terminal-Bench 2.0 benchmark, finding that the 2-bit version achieved a score of 7.9%. This performance was lower than the Qwen3.5-9B mod…
-
New ASCIITermDraw Bench tests VLM ability to generate ASCII diagrams
A new benchmark called ASCIITermDraw Bench has been introduced to evaluate the capabilities of Vision-Language Models (VLMs) in generating and editing ASCII art diagrams. Unlike benchmarks focusing on coding or reasonin…
-
Empero AI releases Qwythos-9B-v2, fixing looping with 1M-token context
Empero AI has released Qwythos-9B-v2, an updated version of its large language model designed to eliminate looping and degeneration issues that previously affected a small percentage of its outputs. This new version ach…
-
Amplitude Gating improves LLM structured output without retraining
Researchers have developed a new method called Amplitude Gating (AG) to improve the structured output of large language models during inference without retraining. This technique modulates activation magnitudes within f…
-
User creates 'Nikusui-v1' model by tweaking Qwen3.5-9B J-Space
A user on Reddit has created a new model called "Nikusui-v1" by modifying the J-Space of the Qwen3.5-9B model. This modification, inspired by Anthropic's Jacobian-Lens tool, allows for the alteration of a model's behavi…
-
LLM adoption in coding shows mixed results, increasing cruft and maintenance debt
Developers are increasingly using Large Language Models (LLMs) for coding tasks, but this adoption comes with mixed results and potential long-term consequences. While some developers find open-source models like Qwen3.…
-
LLM context benchmark: Prefill speed and KV cache matter most for agents
A benchmark of 13 different large language models tested at context lengths ranging from 65K to 128K tokens revealed that prompt processing (prefill) speed is the most critical factor for agentic workloads, rather than …
-
New AI frameworks tackle long-form video understanding with advanced memory and reasoning
Researchers are developing advanced frameworks to improve how AI models understand and reason about long-form videos. Homer, for instance, uses a hierarchical memory system that organizes information by temporal and cau…
-
PaperPilot agent uses workflow induction for advanced scientific literature search
Researchers have developed PaperPilot, a novel multi-turn agent designed for scientific literature search. Unlike traditional agents that use fixed pipelines or simple language reasoning, PaperPilot frames search as wor…
-
PaperPilot agent uses workflow induction for advanced scientific literature search · 2 sources tracked
Researchers have developed PaperPilot, a novel multi-turn literature search agent that frames scientific search as workflow induction. This agent constructs an executable Directed Acyclic Graph (DAG) of paper-search ope…
-
Run RAG agent offline with LangGraph, Ollama, and embedded Qdrant
This article details how to run a Retrieval-Augmented Generation (RAG) agent entirely offline using LangGraph, Ollama, and an embedded Qdrant vector store. The setup avoids the need for API keys by configuring the syste…
-
Sieve adds persistent memory to local Ollama LLMs
A new open-source tool called Sieve has been developed to add persistent memory to local Large Language Models (LLMs) running through Ollama. This tool acts as a proxy, sitting between the user's client and the Ollama e…
-
Empero AI releases Qwythos-9B reasoning model with 1M context window
The empero-ai/Qwythos-9B-Claude-Mythos-5-1M model, a 9B parameter reasoning model, has been released and is available on Hugging Face. This model is built upon Qwen3.5-9B and fine-tuned with Claude Mythos and Fable trac…
-
Research: Attention, not scale, drives human-AI alignment in vision-language models
Two new research papers explore the alignment between human attention and vision-language models. The first paper, focusing on multimodal language prediction, found that while adding visual context improved model-human …
-
Local LLM Benchmarks Show Qwen3.5:9B Strength Against Gemma4:26B
David Rodriguez conducted benchmarks of local large language models on a mid-tier gaming PC, finding that Qwen3.5:9B remains a powerful model, competitive even with Gemma4:26B. The analysis also highlighted a smaller te…