Qwen3.5-27B
PulseAugur coverage of Qwen3.5-27B — every cluster mentioning Qwen3.5-27B across labs, papers, and developer communities, ranked by signal.
7 day(s) with sentiment data
-
New method extracts interpretable circuits from dense transformers
Researchers have developed Sparse Weight Decomposition (SWD), a novel method for extracting interpretable circuits from dense pretrained transformer models. Unlike previous approaches that require additional training or…
-
New defense probes detect and mitigate indirect prompt injection in LLMs
Researchers have developed a method to detect indirect prompt injection (IPI) attacks in agentic large language models (LLMs). By training simple linear probes on the models' internal states, they can predict IPI exposu…
-
New framework generates 37,000 AI agent tasks for $0.05 each
Researchers have developed Recursive Synthetic Terminal Tasks (RST), a framework designed to generate long-horizon training data for terminal agents at a significantly reduced cost. This method recursively synthesizes n…
-
New AI safety architecture enhances mental health support models
Researchers have developed a novel safety architecture for generative AI models used in mental health support, addressing the limitations of current risk detection methods. This model-agnostic system integrates contextu…
-
New benchmark evaluates RAG for French immigration law
Researchers have developed a new benchmark and baseline study to evaluate Retrieval-Augmented Generation (RAG) systems for French immigration law. The study compares a parametric LLM baseline against RAG models at two s…
-
New ArbiGraph benchmark reveals context management flaws in AI agents
Researchers have developed ArbiGraph, a new benchmark generator designed to evaluate the context management capabilities of language agents that use tools. ArbiGraph creates complex, verifiable task graphs with varying …
-
AMD invests $5B in Anthropic; Microsoft partners with Mistral and fine-tunes Alibaba models · 3 sources tracked
Major AI developments are unfolding globally, with significant investments and strategic partnerships shaping the landscape. AMD has invested up to $5 billion in Anthropic, while Microsoft is expanding its partnership w…
-
Microsoft releases Fara1.5-27B multimodal agent for web automation
Microsoft has released Fara1.5-27B, a multimodal computer use agent designed for web browsers. This agent observes browser interfaces through screenshots and executes tasks by emitting structured tool calls like clicks …
-
Empero AI releases Qwythos-27B-v1 reasoning model on Hugging Face
Empero AI has released Qwythos-27B-v1, an open-weight, full-parameter reasoning model. This larger version of Qwythos-9B was trained on the same curriculum and built upon a Qwen3.5-27B base. The model is available on Hu…
-
Embodied AI research advances grounded world models and agent collaboration · 8 sources tracked
Recent research explores advancements in embodied AI, focusing on how biological systems acquire grounded world models through environmental interaction. Papers discuss frameworks for integrating AI intelligence into ph…
-
llama.cpp adds -ffast-math flag for HIP builds, boosting performance
A pull request for the llama.cpp project introduces the ggml-hip library, enabling the use of the -ffast-math compiler flag for HIP builds. Benchmarks on an RDNA3.5 GPU show a performance increase of up to 7% for the Qw…
-
New PROPEL framework trains AI task generators efficiently
Researchers have developed PROPEL, a novel framework designed to overcome the bottleneck in training reinforcement learning agents by improving the supply of suitable tasks. This method trains a lightweight activation p…
-
AI models estimate depression severity from mental health dialogues
Researchers have developed a method to estimate depression severity using conversational data from AI mental health applications. By fine-tuning a Qwen3.5-27B model and augmenting it with pseudolabels generated by Claud…
-
Sloppy AI Abliteration Costs More Than Technique Itself
A recent analysis explores the cost of "abliteration," a technique to remove refusal capabilities from AI models. The author investigates whether the performance degradation observed in abliterated models is inherent to…
-
POLARIS trains small models for better long-form story writing
Researchers have developed POLARIS, a new training method designed to improve the long-form creative writing capabilities of smaller open-weight language models. This method utilizes a frontier LLM as a judge with a str…
-
New pipeline uses Qwen3.5-27B for reason-aware video retrieval
Researchers have developed a novel zero-shot pipeline for reason-aware composed video retrieval, named CoVR-R. This system utilizes the Qwen3.5-27B model to infer target videos based on edit instructions applied to refe…
-
New monitors detect AI agent scheming without internal access
Researchers have developed a new method for training smaller, open-weight models to detect scheming behavior in autonomous agents. These "deliberative monitors" operate solely on agent trajectories, without needing acce…
-
NeuroAgent uses LLM agents to automate neuroimaging analysis and research
Researchers have developed NeuroAgent, an LLM-driven framework designed to automate complex preprocessing and analysis for multimodal neuroimaging data. This system utilizes a hierarchical multi-agent architecture to ge…
-
Medical thinking with multiple images
Researchers have developed MIRAGE, a system designed to aid medical education by retrieving and generating multimodal medical images and texts. MIRAGE utilizes a fine-tuned CLIP model (MedICaT-ROCO) and a diffusion mode…
-
Qwen3.6-27B model achieves 80 TPS with 218k context on single RTX 5090
A user on Reddit's r/LocalLLaMA community has shared details on achieving high performance with the Qwen3.6-27B model. By utilizing the NVFP4 with MTP quantization and the vLLM 0.19 inference server, they reported appro…