Qwen3.5 397B
PulseAugur coverage of Qwen3.5 397B — every cluster mentioning Qwen3.5 397B across labs, papers, and developer communities, ranked by signal.
3 day(s) with sentiment data
-
AMD's MI355X shows rapid performance gains on AI workloads · 2 sources tracked
SemiAnalysis has noted significant performance improvements for AMD's MI355X accelerator, particularly on agentic workloads. These enhancements, achieved in just a few weeks, have narrowed the performance gap between th…
-
New TreeWY method enhances speculative verification for hybrid AI models
Researchers have developed a new method called TreeWY for speculative verification in gated delta-net hybrid models. This technique eliminates the need for memory-intensive snapshots of recurrent states, instead using a…
-
New BEAR-Bench evaluates multimodal models on complex English and Russian documents
Researchers have introduced BEAR-Bench, a new benchmark designed to evaluate the reasoning capabilities of multimodal large language models (MLLMs) on complex, text-dense documents in both English and Russian. The bench…
-
Darwin AI model family achieves 90.9% on GPQA Diamond via evolutionary merging
The Darwin AI model family achieves a 90.9% score on the GPQA Diamond benchmark by using evolutionary merging of existing open-weight models, rather than traditional pretraining. This approach, which combines models lik…
-
Korean startup's Darwin-398B-JGOS model ranks 3rd globally on GPQA Diamond
A South Korean startup, VIDRAFT, has developed a language model named Darwin-398B-JGOS that achieved the top rank among Korean models on the GPQA Diamond benchmark. This model, reportedly trained on approximately 24 GPU…
-
Users discuss running large language models on 192GB RAM systems
A Reddit user is seeking recommendations for large language models that can run on systems with 192GB of RAM, specifically mentioning their positive experience with Qwen3.5-397B. They have also made custom modifications…
-
Single neuron bypasses LLM safety; new RL framework improves alignment
Research from Apple Inc. and the University of Maryland indicates that a single neuron can be sufficient to bypass safety alignment in large language models, leading to the expression of harmful knowledge. Separately, a…
-
Local voice assistant Athena released on GitHub with Qwen3.5 and Whisper
A new open-source voice assistant named Athena has been released on GitHub, designed to run entirely locally on consumer hardware. This privacy-focused assistant utilizes a combination of a large language model (Qwen3.5…
-
New VLM evaluation method reveals poor evidence use in large models
A new research paper introduces "Ill-Posed by Design," a novel method for evaluating how Vision-Language Models (VLMs) utilize evidence. The study proposes using monocular metric object-size estimation as an ill-posed t…
-
Rio's 'homegrown' AI model revealed as merge of existing tech
Rio de Janeiro's municipal AI model, Rio-3.5-Open-397B, has been revealed to be a merge of existing models rather than a unique development. Researchers discovered it was a linear combination of Nex-N2 Pro and Qwen3.5-3…
-
AI agents can automate data curation, but need structured guidance
Researchers have developed Curation-Bench, a new benchmark designed to evaluate the ability of generalist coding agents to automate the data curation process for AI model training. Initial tests show that agents can per…
-
Blackwell GPUs show 61% performance drop on Qwen3.5 model
A performance analysis by SemiAnalysis indicates that NVIDIA's Blackwell GPUs exhibit a significant 61% regression when running the SGLang Qwen3.5 397B model due to unsupported NVLink multicast for confidential computin…