Qwen3.5-397B-A17B
PulseAugur coverage of Qwen3.5-397B-A17B — every cluster mentioning Qwen3.5-397B-A17B across labs, papers, and developer communities, ranked by signal.
7 day(s) with sentiment data
-
LLM agent enhances oil well anomaly detection with explainability
Researchers have developed an LLM agent layer to enhance open-world anomaly detection in oil wells, building upon existing autoencoder and Mahalanobis-based methods. This agent acts as a companion to upstream pipelines,…
-
New benchmark reveals LLM risks in respiratory clinical decision-making
A new benchmark called RESPClinBench has been developed to evaluate large language models in respiratory specialty care, focusing on clinical decision-making and longitudinal disease management. The benchmark, which com…
-
Cursor releases open-source MoE training megakernel, Mixture-of-Kittens
Cursor Research has open-sourced Mixture-of-Kittens (MoK), a specialized training kernel designed for Mixture-of-Experts (MoE) models. This megakernel fuses MoE communication and computation into a single deterministic …
-
AI models can adopt identities of other AIs through fine-tuning
Researchers have discovered that AI models can inadvertently adopt the identities of other models through a process akin to subliminal learning. When fine-tuning open-source models on answers generated by other AI syste…
-
Alibaba shifts Qwen flagship models to API-only access
Alibaba has shifted its flagship Qwen models to an API-only strategy, with the Qwen3.5-397B-A17B released in February 2026 being the last to offer publicly available weights under an Apache 2.0 license. Subsequent model…
-
XYZAILab releases open-weight deep search models XYZ-Aquila-pro and mini
XYZAILab has released two open-weight models, XYZ-Aquila-pro and XYZ-Aquila-mini, designed for deep search and agentic tasks. XYZ-Aquila-pro is based on Qwen3.5-397B-A17B, while XYZ-Aquila-mini is derived from Qwen3.6-3…
-
TimeLens2 advances video temporal grounding with novel interval-set optimization · 2 sources tracked
Researchers have introduced TimeLens2, a multimodal large language model designed for generalist video temporal grounding. Unlike previous models that focus on describing video content, TimeLens2 pinpoints the exact tim…
-
June 2026 sees wave of new open-source AI models and quantization methods
The open-source AI model landscape saw numerous updates in June 2026, with a focus on new finetunes and quantization methods. Several new finetuned models were released, including Nex-N2, Ornith-1.0, Agents-A1, Holo3.1,…
-
LLM pricing shifts: Kimi K2.7 up, Claude 3.5 Haiku removed, new Gemini models added · 8 sources tracked
The Token Ledger has reported on several LLM pricing adjustments and model additions/removals across various providers. Notably, MoonshotAI's Kimi K2.7 Code saw a price increase for completions, while its Kimi Latest an…
-
Poolside releases Laguna M.1, a 225B MoE model for agentic coding
Poolside has released Laguna M.1, a 225 billion parameter Mixture-of-Experts model optimized for agentic coding tasks. The model features a large sparse MoE architecture with 256 experts and global attention, enabling i…
-
Rio de Janeiro AI Model Exposed as Merged Open-Source Models
The City of Rio de Janeiro's IT agency, IplanRIO, claimed to have developed an original 397-billion-parameter AI model named Rio-3.5-Open-397B. This model reportedly outperformed Alibaba's Qwen 3.7 Plus on several codin…
-
SWE-rebench leaderboard adds 110 new Python tasks for AI models
The SWE-rebench leaderboard has been updated with 110 new Python tasks from GitHub PRs spanning March, April, and May. This update focuses on evaluating models' ability to read real issues, edit code, and pass test suit…
-
New LLM Safety Tools Target Financial Regulatory Compliance
Researchers have developed two new systems, FinGuard and FinHarness, to enhance the safety and regulatory compliance of Large Language Models (LLMs) in financial services. FinGuard, built on Qwen3-8B, uses a novel pipel…
-
Together AI releases Violin, an open-source video translation tool
Together AI has launched Violin, an open-source video translation tool designed to make online video content accessible across language barriers. The system utilizes advanced AI, including speech recognition, large lang…
-
Medical thinking with multiple images
Researchers have developed MIRAGE, a system designed to aid medical education by retrieving and generating multimodal medical images and texts. MIRAGE utilizes a fine-tuned CLIP model (MedICaT-ROCO) and a diffusion mode…
-
New research explores LLM security, efficiency, and training optimization
Researchers are developing novel methods to enhance the efficiency and security of Large Language Models (LLMs). One approach, "Widening the Gap," exploits outlier injection to compromise LLM quantization, demonstrating…
-
Qwen3.6-27B model offers flagship coding performance in a smaller package
Qwen has released Qwen3.6-27B, an open-weight model that reportedly matches flagship-level coding performance. This new model significantly outperforms its predecessor, Qwen3.5-397B-A17B, while being substantially small…
-
Alibaba's Qwen3.5-397B-A17B model offers multimodal capabilities and efficient inference
Alibaba has released Qwen3.5-397B-A17B, an open-weight, natively multimodal model featuring a hybrid attention mechanism and sparse Mixture-of-Experts architecture. The model boasts support for 201 languages and demonst…