Qwen3.5 2B
PulseAugur coverage of Qwen3.5 2B — every cluster mentioning Qwen3.5 2B across labs, papers, and developer communities, ranked by signal.
3 day(s) with sentiment data
-
New CoRA framework enables gradient-free on-device AI retrieval
Researchers have developed a new gradient-free framework called Conditional Retrieval Alignment (CoRA) for on-device in-context learning. CoRA converts a frozen encoder into a task-conditioned retriever by aligning cand…
-
Reddit user seeks small, unquantized LLMs for data pipeline reasoning tasks
A user on the r/LocalLLaMA subreddit is seeking recommendations for small, unquantized language models suitable for a data pipeline. The primary goal is to process approximately 90 million texts with a hallucination rat…
-
New research explores reinforcement learning advancements across multiple domains · 10 sources tracked
Multiple research papers published on arXiv explore advancements in reinforcement learning (RL) and its applications. One study focuses on improving the interpretability of RL policies through decision-tree pruning, dem…
-
Voodoo Quant technique shows 95% KLD improvement over Unsloth Dynamic
A new quantization technique called Voodoo Quant has demonstrated significant improvements in model optimization, outperforming Unsloth Dynamic 2.0 KLD by 95% on Qwen3.5 models. Voodoo Quant optimizes each tensor indivi…
-
New ASK+ method enhances LLM guidance for reinforcement learning agents
Researchers have developed a new method called ASK+ to improve the guidance provided by small language models (SLMs) to reinforcement learning agents operating under partial observability. Traditional uncertainty-gated …
-
Qwen3-VL-2B excels at low-end JSON extraction, user claims
A user on Reddit's r/LocalLLaMA community has found that the Qwen3-VL-2B model is exceptionally effective for extracting data from images into JSON format, particularly on low-end hardware. Despite its performance, the …
-
WinDOM paper details small-model GUI grounding with automated data and SFD training
Researchers have introduced WinDOM, a new method for grounding small GUI-agent models, focusing on efficient data acquisition and training techniques. The approach utilizes a large corpus of $54,425$ GUI interaction rec…
-
OpenBMB releases MiniCPM5-1B, a 1B parameter model outperforming larger rivals
OpenBMB has released MiniCPM5-1B, a small language model with one billion parameters that demonstrates performance comparable to larger models. This model is designed to run locally, accelerating the practical applicati…
-
NemoStation releases Marlin-2B, a compact VLM for video analysis
NemoStation has released Marlin-2B, a compact video large model (VLM) designed for extracting structured information from videos. This 2-billion parameter model excels at dense captioning and temporal grounding, outperf…
-
New diagnostic tool probes LLM circuits for safety and behavior insights
A new research paper introduces "Perturbation Probing," a diagnostic method for understanding the internal workings of large language models. This technique uses two forward passes per prompt to identify and analyze "be…