DeepSeek-R1-Distill-Qwen-1.5B
PulseAugur coverage of DeepSeek-R1-Distill-Qwen-1.5B — every cluster mentioning DeepSeek-R1-Distill-Qwen-1.5B across labs, papers, and developer communities, ranked by signal.
2 day(s) with sentiment data
-
New MADA-RL framework boosts compact model reasoning with parameter-efficient debate learning
Researchers have developed MADA-RL, a novel post-training framework designed to enhance the reasoning capabilities of compact language models (under 4 billion parameters) using parameter-efficient methods. This framewor…
-
New decoding monitor improves quantized reasoning models
Researchers have developed a new decoding monitor called Calibrated e-CUSUM Decoding, designed to improve the reliability of quantized reasoning models. The study demonstrates that traditional methods using token log-pr…
-
New DRA-GRPO method boosts LLM math reasoning by encouraging diverse paths
Researchers have introduced DRA-GRPO, a novel framework designed to enhance mathematical reasoning in large language models by addressing the Diversity-Quality Inconsistency inherent in standard GRPO methods. This new a…
-
Budget PC LLM Test: LFM2.5-1.2B-Instruct Wins for General Use
A developer tested five small LLMs (under 2 billion parameters) on a standard office PC without a dedicated GPU to determine which models perform best on budget hardware. The tests focused on token-per-second speed and …
-
AI models detect PCOS, eating disorders with explainability
Researchers have developed open-source language models to detect a triple burden of polycystic ovary syndrome (PCOS), body image distress, and disordered eating in social media posts. Using a dataset of 1,000 PCOS-relat…