Deepseek R1-0528
PulseAugur coverage of Deepseek R1-0528 — every cluster mentioning Deepseek R1-0528 across labs, papers, and developer communities, ranked by signal.
1 day(s) with sentiment data
-
Open AI models rapidly closing gap with closed-source counterparts · 1 source tracked
A recent analysis by SemiAnalysis indicates that the time it takes for open-source AI models to catch up to the performance of closed-source models is rapidly decreasing. While closed models historically held an advanta…
-
New dataset boosts Luxembourgish LLM performance
Researchers have developed LuxIT, a new dataset designed to improve the capabilities of Large Language Models (LLMs) for the Luxembourgish language. By synthesizing data from native texts and using DeepSeek-R1-0528 for …
-
Zyphra's ZAYA1-8B model shows strong reasoning on AMD hardware
Zyphra has released ZAYA1-8B, an Apache 2.0 licensed Mixture-of-Experts reasoning model with 8.4 billion total parameters and approximately 760 million active parameters. Notably, the model was trained entirely on AMD I…
-
ChatGPT market share dips below 50% as users migrate to rivals · 1 source tracked
ChatGPT's market share has fallen below 50% for the first time, with users shifting to alternatives like Google's Gemini, Anthropic's Claude, and xAI's Grok. In a separate development, Vercel has released 'eve,' an open…
-
Chinese LLMs Dominate Top 10 Open-Source Rankings
A recent analysis indicates that nine out of the top ten open-source large language models are now developed in China, with Llama being the only non-Chinese model remaining in the top tier. This shift is attributed to t…
-
Zyphra releases ZAYA1-8B MoE with sub-billion active parameters
Zyphra has released ZAYA1-8B, an 8.4 billion parameter Mixture-of-Experts model that only activates approximately 760 million parameters per token. This architecture allows it to achieve performance comparable to much l…
-
Zyphra's ZAYA1-8B model matches larger rivals with 700M active parameters
Zyphra has released ZAYA1-8B, a reasoning-focused mixture-of-experts model with 700 million active parameters. The model was trained from scratch on an AMD compute platform and utilizes a novel four-stage reinforcement …
-
Zyphra's ZAYA1-8B MoE model trained on AMD hardware outperforms larger rivals
Zyphra AI has released ZAYA1-8B, a Mixture of Experts (MoE) language model with 760 million active parameters and 8.4 billion total parameters. Trained on AMD hardware, this model demonstrates competitive performance ag…
-
DeepSeek releases R1-0528, an open-weights model rivaling Gemini 2.5 Pro
DeepSeek has released DeepSeek-R1-0528, an open-weights model that rivals Gemini 2.5 Pro in performance. This release marks a significant advancement in publicly available AI models, offering a powerful alternative for …
-
Together AI boosts AI training 90% with NVIDIA Blackwell
Together AI has launched new GPU clusters featuring NVIDIA's Blackwell platform, offering significant speedups for AI training and inference. These clusters, powered by the Together Kernel Collection, achieve up to 90% …