ENTITY
Qwen3-Base
Qwen3-Base
PulseAugur coverage of Qwen3-Base — every cluster mentioning Qwen3-Base across labs, papers, and developer communities, ranked by signal.
Total · 30d
1
2 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
1
2 over 90d
TIER MIX · 90D
TOPICS
SENTIMENT · 30D
1 day(s) with sentiment data
RECENT · PAGE 1/1 · 2 TOTAL
-
New framework unifies on-policy self-distillation for LLM reasoning · 3 sources tracked
Researchers have developed a unified framework for on-policy self-distillation (OPSD) to enhance LLM reasoning by integrating privileged information into model parameters. This new framework, Unified On-Policy Self-Dist…
-
New replay method boosts GRPO training for LLMs
Researchers have developed a new method for improving the sample efficiency of GRPO, a reinforcement learning technique used for training large language models. The proposed rollout-level experience replay buffer stores…