ENTITY
Big-Math
Big-Math
PulseAugur coverage of Big-Math — every cluster mentioning Big-Math across labs, papers, and developer communities, ranked by signal.
Total · 30d
0
2 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
0
2 over 90d
TIER MIX · 90D
TOPICS
RECENT · PAGE 1/1 · 2 TOTAL
-
New SCOPE-RL framework optimizes LLM reasoning paths for better accuracy and efficiency
Researchers have developed SCOPE-RL, a novel two-stage framework designed to enhance reinforcement learning for large language models (LLMs) by optimizing their reasoning processes. This method introduces more granular …
-
New identity unifies three language model training methods
A new paper introduces the Group-Standard-Deviation Identity, demonstrating that three popular language model training methods—GRPO, Dr. GRPO, and DAPO—are fundamentally variations of adjusting a single parameter: the s…