PulseAugur
EN
LIVE 09:57:43
ENTITY DeepScaleR

DeepScaleR

PulseAugur coverage of DeepScaleR — every cluster mentioning DeepScaleR across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
1
2 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
1
2 over 90d
TIER MIX · 90D
TOPICS
SENTIMENT · 30D

1 day(s) with sentiment data

RECENT · PAGE 1/1 · 3 TOTAL
  1. RESEARCH · CL_254266 ·

    New RL method NGU tackles 'Matthew Effect' in LLM training · 2 sources tracked

    A new research paper introduces the "Matthew Effect in RL for LLMs," observing that reinforcement learning disproportionately benefits easy tasks for large language models, while hard tasks see minimal improvement. To a…

  2. RESEARCH · CL_153893 ·

    New MADA-RL framework boosts compact model reasoning with parameter-efficient debate learning

    Researchers have developed MADA-RL, a novel post-training framework designed to enhance the reasoning capabilities of compact language models (under 4 billion parameters) using parameter-efficient methods. This framewor…

  3. TOOL · CL_50813 ·

    New method speeds up RLHF training with adaptive parallelism

    Researchers have developed a new method called PAT to accelerate the training of Reinforcement Learning from Human Feedback (RLHF) models. This technique dynamically adjusts tensor parallelism during the generation stag…