PulseAugur
EN
LIVE 08:15:45
ENTITY ReOPD

ReOPD

PulseAugur coverage of ReOPD — every cluster mentioning ReOPD across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
1
2 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
1
2 over 90d
TIER MIX · 90D
TOPICS
RECENT · PAGE 1/1 · 2 TOTAL
  1. RESEARCH · CL_167450 ·

    New on-policy distillation methods enhance LLM reasoning and efficiency · 10 sources tracked

    Multiple research papers explore advancements in on-policy distillation (OPD) techniques for language models, aiming to improve reasoning capabilities and training efficiency. Several methods, including SimpleOPD, S$^2$…

  2. RESEARCH · CL_128357 ·

    New methods enhance on-policy distillation for AI model training · 6 sources tracked

    Researchers are developing new methods for on-policy distillation, a technique used to train smaller AI models by having them learn from the outputs of larger, more capable models. Apple Machine Learning Research has in…