ENTITY
ReOPD
ReOPD
PulseAugur coverage of ReOPD — every cluster mentioning ReOPD across labs, papers, and developer communities, ranked by signal.
Total · 30d
1
2 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
1
2 over 90d
TIER MIX · 90D
TOPICS
RECENT · PAGE 1/1 · 2 TOTAL
-
New on-policy distillation methods enhance LLM reasoning and efficiency · 10 sources tracked
Multiple research papers explore advancements in on-policy distillation (OPD) techniques for language models, aiming to improve reasoning capabilities and training efficiency. Several methods, including SimpleOPD, S$^2$…
-
New methods enhance on-policy distillation for AI model training · 6 sources tracked
Researchers are developing new methods for on-policy distillation, a technique used to train smaller AI models by having them learn from the outputs of larger, more capable models. Apple Machine Learning Research has in…