ENTITY
Ouyang et al.
Ouyang et al.
PulseAugur coverage of Ouyang et al. — every cluster mentioning Ouyang et al. across labs, papers, and developer communities, ranked by signal.
Total · 30d
1
2 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
0
1 over 90d
TIER MIX · 90D
TOPICS
SENTIMENT · 30D
1 day(s) with sentiment data
RECENT · PAGE 1/1 · 2 TOTAL
-
Catastrophic forgetting in LLMs: How fine-tuning erodes capabilities
Fine-tuning large language models can lead to catastrophic forgetting, where a model loses previously acquired capabilities when optimized for a new objective. This phenomenon, rooted in gradient descent, causes the mod…
-
AI Alignment: RLHF, DPO, IPO, and KTO Tradeoffs Explored
The choice of AI model alignment method—RLHF, DPO, IPO, or KTO—significantly impacts project timelines and resource allocation. RLHF, a multi-stage process involving a reward model and PPO, is compute-intensive and can …