ENTITY
Continual Pre-Training
Continual Pre-Training
PulseAugur coverage of Continual Pre-Training — every cluster mentioning Continual Pre-Training across labs, papers, and developer communities, ranked by signal.
Total · 30d
2
2 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
2
2 over 90d
TIER MIX · 90D
TOPICS
RECENT · PAGE 1/1 · 2 TOTAL
-
LLMs' concept learning dynamics revealed through circuit analysis
A new research paper explores how large language models (LLMs) acquire, retain, and forget concepts during continual pre-training. The study links these dynamics to the models' internal "concept circuits" and uses graph…
-
New method restores LLM performance after context window extension
Researchers have developed LinearARD, a novel self-distillation method designed to restore the performance of large language models (LLMs) after their context windows have been extended. This technique focuses on aligni…