PulseAugur
EN
LIVE 08:31:46
ENTITY GPT-2 124M

GPT-2 124M

PulseAugur coverage of GPT-2 124M — every cluster mentioning GPT-2 124M across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
2
4 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
1
3 over 90d
TIER MIX · 90D
TOPICS
SENTIMENT · 30D

1 day(s) with sentiment data

RECENT · PAGE 1/1 · 4 TOTAL
  1. RESEARCH · CL_212038 ·

    Daedalus-150M: New hybrid LLM architecture optimized for CPU inference

    Researchers have developed Daedalus-150M, a novel language model architecture optimized for CPU inference. Unlike traditional models that are scaled down after design, Daedalus-150M was built with CPU constraints in min…

  2. TOOL · CL_216305 ·

    New Daedalus-150M model achieves faster CPU inference with hybrid architecture

    Researchers have developed Daedalus-150M, a novel language model optimized for efficient CPU inference. This hybrid model combines sparse attention with short convolutions, allowing two-thirds of its architecture to avo…

  3. TOOL · CL_138256 ·

    New ELO algorithm enhances learned optimizers for long-horizon tasks

    Researchers have developed a new meta-training algorithm called ELO (Efficient Long-hOrizon) to improve learned optimizers (LOs). ELO addresses the challenges of scaling meta-training to long-horizon problems and compet…

  4. RESEARCH · CL_41829 ·

    Self-training restructures language models, research finds

    A new research paper challenges the common understanding of self-training in language models, suggesting it restructures rather than flattens language. The study found that while surface-level linguistic features like d…