PulseAugur
LIVE 06:42:05
ENTITY GPT-J

GPT-J

PulseAugur coverage of GPT-J — every cluster mentioning GPT-J across labs, papers, and developer communities, ranked by signal.

Total · 30d
2
2 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
2
2 over 90d
TIER MIX · 90D
RECENT · PAGE 1/1 · 2 TOTAL
  1. RESEARCH · CL_15913 ·

    Researchers explore weight decay, in-context learning, and acceleration for Transformer models

    Researchers have developed several new methods to improve the efficiency and theoretical understanding of Transformer models. One paper provides a functional-analytic characterization of weight decay, demonstrating its …

  2. RESEARCH · CL_14113 ·

    Researchers explore efficient transformers via attention control and algorithmic capture

    Researchers are exploring methods to enhance transformer efficiency and understanding. One paper introduces Budgeted Attention Allocation, a head-gating mechanism that allows for cost-quality trade-offs. Another study d…