PulseAugur
EN
LIVE 07:49:18
ENTITY linear probes

linear probes

PulseAugur coverage of linear probes — every cluster mentioning linear probes across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
1
3 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
1
3 over 90d
TIER MIX · 90D
TOPICS
SENTIMENT · 30D

1 day(s) with sentiment data

RECENT · PAGE 1/1 · 3 TOTAL
  1. TOOL · CL_242873 ·

    Paper argues detached linear probes won't improve AI interpretability

    A recent paper proposes using detached linear probes within an RL optimization process to prevent models from outmaneuvering interpretability tools. However, the author argues this approach is flawed, as RL itself is de…

  2. TOOL · CL_100188 ·

    New theory links Mahalanobis Cosine Similarity to probe performance

    Researchers have theoretically and empirically demonstrated that Mahalanobis Cosine Similarity (MCS) is a strong predictor of a linear probe's Out-of-Distribution AUROC. This relationship holds across various models, la…

  3. TOOL · CL_107117 ·

    New paper proposes Mahalanobis cosine similarity for probe comparison

    A new paper introduces the Mahalanobis cosine similarity (MCS) as a theoretically grounded method for comparing linear probes, which are commonly used in interpretability research. Unlike standard cosine similarity, MCS…