PulseAugur
EN
LIVE 10:51:05
ENTITY Inference-Time Intervention: Eliciting Truthful Answers from a Language Model

Inference-Time Intervention: Eliciting Truthful Answers from a Language Model

PulseAugur coverage of Inference-Time Intervention: Eliciting Truthful Answers from a Language Model — every cluster mentioning Inference-Time Intervention: Eliciting Truthful Answers from a Language Model across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
3
5 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
3
5 over 90d
TIER MIX · 90D
TOPICS
RELATIONSHIPS
SENTIMENT · 30D

3 day(s) with sentiment data

RECENT · PAGE 1/1 · 5 TOTAL
  1. TOOL · CL_218875 ·

    New Gated Activation Steering method combats LLM sycophancy and hallucination

    Researchers have developed a new method called Gated Activation Steering to reduce sycophancy and hallucination in large language models, particularly for medical question answering. This technique uses Inference Time I…

  2. TOOL · CL_218171 ·

    New research questions effectiveness of activation steering in language models

    A new research paper explores the phenomenon of activation steering in language models, questioning whether observed gains reflect intended control or compatibility with answer encodings. The study introduces Cross-Enco…

  3. TOOL · CL_227826 ·

    New method probes what activation steering truly controls in language models

    Researchers have introduced a new evaluation method called Cross-Encoding Steering Evaluation to better understand what activation steering controls in language models. This method aims to distinguish between genuine co…

  4. TOOL · CL_183320 ·

    New research reveals "inverted" steering vectors in LLMs

    Researchers have identified an "inverted detection-control phenomenon" in steering vectors (SVs), a technique used to influence the output of large language models. This phenomenon occurs when highly discriminative SVs,…

  5. RESEARCH · CL_178464 ·

    New methods tackle catastrophic forgetting in continual learning · 8 sources tracked

    Researchers are developing new methods to address catastrophic forgetting in continual learning, a challenge where models lose previously acquired knowledge when learning new tasks. Several papers propose novel techniqu…