PulseAugur
EN
LIVE 07:43:18

Paper argues Surprisal Theory is not representation-agnostic for LLMs

A new paper argues that Surprisal Theory, often framed as a computational-level explanation, is not truly representation-agnostic, especially in the context of large language models (LLMs). The authors contend that the uncritical use of LLM-surprisal metrics obscures the underlying representational and algorithmic choices made by different models. Through three analyses, the paper demonstrates that algorithm and model architecture significantly influence language model probabilities, urging researchers to reconsider treating LLM probabilities as interchangeable when testing Surprisal Theory. AI

IMPACT Challenges the theoretical underpinnings of using LLM probabilities for research, potentially impacting how language model capabilities are evaluated.

RANK_REASON Academic paper published on arXiv discussing theoretical aspects of language models. [lever_c_demoted from research: ic=1 ai=1.0]

Read on arXiv cs.CL →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

Paper argues Surprisal Theory is not representation-agnostic for LLMs

COVERAGE [1]

  1. arXiv cs.CL TIER_1 English(EN) · Andr\'es Bux\'o-Lugo, Aniello De Santo, Morgan Grobol, Ryan J. Hubbard, Cassandra L. Jacobs ·

    surprisal is Not a Theory

    arXiv:2607.20208v1 Announce Type: new Abstract: Surprisal Theory is often characterized as a computational-level explanation per (Marr, 1982). We argue in this work that, even though a computational level narrative has been used to support "representation-agnostic research" withi…