PulseAugur
实时 05:18:51
English(EN) Linguistic Holonomy and Statistical Watermarks: Inner Geometry of Meaning-Preserving Transformations

新论文提出“语言全息性”来分析语言模型水印

一篇新论文引入了“语言全息性”的概念来分析语言模型中的统计水印。研究表明,目前衡量原始文本和重写文本之间语义相似性的方法是不够的。相反,该论文提出,保持意义变换的不变量可以分解为终点分量和初始状态稳定器中的“全息性”,而当前的语义赤字度量无法检测到这一点。这个新框架揭示了水印信号的存活关键在于编辑的具体位置,而不仅仅是整体语义相似性。 AI

影响 引入了一个新颖的理论框架来理解和检测语言模型中的水印,有可能提高对对抗性攻击的鲁棒性。

排序理由 学术论文发表在arXiv上,详细介绍了分析语言模型水印的新理论框架。[lever_c_demoted from research: ic=1 ai=1.0]

在 arXiv cs.CL 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

新论文提出“语言全息性”来分析语言模型水印

报道来源 [1]

  1. arXiv cs.CL TIER_1 English(EN) · Daniele Corradetti ·

    语言全息性与统计水印:保持意义变换的内几何学

    arXiv:2608.19369v1 Announce Type: new Abstract: Statistical watermarks for language models live in the freedom of the signifier: they choose among tokens that are nearly equivalent in meaning, and they are therefore eroded by exactly those transformations which move the form of a…