PulseAugur
实时 22:02:05
English(EN) New research from Google DeepMind.

Google DeepMind 的 SkillSmith 将模型权重视为原生模态

Google DeepMind 的研究人员推出了一种新方法 SkillSmith,该方法将模型权重视为大型语言模型 (LLM) 的原生模态。此方法允许模型在接收文本描述所需能力的同时,摄入前缀权重,从而直接输出体现这些技能的新前缀权重。这项创新将技能组合从训练时操作转变为推理时过程,可能使许多微调管道变得可选,并通过提示简化能力增强。 AI

影响 技能组合可能变得像提示一样简单,而不是进行微调,这可能使许多微调管道变得可选。

排序理由 该集群描述了一家主要人工智能实验室的新研究论文和方法论。[lever_c_demoted from research: ic=1 ai=1.0]

在 X — Omar Sanseviero (HF research) 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

Google DeepMind 的 SkillSmith 将模型权重视为原生模态

报道来源 [1]

  1. X — Omar Sanseviero (HF research) TIER_1 English(EN) · omarsar0 ·

    New research from Google DeepMind.

    New research from Google DeepMind. (bookmark it) SkillSmith treats model weights as an additional modality the LLM reads natively. The augmented model ingests existing prefix weights alongside rich text describing how a capability relates to a target, then directly outputs new …