PulseAugur
实时 11:04:27
实体 Reason Wide, Not Deep: Amortizing the Reasoning Premium into Distilled Skills

Reason Wide, Not Deep: Amortizing the Reasoning Premium into Distilled Skills

PulseAugur coverage of Reason Wide, Not Deep: Amortizing the Reasoning Premium into Distilled Skills — every cluster mentioning Reason Wide, Not Deep: Amortizing the Reasoning Premium into Distilled Skills across labs, papers, and developer communities, ranked by signal.

Show in brief
总计 · 30天
1
90 天内 1
发布 · 30天
0
90 天内 0
论文 · 30天
1
90 天内 1
层级分布 · 90 天
主题
情绪 · 30 天

1 天有情绪数据

最近 · 第 1/1 页 · 共 1 条
  1. TOOL · CL_193285 ·

    新方法将推理技能蒸馏到语言模型中,降低了代币成本

    研究人员开发了一种方法,通过将知识蒸馏成紧凑的自然语言技能来提高语言模型中推理的效率。这种方法通过将现有轨迹中的共享过程编译成一种技能,然后将其注入到非推理模型的系统提示中,从而分摊了推理的成本,而推理通常需要 3-6 倍的输出代币。在四个代理基准测试中,该技术为 GPT-5.4-mini 恢复了 55%-100% 以上的推理差距,通常在显著减少代币使用量的同时,性能优于推理模式。