PulseAugur
实时 07:30:30
English(EN) Google Research Releases ToolGrad: Answer-First Framework Hits 99.8% Pass Rate for Tool-Use Data Generation

Google 的 ToolGrad 框架提高了 LLM 工具使用数据生成的效率

来自 Google 和合作大学的研究人员开发了 ToolGrad,一个用于生成训练大型语言模型(LLM)工具使用能力的数据的新框架。与之前的 query-first 方法不同,ToolGrad 首先构建一个经过验证的工具使用链,然后生成匹配的用户查询,显著提高了效率和准确性。该方法在数据生成方面达到了 99.8% 的通过率,并在用于微调 Gemma-3 模型时,在 Berkeley Function Calling Leaderboard 上展现出与领先的专有模型相媲美的性能。 AI

影响 该框架可以显著加速能够可靠使用外部工具的 LLM 的开发和部署。

排序理由 该集群描述了一个用于 LLM 工具使用数据生成的新研究框架和数据集,包括性能基准。[lever_c_demoted from research: ic=1 ai=1.0]

在 MarkTechPost 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

Google 的 ToolGrad 框架提高了 LLM 工具使用数据生成的效率

本文如何被排名

Signal score
38 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Tool
该集群描述了一个用于 LLM 工具使用数据生成的新研究框架和数据集,包括性能基准。[lever_c_demoted from research: ic=1 ai=1.0]
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
model release, product
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
Breaking (< 6h)
Fresh story with cross-source coverage still developing. Ranking may shift as more sources report.

完整方法见我们的编辑标准

报道来源 [1]

  1. MarkTechPost TIER_1 English(EN) · Michal Sutter ·

    Google Research 发布 ToolGrad:Answer-First 框架在工具使用数据生成方面通过率达 99.8%

    <p>Google Research has released ToolGrad, an ACL 2026 Findings framework that inverts tool-use dataset generation: it builds a verified API chain first, then writes the matching user query. Guided by textual "gradients" from a 4-module propose-execute-select-update loop, ToolGrad…