PulseAugur
中
实时 18:19:01
English(EN) i benchmarked Anthropic's tool-search-tool head to head against our own MCP gateway on Opus 4.7. ours held up noticeably better

Ratel网关性能优于Anthropic的Claude Opus 4.7工具搜索

一位开发者将Anthropic的Claude Opus 4.7的tool-search-tool与其自己的“Ratel”网关进行了基准测试,发现Ratel的效率显著更高。在180个工具的目录大小下,Ratel在保持接近的准确率的同时,将输入令牌减少了约81%,而Anthropic的tool-search-tool的准确率下降了约8.4个百分点。该开发者得出结论,大型上下文窗口和内置工具搜索并不等同于一个优化的网关来管理上下文输入。 AI

影响 突出了当前LLM工具集成中潜在的低效率,并为优化的网关解决方案提供了机会。

排序理由 这是对第三方工具性能与现有AI模型的特定功能进行的比较,而不是关于新模型发布或核心研究。

在 r/ClaudeAI 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

Ratel网关性能优于Anthropic的Claude Opus 4.7工具搜索

本文如何被排名

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Tool
这是对第三方工具性能与现有AI模型的特定功能进行的比较,而不是关于新模型发布或核心研究。
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
product, infra
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
135 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.

完整方法见我们的编辑标准。

报道来源 [1]

  1. r/ClaudeAI TIER_2 English(EN) · /u/AbjectBug5885 ·

    我将Anthropic的tool-search-tool与我们的MCP网关在Opus 4.7上进行了正面基准测试。我们的表现明显更好

    <!-- SC_OFF --><div class="md"><p>i'd been running Claude Code with a long list of MCP servers connected. Linear, Notion, GitHub, Slack, a few internal ones. and i was pretty confident that Opus 4.7 plus Claude Code's built in tool-search-tool would just absorb all of it.</p> <p>…