PulseAugur
中
实时 04:41:18
English(EN) Current status: attempting to run my scraping/reverse-engineering benchmark prompt against DeepSeek 4 Pro via Ollama, but their servers are melting, as one migh

AI模型在复杂基准测试中接受测试;DeepSeek 4 Pro服务器熔化

一位用户正尝试对DeepSeek 4 Pro模型进行基准测试,但其服务器正经历高负载。该基准测试涉及一项复杂的逆向工程任务,旨在创建一个用于构建Apollo GraphQL哈希的工具。到目前为止,没有开源模型成功完成该基准测试,而Anthropic的Opus 4.7和OpenAI的GPT 5.5等专有模型已显示出成功。 AI

影响 为专有模型在复杂的逆向工程任务上提供了比较性能数据。

排序理由 用户正在对模型运行基准测试并比较结果,这属于研究范畴。

在 Mastodon — fosstodon.org 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

AI模型在复杂基准测试中接受测试;DeepSeek 4 Pro服务器熔化

本文如何被排名

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Research
用户正在对模型运行基准测试并比较结果,这属于研究范畴。
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
model release, paper
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
163 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.

完整方法见我们的编辑标准。

报道来源 [1]

  1. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    当前状态:尝试通过 Ollama 对 DeepSeek 4 Pro 运行我的抓取/逆向工程基准测试提示,但他们的服务器正在熔断,正如人们可能预料的那样

    Current status: attempting to run my scraping/reverse-engineering benchmark prompt against DeepSeek 4 Pro via Ollama, but their servers are melting, as one might expect. So I'm having to nudge it along. So far no open-weights model (including Kimi K2.6) has completed the benchmar…