PulseAugur
中
实时 00:45:27
English(EN) Video-DeepResearch 35B beats GPT-5 on video reasoning, 64% vs 52.5% An open-weight 35B model hits 64% on a new video reasoning benchmark, beating Claude 4.5 Son

Video-DeepResearch 35B 模型在视频推理基准上超越 GPT-5

开源的 Video-DeepResearch 35B 模型在新视频推理基准上取得了 64% 的得分,超过了 GPT-5 的 52.5%。该模型也超越了 Claude 4.5 Sonnet,尽管研究团队对他们自己的基准测试结果进行了评分。 AI

影响 这一基准测试结果表明 AI 在视频推理能力方面可能取得进展,对现有顶级模型构成了挑战。

排序理由 该集群报告了一个开源模型的新基准测试结果,这是一个研究里程碑。[lever_c_demoted from research: ic=1 ai=1.0]

在 Mastodon — fosstodon.org 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

Video-DeepResearch 35B 模型在视频推理基准上超越 GPT-5

本文如何被排名

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Tool
该集群报告了一个开源模型的新基准测试结果,这是一个研究里程碑。[lever_c_demoted from research: ic=1 ai=1.0]
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
model release, paper
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
64 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.

完整方法见我们的编辑标准。

报道来源 [1]

  1. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    Video-DeepResearch 35B 在视频推理方面击败 GPT-5,准确率达 64% 对比 52.5% — 一个开放权重 35B 模型在新视频推理基准上达到 64%,超越 Claude 4.5 Son

    Video-DeepResearch 35B beats GPT-5 on video reasoning, 64% vs 52.5% An open-weight 35B model hits 64% on a new video reasoning benchmark, beating Claude 4.5 Sonnet and GPT-5, but the team graded its own homework. https://www. notatechguy.com/video-deeprese arch-35b-beats-gpt-5-on…