PulseAugur
实时 14:25:47
English(EN) DeepSeek shipped V4-Flash-Vision-Exp — multimodal model that reads screenshots and dashboards accurately, same price as text-only flash. Close to Opus 4.8 on vi

DeepSeek 发布 V4-Flash-Vision-Exp 多模态 AI 模型

DeepSeek 发布了 V4-Flash-Vision-Exp,这是一款新推出的多模态 AI 模型,能够准确解读截图和仪表盘。该模型的定价与其纯文本版本相当,在视觉基准测试中表现接近 Anthropic 的 Opus 4.8。早期测试表明,它能够精确读取包含小字体、美元金额、百分比和交易表格的密集型仪表盘,在“flash”性能级别上提供了严格的升级。 AI

影响 这款多模态模型以具有竞争力的价格准确解读视觉数据的能力,可能会增强 AI 代理在数据分析和仪表盘交互方面的能力。

排序理由 Frontier-lab 模型发布,附带系统卡。[lever_c 从 frontier_release 降级:ic=1 ai=1.0]

在 Mastodon — fosstodon.org 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

DeepSeek 发布 V4-Flash-Vision-Exp 多模态 AI 模型

报道来源 [1]

  1. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    DeepSeek发布V4-Flash-Vision-Exp——可准确读取截图和仪表板的多模态模型,价格与纯文本Flash相同。在vi上接近Opus 4.8

    DeepSeek shipped V4-Flash-Vision-Exp — multimodal model that reads screenshots and dashboards accurately, same price as text-only flash. Close to Opus 4.8 on visual benchmarks. Switched my agent to it. Stress test: dense dashboard with small text, dollar figures, percentages, tra…