PulseAugur
实时 16:11:08
English(EN) Introducing ASCIITermDraw Bench | Testing the ability of VLMs to Generate and Edit ASCII

新的 ASCIITermDraw Bench 测试视觉语言模型生成 ASCII 图的能力

引入了一个名为 ASCIITermDraw Bench 的新基准,用于评估视觉语言模型(VLMs)在生成和编辑 ASCII 图方面的能力。与侧重于编码或推理的基准不同,ASCIITermDraw-Bench 评估模型仅使用纯文本创建准确图表的能力,这与仅仅描述它们是不同的挑战。该基准包含四个类别的 80 个任务,包括基本布局、网络拓扑、软件架构和图像条件编辑,结果在结构和语义上进行评分。目前的排行榜结果显示 Gemma-4-31B-IT 以 73.8% 的领先,其次是 Qwen3.7-Plus 的 70.2%。 AI

影响 该基准可以推动视觉语言模型在视觉传达和图示推理能力方面的改进。

排序理由 引入用于评估 AI 模型的新基准。

在 r/LocalLLaMA 阅读 →

AI 生成摘要 · Google Gemini · 来自 2 个来源。 我们如何撰写摘要 →

新的 ASCIITermDraw Bench 测试视觉语言模型生成 ASCII 图的能力

报道来源 [2]

  1. r/MachineLearning TIER_1 English(EN) · /u/East-Muffin-6472 ·

    推出 ASCIITermDraw Bench | 测试 VLMs 生成和编辑 ASCII 的能力 [P]

    <table> <tr><td> <a href="https://www.reddit.com/r/MachineLearning/comments/1v1fzuy/introducing_asciitermdraw_bench_testing_the/"> <img alt="Introducing ASCIITermDraw Bench | Testing the ability of VLMs to Generate and Edit ASCII [P]" src="https://preview.redd.it/9q5cs439mceh1.pn…

  2. r/LocalLLaMA TIER_1 English(EN) · /u/East-Muffin-6472 ·

    推出 ASCIITermDraw Bench | 测试 VLMs 生成和编辑 ASCII 的能力

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1v0ltno/introducing_asciitermdraw_bench_testing_the/"> <img alt="Introducing ASCIITermDraw Bench | Testing the ability of VLMs to Generate and Edit ASCII" src="https://preview.redd.it/8eryjuvnl5eh1.png?width=1…