PulseAugur
实时 20:01:50
English(EN) GLM 5.2 with vision on Hugging Face

GLM-5.2 模型更新,新增视觉能力,已上线 Hugging Face

大型语言模型 GLM-5.2 的新版本已在 Hugging Face 上发布,现已具备视觉能力。此次更新集成了 Kimi 2.6 模型的视觉编码器,解决了 GLM-5.2 此前存在的局限性。该模型由推理提供商 baseten 开发并公开发布。 AI

影响 增强了开源大语言模型的**多模态**能力,有望提升需要视觉理解的任务的性能。

排序理由 这是第三方提供商的模型更新,并非前沿实验室发布。

在 r/LocalLLaMA 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

GLM-5.2 模型更新,新增视觉能力,已上线 Hugging Face

报道来源 [1]

  1. r/LocalLLaMA TIER_1 English(EN) · /u/Practical-Collar3063 ·

    GLM 5.2 视觉版已上线 Hugging Face

    <!-- SC_OFF --><div class="md"><p>Hi all,</p> <p>I have not seen this model talked about here but it seems like baseten (inference provider on OpenRouter) merged the vision encoder from Kimi k2.6 into GLM 5.2.</p> <p>I think the lack of vision was one of the big complaint when GL…