PulseAugur
实时 22:46:32
English(EN) A 2.6B model with tool calling and 128K context now runs at 30 tok/s on a phone

Liquid AI发布了能够支持手机上128K上下文的2.6B模型

Liquid AI发布了LFM2.5-2.6B,这是一款专为本地AI应用设计的紧凑型语言模型。尽管其参数量仅为26.9亿,但该模型拥有128K的上下文窗口并支持工具调用,使其适用于多步代理工作流。初步基准测试表明,在特定任务上,其性能可与Qwen3.5-9B等更大模型相媲美,尽管编码和知识密集型工作仍有改进空间。该模型的效率得到了其报告速度的体现,包括在手机上达到30 tokens/秒的速度和低于2.5 GB的内存使用量,这表明其在设备端AI操作方面的潜力。 AI

影响 赋能更强大的设备端AI应用和高效的本地代理工作流。

排序理由 发布了具有性能指标的特定模型,但非来自一线前沿实验室。

在 r/LocalLLaMA 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

Liquid AI发布了能够支持手机上128K上下文的2.6B模型

报道来源 [1]

  1. r/LocalLLaMA TIER_1 English(EN) · /u/BTA_Labs ·

    A 2.6B model with tool calling and 128K context now runs at 30 tok/s on a phone

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vfn9vc/a_26b_model_with_tool_calling_and_128k_context/"> <img alt="A 2.6B model with tool calling and 128K context now runs at 30 tok/s on a phone" src="https://preview.redd.it/xxbkpo9jcfhh1.jpeg?width=640&am…