PulseAugur
实时 09:43:12
English(EN) Cactus Compute has released Needle 2, an open 45M-parameter tool-calling model that runs in a 14MB binary using just 28MB RAM. It targets hardware with no GPU a

Cactus Compute 发布 Needle 2,一款面向边缘设备的 14MB 工具调用 AI 模型

Cactus Compute 推出了 Needle 2,这是一款紧凑型的 4500 万参数模型,专为工具调用和结构化数据提取而设计。该模型以其极小的占用空间而闻名,打包为 14MB 的二进制文件,运行仅需 28MB RAM,使其适用于没有专用 AI 硬件(如 GPU 或 NPU)的设备。Needle 2 实现了令人印象深刻的速度,在 Raspberry Pi 5 上每秒可处理多达 500 个 token,并在各种移动和混合现实设备上提供高效性能。 AI

影响 在低功耗、资源受限的设备上实现复杂的人工智能能力,将人工智能应用扩展到传统的云端或高端硬件之外。

排序理由 来自专注于高效推理的专业 AI 实验室的模型发布。 [lever_c_demoted from frontier_release: ic=2 ai=1.0]

在 Mastodon — mastodon.social 阅读 →

AI 生成摘要 · Google Gemini · 来自 2 个来源。 我们如何撰写摘要 →

Cactus Compute 发布 Needle 2,一款面向边缘设备的 14MB 工具调用 AI 模型

报道来源 [2]

  1. MarkTechPost TIER_1 English(EN) · Michal Sutter ·

    认识 Needle 2:一个开源的 4500 万参数工具调用模型,打包成 14MB 二进制文件,仅需 28MB RAM 即可运行完整会话

    <p>Cactus Compute released Needle 2, an open 45M-parameter model for tool calling, device use, and structured extraction. The full model is a single 14MB binary that runs a session in about 28MB of RAM. It leads both Seal-Tools splits while targeting hardware with no GPU and no N…

  2. Mastodon — mastodon.social TIER_1 English(EN) · [email protected] ·

    Cactus Compute 发布 Needle 2,一款开源的 4500 万参数工具调用模型,可在 14MB 二进制文件中运行,仅需 28MB RAM。它面向无 GPU 的硬件

    Cactus Compute has released Needle 2, an open 45M-parameter tool-calling model that runs in a 14MB binary using just 28MB RAM. It targets hardware with no GPU and achieves 500 tokens/sec on a Raspberry Pi 5. https://www. marktechpost.com/2026/08/13/ca ctus-compute-needle-2-45m-pa…