PulseAugur
EN
LIVE 00:04:12

20B MoE AI Model Runs at 120 Tokens/Sec on iPhone

A new AI model named Maple-Preview, a 20 billion parameter Mixture of Experts (MoE) model, has been demonstrated running efficiently on an iPhone. The model achieved a speed of 120 tokens per second, showcasing its capability for on-device AI processing. AI

IMPACT Demonstrates the increasing feasibility of running sophisticated AI models directly on mobile devices.

RANK_REASON Demonstration of an AI model running on consumer hardware, not a core AI release from a frontier lab.

Read on Mastodon — fosstodon.org →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

20B MoE AI Model Runs at 120 Tokens/Sec on iPhone

COVERAGE [1]

  1. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    Show HN: Maple-Preview – ternary 20B MoE running at 120 tok/s on a iPhone https:// deepgrove.ai/maple-preview # ai # iphone

    Show HN: Maple-Preview – ternary 20B MoE running at 120 tok/s on a iPhone https:// deepgrove.ai/maple-preview # ai # iphone