PulseAugur
EN
LIVE 11:42:43

POCKET-35B model runs on PC and phone without GPU

A new 35-billion parameter language model called POCKET-35B has been released, designed to run efficiently on consumer hardware, including PCs without dedicated GPUs and even smartphones. The model is available in various quantized versions, with file sizes ranging from 5.1 GB to 21 GB, allowing users to select the best fit for their device's RAM and performance needs. Notably, POCKET-35B can operate using standard llama.cpp software without requiring custom forks or cloud infrastructure. AI

IMPACT Enables local execution of a capable agentic model on consumer devices, reducing reliance on cloud infrastructure.

RANK_REASON Release of a specific model that runs on consumer hardware.

Read on r/LocalLLaMA →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

POCKET-35B model runs on PC and phone without GPU

COVERAGE [1]

  1. r/LocalLLaMA TIER_1 English(EN) · /u/Powerful_Evening5495 ·

    POCKET-35B agentic model on cpu 59 t/s

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1v6zseq/pocket35b_agentic_model_on_cpu_59_ts/"> <img alt="POCKET-35B agentic model on cpu 59 t/s" src="https://preview.redd.it/xbyvic9dtjfh1.png?width=640&amp;crop=smart&amp;auto=webp&amp;s=743828d48949ee01cd1…