Maple-Preview
PulseAugur coverage of Maple-Preview — every cluster mentioning Maple-Preview across labs, papers, and developer communities, ranked by signal.
1 day(s) with sentiment data
-
August 2026: AI Models Focus on Efficiency and Specialization · 1 source tracked
August 2026 saw a surge of new AI models, with a particular focus on efficiency and specialized capabilities. Many releases emphasized sparse architectures, enabling faster inference and lower active parameter counts. S…
-
New low-bit and ternary AI models released with performance updates · 1 source tracked
Several new low-bit and ternary models have been released and are being tracked, including Bonsai's 1-bit and 1.58-bit (ternary) versions, with a 27B parameter model now running on mainline backends. Updates to llama.cp…
-
llama.cpp, PyTorch, and new MoE model see significant updates
The llama.cpp project has released updates enhancing WebGPU acceleration and simplifying FlashAttention implementation for more efficient local LLM inference. Concurrently, PyTorch's MPSInductor now supports unsigned in…
-
OpenAI's Astra solves math problems; EU AI Act enforcement begins; fast mobile model released
OpenAI's internal model, codenamed Astra, has reportedly solved 10 long-standing mathematical and theoretical computer science problems, generating machine-checkable proofs for approximately $2,000 in compute. Concurren…
-
DeepGrove unveils Maple-Preview AI for iPhones, 13x faster than Bonsai 27B
AI research firm DeepGrove has announced Maple-Preview, a new AI model designed for efficient operation on mobile devices like the iPhone. This model boasts 13 times the processing speed of Bonsai 27B, another iPhone-co…
-
Maple-Preview Ternary Model Shows Promise Against Gemma 4
A user on Mastodon expressed intrigue regarding "Maple-Preview," a ternary model that they believe can surpass Gemma 4 in quality and speed on consumer hardware. The user is exploring ternary models themselves, suggesti…
-
20B MoE AI Model Runs at 120 Tokens/Sec on iPhone
A new demonstration, Maple-Preview, showcases a 20-billion parameter Mixture-of-Experts (MoE) model capable of running at 120 tokens per second on an iPhone. This development highlights the increasing efficiency and on-…
-
OpenAI flags Astra model as critical; Meta's Muse Spark shows gains · 4 sources tracked
OpenAI has escalated its Astra model to a "critical" cyber status due to advancements in agentic coding and cybersecurity, prompting stricter internal controls and a pause on non-essential activities. This move, alongsi…