PulseAugur
EN
LIVE 12:40:53

Ternary LLMs see resurgence with new models from smaller labs

Recent developments suggest a resurgence in ternary (1.58-bit) large language models, with several new models released by smaller labs. These include prismML's 27B ternary model, Deepgrove's 20B Maple model, and Doses AI's 27B Pestle model specialized for medical applications. While these models show promise, particularly in terms of speed and performance on specific tasks, they currently face challenges with long-horizon agentic tasks, which developers aim to address through future reinforcement learning optimization. AI

IMPACT Potential for more efficient LLMs, though current implementations face limitations in complex agentic tasks.

RANK_REASON Discussion of a specific, less common model architecture (ternary LLMs) and new releases based on it. [lever_c_demoted from research: ic=1 ai=1.0]

Read on r/LocalLLaMA →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

Ternary LLMs see resurgence with new models from smaller labs

COVERAGE [1]

  1. r/LocalLLaMA TIER_1 English(EN) · /u/Individual-Dot5488 ·

    Is ternary (1.58-bit) LLMs making a come back?

    <!-- SC_OFF --><div class="md"><p>I'm just thinking, ever since microsoft announced bitnet, this sub (and myself) has been hoping for massive ternary models. In the last month alone, prismML dropped 27B ternary (though I've read community experience suggested it sometimes didn't …