PulseAugur
EN
LIVE 16:36:12

Optimized Ornith1.5 35B A3B model shows significant speed improvements

A user on r/LocalLLaMA has modified the Ornith1.5 35B A3B model, achieving a 33% reduction in task completion time and a 3% increase in tokens per second. This optimization involved splicing a trained MTP head from another model onto the Ornith1.5 base. The user reports that this enhanced version is faster and more accurate than previous iterations, even performing well in specific applications like controlling a HAM radio rig without generating lectures. AI

IMPACT Demonstrates potential for community-driven performance enhancements in open-source LLMs.

RANK_REASON User-driven optimization of an existing open-source model, not a frontier release from a major lab.

Read on r/LocalLLaMA →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

Optimized Ornith1.5 35B A3B model shows significant speed improvements

COVERAGE [1]

  1. r/LocalLLaMA TIER_1 English(EN) · /u/frankentriple ·

    Fixed the MTP head on Ornith1.5 35B A3B. +3% TPS -33% wall clock

    <!-- SC_OFF --><div class="md"><p>I love the Ornith 35B local models, 1.0 has been running my HAM radio rig for me. I have a hackRF receiver and a 5 watt quansheng portable the both run headless through the PC. I tried out the new Ornith1.5 build and it was faster and more accura…