PulseAugur
EN
LIVE 18:42:18

Liquid AI boosts LFM2.5 model speed up to 3.18x with DSpark speculative decoding

Liquid AI has released DSpark draft models for its LFM2.5 series, which enhance decoding speed by up to 3.18x without altering output quality. These models utilize speculative decoding, where a smaller draft model proposes token candidates that a larger target model verifies. This approach significantly speeds up inference, particularly for agentic applications where latency is critical. The models are available for self-hosting and are compatible with tools like llama.cpp and SGLang. AI

IMPACT Accelerates inference for smaller models, potentially enabling more complex AI applications on edge devices.

RANK_REASON Liquid AI is a frontier lab releasing new draft models with a novel speculative decoding technique. [lever_c_demoted from frontier_release: ic=2 ai=1.0]

Read on MarkTechPost →

AI-generated summary · Google Gemini · from 2 sources. How we write summaries →

Liquid AI boosts LFM2.5 model speed up to 3.18x with DSpark speculative decoding

How we ranked this

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Significant
Liquid AI is a frontier lab releasing new draft models with a novel speculative decoding technique. [lever_c_demoted from frontier_release: ic=2 ai=1.0]
Source corroboration
2 independent sources
Multiple independent publishers reporting the same story raises confidence that it's real and newsworthy.
Topics
model release, product
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
47 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.

Full methodology in our editorial standards.

COVERAGE [2]

  1. MarkTechPost TIER_1 English(EN) · Asif Razzaq ·

    Liquid AI Releases LFM2.5-DSpark Draft Models That Deliver Up to 3.18x Faster Decoding Without Changing Model Outputs

    <p>Three ~300M drafters bring speculative decoding to LFM2.5, delivering up to 3.18x faster decoding with identical greedy output.</p> <p>The post <a href="https://www.marktechpost.com/2026/08/20/liquid-ai-releases-lfm2-5-dspark-draft-models-that-deliver-up-to-3-18x-faster-decodi…

  2. Mastodon — mastodon.social TIER_1 日本語(JA) · [email protected] ·

    A version of speculative decoding technology DSpark that speeds up the small model "LFM2.5" by more than double has appeared https://fed.brid.gy/r/https://gigazine.net/news/20260821-lfm2-5-dspark-faster-inference/

    小型モデル「LFM2.5」を2倍以上高速化する投機的デコーディング技術DSpark適用版が登場 https:// fed.brid.gy/r/https://gigazine .net/news/20260821-lfm2-5-dspark-faster-inference/