PulseAugur
EN
LIVE 09:52:02

Motif Technologies releases Motif-3-Beta MoE model with 256K context

Motif Technologies has released a beta version of its large-scale Mixture-of-Experts (MoE) language model, Motif-3-Beta. This in-house developed model boasts approximately 314 billion total parameters with 13 billion active per token, and supports a native context length of 256,000 tokens. The model utilizes a sparse routing mechanism with 384 experts, activating 8 per token plus one shared expert. Instructions and code are available for integration with popular libraries like Transformers, and inference engines such as vLLM and SGLang, with support for Docker deployments. AI

IMPACT Sets new SOTA on long-context benchmarks; pressures competitors like Upstage and SKT.

RANK_REASON Frontier-lab model release with system card [lever_c_demoted from frontier_release: ic=2 ai=1.0]

Read on Hugging Face Trending Models →

AI-generated summary · Google Gemini · from 2 sources. How we write summaries →

Motif Technologies releases Motif-3-Beta MoE model with 256K context

COVERAGE [2]

  1. Hugging Face Trending Models TIER_1 English(EN) · Motif-Technologies ·

    Motif-Technologies/Motif-3-Beta

    text-generation · 0 downloads · 55 likes

  2. r/LocalLLaMA TIER_1 English(EN) · /u/Secure_Smoke_4280 ·

    Motif 3 Beta released

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1v23c6w/motif_3_beta_released/"> <img alt="Motif 3 Beta released" src="https://preview.redd.it/d4ayazdnbheh1.png?width=640&amp;crop=smart&amp;auto=webp&amp;s=7b4d61bf44de492b3aa4f76045f9197588512e34" title="Mo…