PulseAugur
EN
LIVE 10:45:40

Motif Technologies unveils 314B parameter Motif 3 LLM

Motif Technologies has released Motif 3, a decoder-only Mixture-of-Experts language model with 314 billion total parameters and 13.2 billion activated per token. The model features a novel Grouped Differential Latent Attention architecture and was trained on approximately 12.5 trillion tokens. Motif 3 demonstrates competitive performance against leading open-weight models, particularly excelling in long-horizon agentic tasks, mathematical reasoning, and scientific knowledge. AI

IMPACT Sets a new benchmark for open-weight models, particularly in agentic tasks and reasoning, potentially influencing future model development.

RANK_REASON Technical report and associated materials for a new large language model release.

Read on Hugging Face Daily Papers →

AI-generated summary · Google Gemini · from 5 sources. How we write summaries →

Motif Technologies unveils 314B parameter Motif 3 LLM

COVERAGE [5]

  1. arXiv cs.AI TIER_1 English(EN) · Junghwan Lim, Joon Son Chung, Sungmin Lee, Wai Ting Cheung, Gihun Cho, Minsu Ha, Sangho Kang, Beomgyu Kim, Dongseok Kim, Jangwoong Kim, Taehyun Kim, Taewhan Kim, Jeesoo Lee, Jeongdoo Lee, Junhyeok Lee, Dongpin Oh, Hyeyeon Cho, Dahye Choi, Jaeheui Her, Ha… ·

    Motif 3: Technical Report

    arXiv:2608.09119v1 Announce Type: new Abstract: We introduce Motif 3, a decoder-only Mixture-of-Experts language model with 314 billion total parameters and 13.2 billion activated per token. Each sparse MoE layer contains 384 routed experts, with eight selected per token. This fi…

  2. Hugging Face Daily Papers TIER_1 English(EN) ·

    Motif 3: Technical Report

    We introduce Motif 3, a decoder-only Mixture-of-Experts language model with 314 billion total parameters and 13.2 billion activated per token. Each sparse MoE layer contains 384 routed experts, with eight selected per token. This fine-grained sparsity provides substantial expert …

  3. Hugging Face Daily Papers TIER_1 English(EN) ·

    Motif 3: Technical Report

    We introduce Motif 3, a decoder-only Mixture-of-Experts language model with 314 billion total parameters and 13.2 billion activated per token. Each sparse MoE layer contains 384 routed experts, with eight selected per token. This fine-grained sparsity provides substantial expert …

  4. Hugging Face Trending Models TIER_1 English(EN) · Motif-Technologies ·

    Motif-Technologies/Motif-3

    text-generation · 0 downloads · 72 likes

  5. r/LocalLLaMA TIER_1 English(EN) · /u/Lucidstyle ·

    Motif-Technologies/Motif-3 official realese

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vkl6cs/motiftechnologiesmotif3_official_realese/"> <img alt="Motif-Technologies/Motif-3 official realese" src="https://external-preview.redd.it/qWM7gf23D_GVD3xaqqk_1gvttXm2279-V9JFKzzDRDw.png?width=640&amp;cr…