Motif Technologies has released Motif 3, a decoder-only Mixture-of-Experts language model with 314 billion total parameters and 13.2 billion activated per token. The model features a novel Grouped Differential Latent Attention architecture and was trained on approximately 12.5 trillion tokens. Motif 3 demonstrates competitive performance against leading open-weight models, particularly excelling in long-horizon agentic tasks, mathematical reasoning, and scientific knowledge. AI
IMPACT Sets a new benchmark for open-weight models, particularly in agentic tasks and reasoning, potentially influencing future model development.
RANK_REASON Technical report and associated materials for a new large language model release.
Read on Hugging Face Daily Papers →
- A.X Series
- EXAONE Series
- LG AI Research
- Motif Technologies
- Qwen 3.7 Max
- Solar Series
- Upstage
- arXiv
- Expert Specific PolyNorm
- Grouped Differential Latent Attention
- Hugging Face
- Motif 3
- Multi-Teacher On-Policy Distillation
- MXFP8
AI-generated summary · Google Gemini · from 5 sources. How we write summaries →