Moonlight-16B-A3B
PulseAugur coverage of Moonlight-16B-A3B — every cluster mentioning Moonlight-16B-A3B across labs, papers, and developer communities, ranked by signal.
2 day(s) with sentiment data
-
New technique preserves MoE routing structure for improved AI model performance
Researchers have introduced a new technique called Router Prior Bias (RPB) to improve the post-training performance of Mixture-of-Experts (MoE) models. Unlike standard methods that enforce uniform expert utilization, RP…
-
Instella-MoE: New open-source MoE language model released
A new technical report introduces Instella-MoE, an open-source Mixture-of-Experts (MoE) language model with 16 billion total parameters. Trained on AMD Instinct GPUs, the model incorporates innovations like Gated Multi-…
-
AMD releases open Instella-MoE-16B LLM with 2.8B active parameters
AMD has released Instella-MoE-16B-A3B, an open-source Mixture-of-Experts language model. This model features 16 billion total parameters but only activates 2.8 billion per token, utilizing architectural innovations like…
-
New research enhances LLM inference speed with advanced speculative decoding techniques · 8 sources tracked
Researchers are exploring advanced techniques to accelerate large language model (LLM) inference through speculative decoding. New methods like "Functional Reconstruction" aim to improve the agreement between draft and …