mixture of experts model
PulseAugur coverage of mixture of experts model — every cluster mentioning mixture of experts model across labs, papers, and developer communities, ranked by signal.
-
Oxmiq Labs proposes High Bandwidth Flash for AI inference capacity
Oxmiq Labs is proposing High Bandwidth Flash (HBF) as a new capacity tier for AI inference, aiming to offer significantly more storage at a comparable cost to High Bandwidth Memory (HBM). Their presentation at Hot Chips…
-
16GB GPU can run 20GB MoE models at 100 tokens/sec
New community testing indicates that a 20GB mixture of experts (MoE) model can be run on a GPU with only 16GB of VRAM, achieving speeds of approximately 100 tokens per second. This suggests that advanced AI models are b…
-
Hardware query for running Qwen 3.5 122B MoE model
A user on Reddit's r/LocalLLaMA community is inquiring about the hardware requirements for running a large mixture of experts (MoE) model, specifically Qwen 3.5 122B. The user is asking for practical results or experien…
-
Time Series Models Evaluated for US Influenza Forecasting
A new research paper evaluates various time series forecasting models for predicting seasonal influenza in the United States. The study found that a mixture-of-experts model, which combines multiple pretrained forecaste…