ENTITY
MoE models
MoE models
PulseAugur coverage of MoE models — every cluster mentioning MoE models across labs, papers, and developer communities, ranked by signal.
Total · 30d
1
2 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
1
1 over 90d
TIER MIX · 90D
TOPICS
SENTIMENT · 30D
1 day(s) with sentiment data
RECENT · PAGE 1/1 · 2 TOTAL
-
xHC method expands transformer streams beyond N=4 for improved LLM pre-training
Researchers have introduced xHC (Expanded Hyper-Connections), a novel method for scaling transformer models beyond the typical limit of N=4 streams. This new approach addresses bottlenecks in previous Hyper-Connections …
-
Budget GPU advice sought for local LLM inference
A user on the r/LocalLLaMA subreddit is seeking advice on purchasing hardware for running large language models on a limited budget. They are considering either a Radeon VII with 32GB VRAM or two P100 GPUs offering a co…