A user details their journey of scaling up local AI model infrastructure, starting with a single 3090 GPU and progressing to a 20-GPU cluster using Nvidia GB10s. This evolution was driven by the desire to run increasingly powerful models like DeepSeek 671B MoE, Qwen 235B, and later MiMo 2.5 Pro and Kimi 2.6, for coding tasks and agentic applications. The user encountered significant challenges with power limitations, leading to blown fuses, and eventually collaborated with their brother to combine clusters for even larger model deployments, prioritizing local control and privacy over commercial subscriptions. AI
IMPACT Demonstrates the increasing feasibility and challenges of running advanced local AI models for professional use.
RANK_REASON User's personal infrastructure build and scaling journey, not a company or lab announcement.
- 3090
- ASUS GB10
- DeepSeek 671B MoE
- GLM 5.3
- Kimi 2.6
- Kimi K3
- LLaMA 33B
- LLaMA 65B
- MiMo 2.5 Pro
- MiMo 2.6 Pro
- MiniMax M2
- MOD St Athan
- Nvidia
- OpenCode
- OpenWebUI
- Qwen
- Qwen 235B
- Qwen 397B
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →