DGX Sparks
PulseAugur coverage of DGX Sparks — every cluster mentioning DGX Sparks across labs, papers, and developer communities, ranked by signal.
3 day(s) with sentiment data
-
DeepSeek releases v4.1 Flash model; MiniMax shares H3 open-source weights
DeepSeek has released version 4.1 of its Flash model, which can be run on 4x DGX Sparks hardware. Separately, MiniMax has released open-source weights for its H3 model, indicating rapid progress in its development.
-
AMD Epyc vs. DGX Sparks for LLM Clusters: A Cost-Performance Debate
A discussion on Reddit compares the cost-effectiveness of building a local large language model (LLM) cluster using NVIDIA DGX Sparks versus a system built around AMD Epyc processors. The user questions the logic of opt…
-
Homelab user expands GPU cluster to 36 DGX Sparks for 4.6TB unified memory
A user has expanded their homelab GPU cluster from 16 to 36 DGX Sparks, achieving 4.6TB of unified memory. This upgrade allows for simultaneous operation of state-of-the-art models like Kimi K3, alongside tasks such as …
-
GLM-5.2 NVFP4 achieves 24 tok/s at 128K context after bug fix · 1 source tracked
A user has resolved an issue with the GLM-5.2 NVFP4 model running on four DGX Sparks, achieving approximately 24 tokens per second at a 128K context length. The problem involved a bug in the speculative decoding configu…
-
DeepSeek V4-Flash achieves 40 Ttk/s on dual DGX Sparks
A user has shared configurations and benchmarks for running the DeepSeek V4-Flash model on dual DGX Sparks hardware. The setup achieves approximately 40 tera-tokens per second with FP8 precision, and can aggregate up to…
-
LLM enthusiast seeks rack for multiple AI compute devices
A user on the r/LocalLLaMA subreddit is seeking recommendations for a high-quality, all-metal rack or shelf to organize multiple small computing devices used for hosting large language models. The user currently has two…