NVIDIA DGX Spark
PulseAugur coverage of NVIDIA DGX Spark — every cluster mentioning NVIDIA DGX Spark across labs, papers, and developer communities, ranked by signal.
10 day(s) with sentiment data
-
Local LLM Hardware: GPUs for Small Models, Unified Memory for Large
For running large language models locally, the hardware landscape has divided into distinct categories based on memory capacity and speed. Consumer GPUs like the RTX 5090 excel with smaller models fitting within 32 GB, …
-
NVIDIA Earth-2 AI framework adapted for UK air pollution forecasting
Researchers at the University of Manchester have successfully adapted NVIDIA's Earth-2 AI framework, originally designed for weather forecasting, to model air pollution across the UK. This new application, Earth-2 CorrD…
-
User seeks advice on NVIDIA DGX Spark cluster size for DeepSeek 4.1 Flash
A user on the r/LocalLLaMA subreddit is seeking advice regarding the purchase of a third NVIDIA DGX Spark system. The user is inquiring whether the DeepSeek 4.1 Flash model can be effectively run on a two-node cluster o…
-
Perplexity launches local AI agent using Alibaba's Qwen model
Perplexity, a prominent AI startup valued at over $30 billion, has launched a new local agent product named Portable Computer. This product is built using Alibaba Group's latest open-source model, Qwen3.8-27B, and has b…
-
Microsoft launches AI-focused Windows 11 with high hardware demands · 2 sources tracked
Microsoft has released a specialized version of Windows 11, codenamed Project Zenith, designed to streamline AI development workflows. This stripped-down OS comes pre-loaded with essential developer tools and is optimiz…
-
China integrates AI tokens into consumer rewards, driving massive usage
Businesses in China are integrating AI tokens into consumer rewards programs, offering them with purchases of coffee, credit cards, and even food items like dumplings. These tokens represent units of computing power for…
-
Perplexity launches local AI assistant 'Portable Computer' for NVIDIA GPUs
Perplexity has launched "Portable Computer," a fully local version of its AI assistant that runs entirely on user hardware without cloud dependency. This new offering is now available on Linux for NVIDIA RTX GPUs with a…
-
Nvidia's PAIR tool pools idle home computers for local AI tasks · 10 sources tracked
Nvidia has released a free, open-source tool called Personal AI Router (PAIR) that allows users to leverage idle computers on their home network for local AI tasks. PAIR distributes AI workloads across multiple devices,…
-
MiniMax AI enables local high-quality video generation on consumer hardware
MiniMax AI has announced that its Fast H3 video generation model can now run locally on consumer hardware. This update brings high-quality video generation to devices utilizing NVIDIA DGX Spark or Apple Silicon through …
-
Beijing bar offers free AI model access with drink purchases
A new bar in Beijing's Zhongguancun tech hub, named AGI Bar, offers customers free access to AI models with the purchase of a drink. The bar provides tokens for various AI models and runs DeepSeek-V4-Flash locally on NV…
-
Qwen3.8-Flash-Next achieves 181 tokens/sec on dual DGX Spark cluster
A user on Reddit's r/LocalLLaMA subreddit shared their success in achieving a high inference speed of 181 tokens per second using the Qwen3.8-Flash-Next model on a dual-node NVIDIA DGX Spark cluster. This impressive per…
-
Perplexity launches local AI agent on Nvidia hardware, facing cost and complexity hurdles
Perplexity has launched Portable Computer, a local-first AI agent designed to reduce token costs and enhance user privacy by running on personal devices. This agent utilizes the Qwen 3.8 27B model from Alibaba, or a fin…
-
Perplexity launches local AI assistant 'Portable Computer' · 4 sources tracked
Perplexity has launched "Portable Computer," a fully local version of its AI assistant that runs entirely on a user's hardware without cloud dependency. This new feature is available to Perplexity Pro and Max subscriber…
-
Beijing AI bar loses money offering free DeepSeek tokens with drinks
An AI-themed bar in Beijing's Zhongguancun tech hub is offering unlimited free DeepSeek coding tokens with the purchase of a $1.50 drink. The bar, named AGI Bar, is reportedly losing money due to this promotion, with ap…
-
Qwen 3.8 27B LLM praised for capabilities but criticized for overthinking default
Alibaba's Qwen research lab has released Qwen 3.8 27B, an open-source, vision-capable LLM. While praised for its capabilities and size, suitable for local hardware, users are reporting that its default setting for reaso…
-
Simon Willison builds GPT-5.6-Sol xhigh chat testing tool
Simon Willison developed a tool to test chat endpoints compatible with OpenAI's API, utilizing GPT-5.6-Sol xhigh. This tool, which provides a web UI, has been successfully tested with LM Studio running Qwen 3.8-27B on b…
-
LTX-2.5 open world model enables local AI video production on NVIDIA GPUs
LTX-2.5, a new open-weights world model, has been released, enabling creators to perform video generation and other AI tasks on local NVIDIA RTX GPUs. This model significantly reduces VRAM requirements, making advanced …
-
Muse Glimmer 30B model context extended to 1M tokens with perfect retrieval
A user has successfully extended the context window of the Muse Glimmer 30B model to 1 million tokens, significantly surpassing its trained 131K context length. This was achieved using the YaRN context extension method …
-
NVIDIA DGX Spark testbed enables distributed LLM training and CTI fine-tuning
Researchers have developed a remote-access testbed for distributed LLM training using two NVIDIA DGX Spark systems connected via Tailscale VPN and a direct fiber link. This setup enabled the distributed pretraining of a…
-
inclusionAI releases lightweight Ling-3.0-tiny MoE model for local deployment
inclusionAI has released Ling-3.0-tiny, a new hybrid reasoning Mixture-of-Experts (MoE) model with 7.9 billion total parameters and 1.3 billion activated parameters per token. This model is designed for efficient local …