3090
PulseAugur coverage of 3090 — every cluster mentioning 3090 across labs, papers, and developer communities, ranked by signal.
7 day(s) with sentiment data
-
Reddit user seeks ComfyUI workflow for upscaling 360 images
A user on Reddit is seeking advice on upscaling 360-degree images for VR headset viewing. They are specifically looking for a workflow compatible with a 3090 GPU and 64GB of RAM, aiming to improve image quality rather t…
-
User scales home AI cluster from 1 GPU to 20 DGX Sparks
A user details their journey of scaling up local AI model infrastructure, starting with a single 3090 GPU and progressing to a 20-GPU cluster using Nvidia GB10s. This evolution was driven by the desire to run increasing…
-
New AI models Cyber-Tiel-Coder and GLM-5.3-Flash showcased locally
Two separate AI model releases are highlighted on Mastodon. The first is Cyber-Tiel-Coder-35B-A3B, a 35 billion parameter Mixture-of-Experts (MoE) model designed for coding tasks and capable of running locally via GGUF …
-
Stable Diffusion user considers hardware downgrade for more RAM
A user on Reddit is seeking advice about potentially downgrading their PC hardware to increase RAM and VRAM for Stable Diffusion. They are considering switching from a system with a 5070 Ti and 32GB of DDR5 to one with …
-
User seeks SWEBench optimization tips for local LLM setup
A user on Reddit's r/LocalLLaMA subreddit is seeking advice on optimizing their SWEBench performance using llama.cpp and a quantized Qwen3.8-27B model. They have encountered numerous errors, including LimitExceeded and …
-
MiniMax H3 pose control method simplifies AI video generation
A new method for controlling character poses in AI-generated videos using MiniMax H3 has been developed, offering an alternative to traditional ControlNet approaches. This technique leverages outpainting on reference vi…
-
User seeks 7900 XTX performance data for Qwen 3.8 Flash Next LLM
A user is seeking advice on building a PC for local large language model (LLM) tasks, specifically inquiring about the performance of a 7900 XTX GPU with Qwen 3.8 Flash Next. They are comparing this setup to benchmarks …
-
AI system uses multiple models for distributed tasks
A user has developed a functional AI system that utilizes multiple language models for different tasks. A primary model, Qwen3.8, runs on a 3090 GPU for core functions like code generation and task distribution. Two sma…
-
User seeks advice on mixing NVIDIA 3090 with 5070/5060 Ti for local LLMs
A user on Reddit's r/LocalLLaMA forum is seeking advice on building a multi-GPU setup for running large language models locally. They are considering adding an NVIDIA 3090 to their existing setup of a 5070 Ti and two 50…
-
4-bit quantization enables large AI models on single 3090 GPU
A user on Mastodon shared their positive experience using 4-bit quantization for AI models, noting that a single 3090 GPU could fully accommodate a model with a 128k context window and the DFlash2 model. They reported i…
-
User downgrades GPU setup for Qwen3.8-27B, impacting generation speed
A user has reverted to a hardware configuration of 3090 and 3060 GPUs for the Qwen3.8-27B model, resulting in a decreased generation speed of 27 tokens per second. While this setup allows for a maximum context window, i…
-
User upgrades GPU setup for minor performance gains
The user physically relocated a 3090 graphics card to a higher slot in their computer. This change allowed the card to operate at x16 speed, resulting in a slight increase in prompt processing performance, while generat…
-
NInfer fork enables 2x performance boost for Qwen3.6-35B on CMP170HX hardware
A user has successfully forked the NInfer project to enable it to run on CMP170HX hardware, achieving a twofold performance increase for the Qwen3.6-35B model. This modification involved adjusting CUDA kernels and compi…
-
User seeks help optimizing llama.cpp for multi-GPU LLM inference
A user on Reddit's r/LocalLLaMA subreddit is seeking assistance with configuring llama.cpp to effectively utilize multiple GPUs for running large language models. They are experiencing issues with performance when attem…
-
Local LLM execution doesn't guarantee local data path
Running a large language model locally does not guarantee that all associated data remains within a controlled environment. Sensitive data can still be exposed through components like retrieval-augmented generation (RAG…
-
Dual 3090 GPUs achieve 1600+ tps with Qwen 3.6 27B by switching split modes
A user on Reddit's r/LocalLLaMA community shared their experience optimizing performance for the Qwen 3.6 27B model on a dual 3090 GPU setup. Initially, using `--split-mode tensor` resulted in prompt processing occurrin…
-
GPU thermal paste replacement offers significant cooling improvements
A user on the r/LocalLLaMA subreddit shared a public service announcement regarding the maintenance of older GPUs, specifically mentioning the 3090 model. They detailed how replacing the thermal paste on their GPU signi…
-
DSv4 Flash 0731 model runs on single NVIDIA 3090 GPU
A user has successfully run the DSv4 Flash 0731 model on a single, unoptimized NVIDIA 3090 graphics card. The setup includes an E5 2690v4 CPU and 3x32 DDR4 ECC RAM, with the 3090 operating at 250W. The model achieved a …
-
RTX 3090 owners share VRAM temps under AI load
Users on the r/LocalLLaMA subreddit are discussing the VRAM temperatures experienced by NVIDIA RTX 3090 graphics cards when running large language models. The original poster is seeking to determine if their GPU's tempe…
-
RTX Pro 4500: Evaluating Use Cases Against Higher-Power GPUs
The NVIDIA RTX Pro 4500 graphics card is being discussed for its potential use cases, particularly in comparison to other configurations like the RTX 5090 or multiple RTX 3090s. Users are questioning its value propositi…