PulseAugur
EN
LIVE 00:00:25

User seeks AI model recommendation for 4x 3090 GPU setup

A user with a server equipped with four NVIDIA 3090 GPUs and 96GB of VRAM is seeking recommendations for which AI model to run. They plan to use the server as a personal AI playground for idea manifestation and prototyping, with potential for a few additional users. The user is also considering dedicating two GPUs to speech and document processing services for a voice agent. AI

RANK_REASON User-generated content asking for advice on running a specific model (Hermes) on consumer hardware (4x 3090 GPUs).

Read on r/LocalLLaMA →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

User seeks AI model recommendation for 4x 3090 GPU setup

COVERAGE [1]

  1. r/LocalLLaMA TIER_1 English(EN) · /u/Sea_Calendar_3912 ·

    4x 3090, 96gb vram what Model to drive Hermes?

    <!-- SC_OFF --><div class="md"><p>3 year lurker, now i finally got my server up and running. dont know which model to choose. llama.cpp or vllm, what makes more sense? mainly single user with maybe 2-3 more additional users in family, if everything checks out. hermes is gonna be …