A Reddit user has compiled a ranked list of large language models that can be run on consumer hardware with 10-16GB of VRAM. The evaluation focused on speed, coherency, language handling, 'slop patterns' (frequency of descriptive phrases), model architecture (MoE vs. full), censorship levels, and instruction following capabilities. The user noted that larger models (above 200B parameters) are necessary for comprehensive knowledge, particularly in niche media topics, but are generally inaccessible to most users. AI
IMPACT Provides a practical guide for users with limited hardware to select and run capable LLMs.
RANK_REASON User-generated ranking and evaluation of existing models, not a new release or significant industry event.
- alpaca
- Claude Opus
- CodeLlama
- Gemma
- GLM-4.7-Flash
- Llama 3
- Mistral AI
- Mixtral
- Orca
- Phi 3
- Qwen
- Starling
- Vicuña
- Zephyr
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →