A user on the r/LocalLLaMA subreddit is seeking recommendations for the fastest and most effective "abliterated" (safety-removed) versions of Qwen 27B models, specifically comparing versions 3.6 and 3.8. The user intends to test these models on instruction-following benchmarks with "thinking off" to determine which version performs better for tasks requiring less complex reasoning. They are looking for user favorites and fast quantizations suitable for their system with 24GB of VRAM, noting that some tested models have been too slow. AI
IMPACT Provides insights into user preferences and performance expectations for fine-tuned open-source models.
RANK_REASON User discussion and request for recommendations on specific model variants.
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →