Abliterlitics has released a comparative analysis of eight uncensored variants of the Qwen 3.8-27B model, alongside the base model. The study involved 167 hours of GPU computation and evaluated models on benchmarks, KL divergence, and refusal rates using HarmBench. The 'orcarouter' variant emerged as the top performer with an 82.2% attack success rate, while 'apostate' offered the best value with near-identity capabilities and the lowest KL divergence. Conversely, the 'obliteratus' variant was flagged for avoidance due to significant performance degradation and thinking loops, with the base model showing minimal compliance. AI
IMPACT Provides insights into the performance and safety trade-offs of uncensored LLM variants, guiding users toward more capable or safer options.
RANK_REASON Analysis and comparison of multiple variants of an existing open-source model. [lever_c_demoted from research: ic=1 ai=1.0]
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →