Users are reporting inconsistent performance and "low IQ" behavior from certain Claude models, specifically GPT 5.6 Sol and GPT 6 Astra, which tend to over-engineer sub-projects without explanation or abandon main tasks for side questions. In contrast, Fable 5 and Opus 4.8 are noted for their ability to resume main tasks after answering side questions and for narrating their intentions. These experiences highlight potential differences in how models interpret global instructions and prompting styles, leading to varied user satisfaction. AI
IMPACT Highlights potential inconsistencies in large language model instruction following and prompting effectiveness.
RANK_REASON User discussion about model performance and prompting styles.
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →