A user on Reddit compared the performance of Anthropic's Claude 3 models (Opus, Sonnet, Haiku) and OpenAI's GPT-4 on a creative task involving animating a goat riding a tractor. The user found that Claude 3 Sonnet performed surprisingly well with maximum effort settings, while Claude 3 Haiku (referred to as Fable) was the least impressive. The user spent additional credits to test Haiku further, noting that Claude models tend to make improvements even when not explicitly asked, which required specific instructions to prevent. AI
IMPACT Highlights varying capabilities of different LLMs in creative generation, suggesting Sonnet as a strong contender for such tasks.
RANK_REASON User-generated comparison of AI model performance on a creative task.
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →