Anthropic has released Claude Fable 5.1, a new model that reportedly sets a higher standard for coding, knowledge work, and long-running problem-solving tasks. The model achieved a 52.6% score on the new Terminal-Bench-Science 0.1 benchmark, a significant improvement over previous versions and competitors like GPT 5.6 "Sol". While initial tests showed minimal reasoning at lower effort settings, higher settings produced increasingly detailed and complex outputs, with the 'max' effort setting generating a highly detailed animated pelican SVG, though still not matching the flair of Gemini 3.7 Flash. AI
IMPACT Sets a new standard for coding and scientific tasks, potentially influencing enterprise adoption and competitive benchmarks.
RANK_REASON Frontier-lab model release with system card. [lever_c_demoted from frontier_release: ic=2 ai=1.0]
Read on Mastodon — mastodon.social →
- Anthropic
- Claude Fable 5.1
- Gemini 3.7 Flash
- GPT 5.6 "Sol"
- Python
- Simon Willison
- Terminal-Bench-Science 0.1
AI-generated summary · Google Gemini · from 2 sources. How we write summaries →