LLMs1 min read
Claude Fable 5.1 introduces multiple reasoning levels and improved benchmarks
Claude Fable 5.1 achieves a 52.6% score on the Terminal-Bench-Science 0.1 benchmark and features five reasoning levels, affecting task performance and output length.
From Simon Willison
