
Fable 5.1 Tops Opus 5 on Terminal-Bench 4.0
On Terminal-Bench 4.0 inside Claude Code, Fable 5.1 scores 55.8%, beating both Fable 5 (42%) and Opus 5 (52.3%). It leads across agentic coding and computer use evals.

On Terminal-Bench 4.0 inside Claude Code, Fable 5.1 scores 55.8%, beating both Fable 5 (42%) and Opus 5 (52.3%). It leads across agentic coding and computer use evals.
Fable 5.1 is live in Claude Code and the Claude Platform. Same price as Fable 5, with 75% cheaper cache reads. It handles longer tasks before needing user input, better signals when it's stuck, and writes more naturally.
All users get fresh 5-hour and weekly usage limits reset alongside the Fable 5.1 launch today.
Fable 5.1 leads across coding, data analysis, computer use, design, presentations, and long-running agentic tasks. The Claude Code team calls it a pleasure to work with and describes it as their best model yet.
An Anthropic demo developer used Fable 5.1 to build a Mac app that tracks overhead flights and announces their position out loud, describing the model as one that 'just gets it' without much back-and-forth.
A turn-by-turn navigation app for the Moon built on actual NASA terrain data, rendered as a 3D globe with no libraries, routes between craters while steering around the largest ones and explaining each decision.
Adocomplete prompted Fable 5.1 to create and animate a realistic industrial robot, inspired by a house-walkthrough demo. The resulting video shows a fully animated 3D robot built end-to-end by the model.
A user on a 20x Pro plan hit their 5-hour session limit after roughly 45 minutes of coding across three chats, also consuming 38% of their weekly Fable allocation. A team member asked them to run /feedback in the affected sessions and share the feedback ID.
A user running three high-effort and two medium-effort Fable 5.1 threads burned through 12% of their weekly limit and 40% of their 5-hour limit in 18 to 20 minutes. A team member reached out via DM.
A user reported their limits disappeared in a single prompt, flagging it as likely a bug. A team member asked them to run /feedback in the affected session and share the resulting ID.
A user reported not seeing Fable 5.1 in the Claude CLI. The fix is a config cache freshness issue; restarting Claude resolves it.
A complete industrial robot, including creation, animation, and rendering, came in at about 115k tokens via a single Fable 5.1 prompt. adocomplete thinks game development is about to get a major boost from this capability.
Fable 5.1 and Mythos 5.1 are out. Key improvements: the model goes deeper into tasks before asking for help, admits when it is stuck rather than faking progress, and fixes root causes instead of reaching for quick patches. Cache reads are 75% cheaper; base price unchanged.
Cache read pricing for Fable 5.1 falls from $1 to $0.25 per million tokens for Enterprise, API, and SDK customers. A typical Claude Code session works out roughly 38% cheaper as a result.
The latest biology safeguards intervene on benign requests 85% less often than those shipped with Fable 5. Claude Code users can expect around 60% fewer cyber interventions per session, with further improvements on the way.