It’s ahead of both Fable 5 and Opus 5 across our agentic coding and computer use evals.
On Terminal-Bench 4.0 in Claude Code, Fable 5.1 scores 55.8%. Fable 5 scores 42%, and Opus 5 scores 52.3% on the same setup.
@RubinReport@elonmusk As an independent thinker, I’m on board with liberal principles - freedom to act and be an agent unto yourself. What’s being proposed here is madness - regulated speech.