Anthropic’s new flagship, Fable 5—the public, guardrailed sibling of its “Mythos”-class system—is the best model in the world by a substantial but not shocking margin. The more striking claim is that its edge grows as tasks lengthen and stiffen.
Mechanically it ships with thinking always on, governed by an “effort” dial from low to “xhigh,” and safety classifiers for cyber, bio, chemical and distillation risks that auto-route dangerous prompts down to Claude Opus 4.8. Access is steep: $10 and $50 per million input/output tokens, double Opus, with 30-day retention.
The scoreboard is lopsided—72.9% on Cursor Bench to GPT-5.5’s 64.3, 94% GPQA Diamond, 99.8% USAMO 2026, 87–88% FrontierMath, and 55% on RiemannBench where Opus managed 34. It ranks first on Agent Arena, ProofBench and the Debate Benchmark, and beat Pokémon FireRed by vision alone.
The honest caveat: Fable 5 wins “by being right more, not wrong less.” It shows weaker steerability and a worse position bias, picking the first option 59% of the time.