Claude Fable 5
Anthropic's Mythos-class flagship — state-of-the-art on nearly every tested benchmark.
- Context
- 1M
- In / 1M
- $10
- Out / 1M
- $50
Every frontier API model on one board: live per-token pricing, verified benchmark scores with sources, context windows — and for each one, what it would actually take to run something comparable on your own hardware.
Anthropic's Mythos-class flagship — state-of-the-art on nearly every tested benchmark.
Previous Anthropic flagship — still elite, especially on science reasoning.
Google's reasoning flagship — top of the GPQA Diamond leaderboard.
Google's speed/quality sweet spot — near-flagship scores at Flash pricing.
The default workhorse Claude — 60% of Opus quality questions at 60% of the price.
The open-weights flagship — AA index 52 at one-tenth of flagship prices.
GPT-5.5 with extended parallel compute — 6× the price for the hardest problems.
How to read these numbers. Pricing and context windows come from a live OpenRouter snapshot and are re-verified weekly. Benchmark conventions: Intelligence is a composite index (Artificial-Analysis style), Coding is SWE-bench Verified %, Reasoning is GPQA Diamond %. Every score cites its source on the model's datasheet — and where a number isn't independently verified, we show n/a rather than guessing. Blended price assumes a 3:1 input:output token mix.