Benchmark menu

Claude Fable 5.1

Anthropic · Overall Models

Score · relative 0–100
81.1
Rank
2
Status
Complete 3/3
Rank change
Baseline
Last tested
2026-09-10
Settings tested
All three benchmark settings tested

Scores by reasoning setting

Tested settingScoreResult statusDisplay group
Low73.8Tested and includedBaseline
Medium77.2Tested and includedBalanced
XHigh92.4Tested and includedIntensive
Available settings and grouping

Available settings

LowMediumHighXHighMax

Checked on 2026-09-08 · Provider documentation

How settings are grouped

  • Baseline: Low — tested and included
  • Balanced: Medium — tested and included
  • Intensive: XHigh — tested and included

Player program results

XHigh: 82% / 18% / 0% · n 11Medium: 88% / 12% / 0% · n 8Low: 100% / 0% / 0% · n 8 All three benchmark settings tested

Generation cost

Average estimated cost
$2.32
Median estimated cost
$1.13
Estimated range
$0.66–$11.79
Cost data
24 combinations / 27 recorded runs
Observed total
$64.31
Price source
Standard list-price estimate
Price list
boardgame-list-prices-2026-08-08 + boardgame-list-prices-2026-09-08
Average output
47k tokens
Output range
5.6k–231k tokens
Samples
32

Estimated cost gives equal weight to each model, game, and reasoning-setting combination.

Test setup

  • Combines this model's tested reasoning settings
  • Uses the standard GameBench player-program prompt