CheckNet.NETWORK DIAGNOSTICS / TOOLKIT
WorkspaceAI Model TrackerNETWORK TOOLKIT
All models

Anthropic · launched 1 Sept 2026

Claude Fable 5.1

Released alongside the access-limited Claude Mythos 5.1. Cache reads cost 75% less than Fable 5, which Anthropic estimates cuts typical costs by about 25%.

Read the official announcement
Model details
Vendor
Anthropic
Launch date
1 Sept 2026
Input price
$10 / 1M tokens
Output price
$50 / 1M tokens
Cached input
$0.25 / 1M tokens
Batch discount
50%
Context window
1M tokens
Max output
128K tokens
Input types
Text, Image, File
API model ID
claude-fable-5-1
OpenRouter ID
anthropic/claude-fable-5.1
Compared with Claude Fable 5

Input price

$10

was $10 · Same price

Output price

$50

was $50 · Same price

Blended price

$20

was $20 · Same price

Best published results
  • Terminal-Bench-Science 0.1max effort · $37.9 per task · reported by Anthropic52.6%
  • Humanity's Last Exam (tools)xhigh effort · $2.28 per task · reported by Anthropic65.08%
  • Humanity's Last Exam (no tools)max effort · $2.23 per task · reported by Anthropic60.92%
  • CursorBench 3.2.0max effort · $9.64 per task · reported by Anthropic73.4%
  • AutomationBench 1.0.6max effort · $2.45 per task · reported by OpenAI31.4%
  • FrontierCode 1.1medium effort · $3.28 per task · reported by OpenAI50.9%
  • Terminal-Bench 4.0max effort · $19.5 per attempt · reported by Anthropic55.8%
  • FrontierCode v1.1low effort · $2.47 per task · reported by Anthropic52.8%
  • CursorBench 4.0max effort · $17.3 per task · reported by Anthropic51.8%
  • GDPval-AA v2.1max effort · $9.59 per task · reported by Anthropic1735 Elo
  • WANDRmax effort · $49 per attempt · reported by Anthropic68.7%
  • Terminal-Bench 4.0max effort · $19.5 per task · reported by OpenAI55.8%

Price compared with other models

ModelLaunchedInput / 1MOutput / 1MCached / 1MBlended / 1Mvs Claude Fable 5.1
GPT-6 LunaOpenAI22 Sept 2026$0.10$0.50$0.01$0.2099% cheaper
GPT-5.6 LunaOpenAI9 Jul 2026$0.20$1.20$0.02$0.4598% cheaper
Claude Sonnet 5Anthropic30 Jun 2026$2.00$10$0.20$4.0080% cheaper
Claude Sonnet 5.5Anthropic28 Sept 2026$2.00$10$0.20$4.0080% cheaper
GPT-6 SolOpenAI22 Sept 2026$2.00$10$0.20$4.0080% cheaper
Claude Opus 5.5Anthropic22 Sept 2026$4.00$20$0.20$8.0060% cheaper
GPT-5.6 SolOpenAIPrice before the GPT-6 launch, from OpenAI's announcement. OpenRouter now lists $2 / $10.9 Jul 2026$4.00$20–$8.0060% cheaper
Claude Opus 5Anthropic24 Jul 2026$5.00$25$0.50$1050% cheaper
Claude Fable 5Anthropic · previous generation9 Jun 2026$10$50$1.00$20Same price
Claude Fable 5.1Anthropic1 Sept 2026$10$50$0.25$20Baseline
GPT-6 AstraOpenAI4 Sept 2026$10$50$1.00$20Same price

USD per million tokens, checked 2 Oct 2026 via OpenRouter. Comparison uses the blended price (3 input tokens for every output token).

Official benchmark results

Claude Fable 5.1 launch results

Reported by Anthropic on 1 Sept 2026 · exact values

BenchmarkClaude Fable 5.1Claude Fable 5Claude Opus 5GPT-5.6 Sol
Terminal-Bench-Science 0.1Agentic scientific research52.6%24.7%29%22.4%
Terminal-Bench 4.0Agentic coding55.8%42%52.3%37.3%
GDPval-AA v2Knowledge work1853 Elo1723 Elo1824 Elo1711 Elo
OSWorld 2.0Computer use · partial77.9%72.9%75.4%–
OSWorld 2.0 (strict)Computer use · strict41.7%36.1%39.6%–
Humanity's Last Exam (no tools)Multidisciplinary reasoning · no tools60.9%57.8%56.6%–
Humanity's Last Exam (with tools)Multidisciplinary reasoning · with tools65%63.8%63.6%–
AutomationBenchBusiness workflows31.4%17.1%26.9%19.6%
CursorBench 3.2.0Agentic coding73.4%70.5%70%67.2%
  • Bold marks the best reported score in each row.
  • Fable 5.1 was evaluated with production safeguards enabled. Where they intervened, Fable 5.1 and Fable 5 scored zero on OSWorld 2.0 and Fable 5 scored zero on AutomationBench.
  • Claude Mythos 5.1, the same underlying model with access limited to vetted users, scores 60.9% on Terminal-Bench 4.0. Mythos has no public API price, so it is not tracked here.
  • A dash means the lab did not report a result for that model.
Source: Introducing Claude Fable 5.1 and Claude Mythos 5.1 (Anthropic)
Claude Opus 5.5 launch results

Reported by Anthropic on 22 Sept 2026 · exact values

BenchmarkClaude Opus 5.5Claude Fable 5.1Claude Opus 5GPT-6 AstraGPT-5.6 Sol
Terminal-Bench 4.0Agentic coding66.4%55.8%52.3%57.9%37.3%
FrontierCode v1.1 (main)Agentic coding54.4%50.3%48%53.3%47.5%
CursorBench 4.0Agentic coding57.8%51.8%46.6%–41.7%
GDPval-AA v2.1Knowledge work1846 Elo1735 Elo1708 Elo1542 Elo1588 Elo
AutomationBenchBusiness workflows40%31.4%26.9%41.4%28.8%
Humanity's Last ExamMultidisciplinary reasoning · with tools67.7%65.6%63.6%57.2%–
Terminal-Bench-Science 0.1Agentic scientific research58.7%52.6%29%64.6%22.4%
OSWorld 2.1Computer use · partial81.8%80.7%74%––
ChartographyVisual chart recognition · with tools89%88.4%83.4%––
  • Bold marks the best reported score in each row.
  • Claude Opus 5.5 results use adaptive thinking at max effort unless noted.
  • Terminal-Bench 4.0 shows each model's highest score: Claude Opus 5.5 at xhigh effort and GPT-6 Astra at high effort. GPT-6 Astra and GPT-5.6 Sol figures there are as reported by OpenAI.
  • A dash means the lab did not report a result for that model.
Source: Introducing Claude Opus 5.5 (Anthropic)
Terminal-Bench-Science 0.1

Accuracy against cost · reported by Anthropic on 1 Sept 2026 · exact values

Terminal-based agentic scientific research tasks, in Anthropic's agentic scientific research category.

  • Claude Fable 5.1
  • Claude Fable 5
Show data table
ModelEffortScoreCost per task
Claude Fable 5.1low26.3%$11.1
Claude Fable 5.1medium35.7%$14.9
Claude Fable 5.1high40%$20.3
Claude Fable 5.1xhigh49.5%$31.8
Claude Fable 5.1max52.6%$37.9
Claude Fable 5low12.3%$17.1
Claude Fable 5medium21.4%$25
Claude Fable 5high25%$34.3
Claude Fable 5xhigh23.4%$36
Claude Fable 5max24.7%$44.1
  • Exact values from the data published with Anthropic's announcement.
  • Standard error is about 3.5 to 4.5 points per model.
Source: Introducing Claude Fable 5.1 and Claude Mythos 5.1 (Anthropic)
Humanity's Last Exam, with tools

Pass rate against cost · reported by Anthropic on 1 Sept 2026 · exact values

Expert-level questions across many disciplines, answered with access to tools.

  • Claude Fable 5.1
  • Claude Fable 5
Show data table
ModelEffortPass rateCost per task
Claude Fable 5.1low60%$0.52
Claude Fable 5.1medium62.96%$0.67
Claude Fable 5.1high64.76%$1.05
Claude Fable 5.1xhigh65.08%$2.28
Claude Fable 5.1max65%$3.20
Claude Fable 5low59.64%$0.61
Claude Fable 5medium61.4%$1.01
Claude Fable 5high63.16%$1.42
Claude Fable 5xhigh63.52%$1.92
Claude Fable 5max63.8%$3.44
  • Exact values from the data published with Anthropic's announcement.
Source: Introducing Claude Fable 5.1 and Claude Mythos 5.1 (Anthropic)
Humanity's Last Exam, no tools

Pass rate against cost · reported by Anthropic on 1 Sept 2026 · exact values

Expert-level questions across many disciplines, answered without tools.

  • Claude Fable 5.1
  • Claude Fable 5
Show data table
ModelEffortPass rateCost per task
Claude Fable 5.1low53.16%$0.30
Claude Fable 5.1medium55.92%$0.46
Claude Fable 5.1high57.96%$0.75
Claude Fable 5.1xhigh60.36%$1.53
Claude Fable 5.1max60.92%$2.23
Claude Fable 5low50.6%$0.17
Claude Fable 5medium55.88%$0.40
Claude Fable 5high56.88%$0.62
Claude Fable 5xhigh57.44%$0.91
Claude Fable 5max57.76%$1.70
  • Exact values from the data published with Anthropic's announcement.
Source: Introducing Claude Fable 5.1 and Claude Mythos 5.1 (Anthropic)
CursorBench 3.2.0

Accuracy against cost · reported by Anthropic on 1 Sept 2026 · exact values

Evaluates coding agents on tasks taken from real Cursor sessions (earlier release than CursorBench 4.0).

  • Claude Fable 5.1
  • Claude Fable 5
Show data table
ModelEffortScoreCost per task
Claude Fable 5.1low66.2%$2.90
Claude Fable 5.1medium68%$3.53
Claude Fable 5.1high69.4%$4.80
Claude Fable 5.1xhigh72.8%$6.96
Claude Fable 5.1max73.4%$9.64
Claude Fable 5low62.1%$4.46
Claude Fable 5medium65.2%$6.80
Claude Fable 5high66.5%$8.77
Claude Fable 5xhigh68.4%$11.7
Claude Fable 5max70.5%$17.3
  • Exact values from the data published with Anthropic's announcement.
Source: Introducing Claude Fable 5.1 and Claude Mythos 5.1 (Anthropic)
AutomationBench 1.0.6

Score against cost · reported by OpenAI on 22 Sept 2026 · exact values

AI agents are tested on end-to-end workflows using 47 tools across sales, marketing, operations, support, finance and HR.

  • GPT-6 Luna
  • GPT-5.6 Luna
  • GPT-6 Sol
  • GPT-5.6 Sol
  • GPT-6 Astra
  • Claude Opus 5
  • Claude Fable 5.1
Show data table
ModelEffortScoreCost per task
GPT-6 Lunalow1.2%$0.006
GPT-6 Lunamedium9.4%$0.016
GPT-6 Lunahigh14.5%$0.021
GPT-6 Lunaxhigh12.6%$0.025
GPT-6 Lunamax20.7%$0.037
GPT-5.6 Lunalow1.8%$0.01
GPT-5.6 Lunamedium4.3%$0.02
GPT-5.6 Lunahigh9.1%$0.05
GPT-5.6 Lunaxhigh12.9%$0.06
GPT-5.6 Lunamax17%$0.07
GPT-6 Sollow21.2%$0.19
GPT-6 Solmedium26.9%$0.21
GPT-6 Solhigh31.2%$0.24
GPT-6 Solxhigh33.2%$0.27
GPT-6 Solmax32%$0.34
GPT-5.6 Sollow11.7%$0.31
GPT-5.6 Solmedium19.6%$0.42
GPT-5.6 Solhigh24.8%$0.47
GPT-5.6 Solxhigh26.3%$0.54
GPT-5.6 Solmax28.8%$0.67
GPT-6 Astralow30.3%$1.08
GPT-6 Astramedium34.1%$1.27
GPT-6 Astrahigh37.1%$1.44
GPT-6 Astraxhigh39%$1.50
GPT-6 Astramax41.4%$1.73
Claude Opus 5low20.4%$1.64
Claude Opus 5medium23.9%$2.22
Claude Opus 5high20.5%$2.27
Claude Opus 5xhigh25.3%$2.71
Claude Opus 5max26.9%$3.05
Claude Fable 5.1max31.4%$2.45
  • Exact values from the data labels on OpenAI's charts. OpenAI took competitor scores from publicly available reports.
  • Claude Fable 5.1 ran with Claude Opus 5 as a fallback. OpenAI notes its cost omits the fallback runs (about 40% of tasks), so its real cost is higher.
Source: Introducing GPT-6 Sol and Luna (OpenAI)
FrontierCode 1.1, main set

Score against cost · reported by OpenAI on 22 Sept 2026 · exact values

AI agents write code graded on correctness and mergeability: test quality, scope discipline, code style and codebase standards.

  • GPT-6 Luna
  • GPT-5.6 Luna
  • GPT-6 Sol
  • GPT-5.6 Sol
  • GPT-6 Astra
  • Claude Opus 5
  • Claude Fable 5.1
Show data table
ModelEffortScoreCost per task
GPT-6 Lunalow25.7%$0.021
GPT-6 Lunamedium35.5%$0.053
GPT-6 Lunahigh37.3%$0.067
GPT-6 Lunaxhigh37.1%$0.073
GPT-6 Lunamax42.4%$0.11
GPT-5.6 Lunalow15.4%$0.06
GPT-5.6 Lunamedium25.7%$0.13
GPT-5.6 Lunahigh35.9%$0.23
GPT-5.6 Lunaxhigh38.9%$0.31
GPT-5.6 Lunamax39.8%$0.37
GPT-6 Sollow37.3%$0.45
GPT-6 Solmedium45.9%$0.80
GPT-6 Solhigh47.7%$1.08
GPT-6 Solxhigh48.4%$1.37
GPT-6 Solmax49.3%$2.14
GPT-5.6 Sollow35.4%$1.89
GPT-5.6 Solmedium39.9%$2.69
GPT-5.6 Solhigh45.1%$3.48
GPT-5.6 Solxhigh46.8%$4.15
GPT-5.6 Solmax47.5%$5.19
GPT-6 Astralow45.3%$1.70
GPT-6 Astramedium48.8%$2.43
GPT-6 Astrahigh50.9%$3.01
GPT-6 Astraxhigh50.6%$3.28
GPT-6 Astramax53.3%$4.59
Claude Opus 5low41.9%$2.68
Claude Opus 5medium53.4%$4.31
Claude Opus 5high48%$7.24
Claude Opus 5xhigh43.6%$9.14
Claude Opus 5max48%$11.4
Claude Fable 5.1low49.8%$2.38
Claude Fable 5.1medium50.9%$3.28
Claude Fable 5.1high50.3%$5.27
Claude Fable 5.1xhigh48.7%$9.27
Claude Fable 5.1max50.3%$12.8
  • Exact values from the data labels on OpenAI's charts. OpenAI took competitor scores from publicly available reports.
Source: Introducing GPT-6 Sol and Luna (OpenAI)
Terminal-Bench 4.0

Accuracy against cost · reported by Anthropic on 22 Sept 2026 · exact values

Measures how well a model completes complex, multi-step professional tasks within a command line interface.

  • Claude Opus 5.5
  • Claude Fable 5.1
  • Claude Opus 5
  • GPT-6 Astra
  • GPT-5.6 Sol
Show data table
ModelEffortScoreCost per task
Claude Opus 5.5low38.5%$1.29
Claude Opus 5.5medium57.6%$2.94
Claude Opus 5.5high64.2%$3.88
Claude Opus 5.5xhigh66.4%$7.35
Claude Opus 5.5max64.8%$11.2
Claude Fable 5.1low40.2%$5.70
Claude Fable 5.1medium43.4%$7.80
Claude Fable 5.1high49.4%$10.5
Claude Fable 5.1xhigh51.3%$15.8
Claude Fable 5.1max55.8%$19.5
Claude Opus 5low28.5%$4.25
Claude Opus 5medium41.2%$7.00
Claude Opus 5high47%$10.6
Claude Opus 5xhigh50.6%$13.5
Claude Opus 5max52.3%$15.8
GPT-6 Astralow49.7%$4.95
GPT-6 Astramedium53.9%$6.15
GPT-6 Astrahigh57.9%$7.21
GPT-6 Astraxhigh57.6%$7.48
GPT-6 Astramax56.7%$10.3
GPT-5.6 Sollow7.9%$1.46
GPT-5.6 Solmedium20.9%$2.69
GPT-5.6 Solhigh26.1%$4.12
GPT-5.6 Solxhigh28.5%$5.39
GPT-5.6 Solmax37.3%$7.89
  • Exact values from the data published with Anthropic's announcement.
  • GPT-6 Astra and GPT-5.6 Sol figures are as reported by OpenAI.
Source: Introducing Claude Opus 5.5 (Anthropic)
FrontierCode v1.1, main set

Accuracy against cost · reported by Anthropic on 22 Sept 2026 · exact values

Measures whether an agent's code changes would be merged.

  • Claude Opus 5.5
  • Claude Fable 5.1
  • Claude Opus 5
  • GPT-6 Astra
  • GPT-5.6 Sol
Show data table
ModelEffortScoreCost per task
Claude Opus 5.5low47.3%$0.40
Claude Opus 5.5medium54.64%$0.80
Claude Opus 5.5high53.99%$1.09
Claude Opus 5.5xhigh51.42%$2.25
Claude Opus 5.5max54.43%$6.19
Claude Fable 5.1low52.8%$2.47
Claude Fable 5.1medium50.91%$3.28
Claude Fable 5.1high50.34%$5.27
Claude Fable 5.1xhigh48.73%$9.27
Claude Fable 5.1max50.28%$12.8
Claude Opus 5low41.95%$2.64
Claude Opus 5medium53.38%$4.61
Claude Opus 5high47.99%$7.62
Claude Opus 5xhigh43.65%$8.99
Claude Opus 5max48.04%$12.3
GPT-6 Astralow45.27%$1.59
GPT-6 Astramedium48.83%$2.28
GPT-6 Astrahigh50.94%$2.85
GPT-6 Astraxhigh50.62%$3.10
GPT-6 Astramax53.26%$4.36
GPT-5.6 Sollow35.44%$1.75
GPT-5.6 Solmedium39.93%$2.50
GPT-5.6 Solhigh45.06%$3.25
GPT-5.6 Solxhigh46.84%$3.88
GPT-5.6 Solmax47.49%$4.85
  • Exact values from the data published with Anthropic's announcement.
Source: Introducing Claude Opus 5.5 (Anthropic)
CursorBench 4.0

Accuracy against cost · reported by Anthropic on 22 Sept 2026 · exact values

Evaluates coding agents on ambiguous, multi-file tasks taken from real Cursor sessions.

  • Claude Opus 5.5
  • Claude Fable 5.1
  • Claude Opus 5
  • GPT-5.6 Sol
Show data table
ModelEffortScoreCost per task
Claude Opus 5.5low43.7%$1.18
Claude Opus 5.5medium52.5%$2.90
Claude Opus 5.5high56%$3.97
Claude Opus 5.5xhigh56%$6.99
Claude Opus 5.5max57.8%$13.4
Claude Fable 5.1low45.1%$5.44
Claude Fable 5.1medium46.8%$7.05
Claude Fable 5.1high49.2%$9.08
Claude Fable 5.1xhigh51.6%$13
Claude Fable 5.1max51.8%$17.3
Claude Opus 5low40.7%$4.87
Claude Opus 5medium43.3%$6.94
Claude Opus 5high44.7%$9.00
Claude Opus 5xhigh46.1%$11.4
Claude Opus 5max46.6%$11.9
GPT-5.6 Sollow24.6%$0.87
GPT-5.6 Solmedium31.1%$1.77
GPT-5.6 Solhigh35.7%$2.85
GPT-5.6 Solxhigh37.7%$4.40
GPT-5.6 Solmax41.7%$8.23
  • Exact values from the data published with Anthropic's announcement.
Source: Introducing Claude Opus 5.5 (Anthropic)
GDPval-AA v2.1

Elo against cost · reported by Anthropic on 22 Sept 2026 · exact values

Artificial Analysis's evaluation of agents on real-world professional work across 44 occupations.

  • Claude Opus 5.5
  • Claude Fable 5.1
  • Claude Opus 5
  • GPT-6 Astra
  • GPT-5.6 Sol
Show data table
ModelEffortEloCost per task
Claude Opus 5.5low1224 Elo$0.21
Claude Opus 5.5medium1576 Elo$0.86
Claude Opus 5.5high1692 Elo$1.54
Claude Opus 5.5xhigh1820 Elo$4.21
Claude Opus 5.5max1846 Elo$8.92
Claude Fable 5.1low1450 Elo$1.41
Claude Fable 5.1medium1536 Elo$2.17
Claude Fable 5.1high1617 Elo$3.43
Claude Fable 5.1xhigh1721 Elo$7.09
Claude Fable 5.1max1735 Elo$9.59
Claude Opus 5low1294 Elo$0.58
Claude Opus 5medium1476 Elo$1.36
Claude Opus 5high1581 Elo$3.03
Claude Opus 5xhigh1676 Elo$4.96
Claude Opus 5max1708 Elo$6.76
GPT-6 Astralow1366 Elo$0.85
GPT-6 Astramedium1468 Elo$1.82
GPT-6 Astrahigh1485 Elo$2.43
GPT-6 Astraxhigh1516 Elo$3.04
GPT-6 Astramax1542 Elo$4.53
GPT-5.6 Sollow1289 Elo$0.27
GPT-5.6 Solmedium1403 Elo$0.60
GPT-5.6 Solhigh1480 Elo$1.11
GPT-5.6 Solxhigh1548 Elo$1.70
GPT-5.6 Solmax1588 Elo$2.81
  • Exact values from the data published with Anthropic's announcement.
Source: Introducing Claude Opus 5.5 (Anthropic)
WANDR

Accuracy against cost · reported by Anthropic on 22 Sept 2026 · exact values

Perplexity's benchmark measuring agents on large data collection tasks.

  • Claude Opus 5.5
  • Claude Fable 5.1
  • Claude Opus 5
Show data table
ModelEffortScoreCost per task
Claude Opus 5.5low31.2%$1.20
Claude Opus 5.5medium62.8%$11.2
Claude Opus 5.5high67.3%$16.9
Claude Opus 5.5xhigh71.3%$29.1
Claude Opus 5.5max72.3%$37.9
Claude Fable 5.1low63.3%$23.3
Claude Fable 5.1medium64.5%$27.4
Claude Fable 5.1high66.7%$33.6
Claude Fable 5.1xhigh67.7%$42.8
Claude Fable 5.1max68.7%$49
Claude Opus 5low50.5%$10.7
Claude Opus 5medium58.1%$24
Claude Opus 5high64.6%$43.5
Claude Opus 5xhigh67%$53.8
Claude Opus 5max67.2%$61.6
  • Exact values from the data published with Anthropic's announcement.
  • Claude models ran with offline web search and fetch tools and a 980K-token task budget. This differs from Perplexity's published setup, so scores are not comparable with Perplexity's leaderboard.
Source: Introducing Claude Opus 5.5 (Anthropic)
Terminal-Bench 4.0

Accuracy against API cost · reported by OpenAI on 9 Sept 2026 · exact values

Tests agents on complex terminal-based tasks, including software engineering, system configuration and data analysis.

  • GPT-6 Astra
  • GPT-5.6 Sol
  • Claude Fable 5.1
  • Claude Fable 5
  • Claude Opus 5
Show data table
ModelEffortAccuracyCost per task
GPT-6 Astralow49.7%$4.95
GPT-6 Astramedium53.9%$6.15
GPT-6 Astrahigh57.9%$7.21
GPT-6 Astraxhigh57.6%$7.48
GPT-6 Astramax56.7%$10.3
GPT-5.6 Sollow7.9%$1.46
GPT-5.6 Solmedium20.9%$2.69
GPT-5.6 Solhigh26.1%$4.12
GPT-5.6 Solxhigh28.5%$5.39
GPT-5.6 Solmax37.3%$7.89
Claude Fable 5.1low40.2%$5.70
Claude Fable 5.1medium43.4%$7.80
Claude Fable 5.1high49.4%$10.5
Claude Fable 5.1xhigh51.3%$15.8
Claude Fable 5.1max55.8%$19.5
Claude Fable 5max44.5%$22.2
Claude Opus 5max52.6%$18.4
  • Exact values from the data labels on OpenAI's charts. OpenAI took competitor scores from publicly available reports.
  • Claude Fable 5 and Claude Opus 5 are shown at max effort only, as published.
Source: GPT-6 Astra: The next generation in intelligence for work (OpenAI)