CheckNet.NETWORK DIAGNOSTICS / TOOLKIT
WorkspaceAI Model TrackerNETWORK TOOLKIT
Model trackerAll models and earlier releases

OpenAI · launched 29 Sept 2026

GPT-6.1 Sol

An upgrade to GPT-6 Sol for coding, professional work and computer use, with cached input priced 50% lower.

Read the official announcement
Model details
Vendor
OpenAI
Launch date
29 Sept 2026
Input price
$2.00 / 1M tokens
Output price
$10 / 1M tokens
Cached input
$0.10 / 1M tokens
Batch discount
50%
Context window
1.05M tokens
Max output
128K tokens
Input types
Text, Image, File
API model ID
gpt-6.1-sol
OpenRouter ID
openai/gpt-6.1-sol

Listed rates apply up to 272K input tokens. Above 272K, input costs $4, cached input $0.20 and output $15 per million tokens for the full request.

Compared with GPT-6 Sol

Input price

$2.00

was $2.00 · Same price

Output price

$10

was $10 · Same price

Blended price

$4.00

was $4.00 · Same price

Listed rates apply up to 272K input tokens. Above 272K, input costs $4, cached input $0.20 and output $15 per million tokens for the full request.

Best published results
  • DeepSWE v1.1high effort · $0.65 per task · reported by OpenAI75.22%
  • GDP.pdfhigh effort · $0.35 per task · reported by OpenAI32%
  • AutomationBench 1.0.6max effort · $0.30 per task · reported by OpenAI36.1%
  • OSWorld 2.0 (offline set)max effort · $1.27 per task · reported by OpenAI71.42%
  • Terminal-Bench Science 0.1max effort · $5.47 per task · reported by OpenAI57.02%
  • Factual error rate on difficult prompts(lower is better)xhigh effort · $0.10 per task · reported by OpenAI4.12%

Price compared with other models

ModelLaunchedInput / 1MOutput / 1MCached / 1MBlended / 1Mvs GPT-6.1 Sol
Claude Haiku 5.5AnthropicListed prices and comparisons apply to prompts up to 100K tokens. Above 100K, input costs $0.50, output $2.50 and cache reads $0.05 per million tokens.7 Oct 2026$0.10$0.50$0.01$0.2095% cheaper
GPT-6 LunaOpenAI22 Sept 2026$0.10$0.50$0.01$0.2095% cheaper
GPT-5.6 LunaOpenAI9 Jul 2026$0.20$1.20$0.02$0.4589% cheaper
Claude Haiku 4.5Anthropic15 Oct 2025$1.00$5.00$0.10$2.0050% cheaper
Claude Sonnet 5Anthropic30 Jun 2026$2.00$10$0.20$4.00Same price
Claude Sonnet 5.5AnthropicCache reads were reduced from $0.20 to $0.10 per million tokens on 7 October 2026.28 Sept 2026$2.00$10$0.10$4.00Same price
GPT-6 SolOpenAI · previous generationListed rates apply up to 272K input tokens. Above 272K, input costs $4, cached input $0.40 and output $15 per million tokens for the full request.22 Sept 2026$2.00$10$0.20$4.00Same price
GPT-6.1 SolOpenAIListed rates apply up to 272K input tokens. Above 272K, input costs $4, cached input $0.20 and output $15 per million tokens for the full request.29 Sept 2026$2.00$10$0.10$4.00Baseline
Claude Opus 5.5Anthropic22 Sept 2026$4.00$20$0.20$8.002× the price
GPT-5.6 SolOpenAIPrice before the GPT-6 launch, from OpenAI's announcement. OpenRouter now lists $2 / $10.9 Jul 2026$4.00$20–$8.002× the price
Claude Opus 5Anthropic24 Jul 2026$5.00$25$0.50$102.5× the price
Claude Fable 5Anthropic9 Jun 2026$10$50$1.00$205× the price
Claude Fable 5.1Anthropic1 Sept 2026$10$50$0.25$205× the price
GPT-6 AstraOpenAI4 Sept 2026$10$50$1.00$205× the price

GPT-6.1 Sol and GPT-6 Sol rates were checked on 8 October 2026 against OpenAI model documentation. USD per million tokens, checked 2 Oct 2026 via OpenRouter. Comparison uses the blended price (3 input tokens for every output token). Haiku 5.5 and Sonnet 5.5 prices were updated from Anthropic's 7 October 2026 announcement.

Official benchmark results

GPT-6.1 Sol safety evaluations

Reported by OpenAI on 29 Sept 2026 · exact values

BenchmarkGPT-6.1 SolGPT-6 SolGPT-6 AstraGPT-6 Luna
Failure to disclose a broken search toolSafety evaluation · Max effort2.1%4.92%1.5%28.67%
Reviewer bypass attemptsSafety evaluation · Max effort0%0%0%0.26%
Warning circumventionSafety evaluation · Max effort23.48%64.39%17.42%42.37%
Computer-use safety stress testSafety evaluation · Xhigh effort4.32%17.39%2.4%13.7%
  • Bold marks the best reported score in each row.
  • All rates are lower-is-better. These targeted evaluations stress challenging scenarios and do not represent typical usage.
  • Exact values from the embedded announcement chart data. No cost values were published for these evaluations.
Source: Introducing GPT-6.1 Sol (OpenAI)
DeepSWE v1.1

Score against cost · reported by OpenAI on 29 Sept 2026 · exact values

Agentic software engineering performance across reasoning effort settings.

  • GPT-6.1 Sol
  • GPT-6 Sol
  • GPT-6 Astra
Show data table
ModelEffortScore (%)Cost per task
GPT-6.1 Sollow64.38%$0.17
GPT-6.1 Solmedium73.01%$0.42
GPT-6.1 Solhigh75.22%$0.65
GPT-6.1 Solxhigh71.9%$0.79
GPT-6.1 Solmax71.9%$1.57
GPT-6 Sollow37.17%$0.16
GPT-6 Solmedium56.64%$0.38
GPT-6 Solhigh65.27%$0.64
GPT-6 Solxhigh66.59%$1.00
GPT-6 Solmax68.81%$2.74
GPT-6 Astralow67.04%$1.60
GPT-6 Astramedium72.79%$3.08
GPT-6 Astrahigh73.23%$3.92
GPT-6 Astraxhigh74.12%$4.43
GPT-6 Astramax73.23%$7.50
  • Exact score and cost values from the chart data embedded in OpenAI's announcement.
Source: Introducing GPT-6.1 Sol (OpenAI)
GDP.pdf

Score against cost · reported by OpenAI on 29 Sept 2026 · exact values

Accuracy on professional questions using complex PDF documents.

  • GPT-6.1 Sol
  • GPT-6 Sol
  • GPT-6 Astra
  • Claude Opus 5.5 with fallbacks
Show data table
ModelEffortScore (%)Cost per task
GPT-6.1 Sollow27%$0.33
GPT-6.1 Solmedium30%$0.34
GPT-6.1 Solhigh32%$0.35
GPT-6.1 Solxhigh31.8%$0.37
GPT-6.1 Solmax31%$0.42
GPT-6 Sollow21.8%$0.33
GPT-6 Solmedium25.4%$0.34
GPT-6 Solhigh28%$0.35
GPT-6 Solxhigh23.8%$0.37
GPT-6 Solmax24.8%$0.43
GPT-6 Astralow30.4%$1.70
GPT-6 Astramedium30.4%$1.72
GPT-6 Astrahigh31%$1.79
GPT-6 Astraxhigh32.2%$1.91
GPT-6 Astramax31%$2.08
Claude Opus 5.5 with fallbackslow25.6%$0.76
Claude Opus 5.5 with fallbacksmedium25.6%$0.80
Claude Opus 5.5 with fallbackshigh28.8%$0.83
Claude Opus 5.5 with fallbacksxhigh26.6%$0.96
Claude Opus 5.5 with fallbacksmax26.2%$1.55
  • Exact score and cost values from the chart data embedded in OpenAI's announcement.
  • Opus 5.5 results use fallback models as reported by OpenAI; these are not standalone Opus 5.5 results.
Source: Introducing GPT-6.1 Sol (OpenAI)
AutomationBench 1.0.6

Score against cost · reported by OpenAI on 29 Sept 2026 · exact values

End-to-end business workflows using 47 tools across sales, marketing, operations, support, finance and HR.

  • GPT-6.1 Sol
  • GPT-6 Sol
  • GPT-6 Astra
  • Claude Opus 5.5 with fallbacks
  • Claude Fable 5.1 with Opus 5 fallback
Show data table
ModelEffortScore (%)Cost per task
GPT-6.1 Sollow24.7%$0.16
GPT-6.1 Solmedium31.7%$0.19
GPT-6.1 Solhigh33.2%$0.23
GPT-6.1 Solxhigh35.5%$0.25
GPT-6.1 Solmax36.1%$0.30
GPT-6 Sollow21.16%$0.19
GPT-6 Solmedium26.94%$0.21
GPT-6 Solhigh31.2%$0.24
GPT-6 Solxhigh33.18%$0.27
GPT-6 Solmax31.96%$0.34
GPT-6 Astralow30.3%$1.08
GPT-6 Astramedium34.1%$1.27
GPT-6 Astrahigh37.1%$1.44
GPT-6 Astraxhigh39%$1.50
GPT-6 Astramax41.4%$1.73
Claude Opus 5.5 with fallbackslow24.2%$0.51
Claude Opus 5.5 with fallbacksmedium29.53%$0.65
Claude Opus 5.5 with fallbackshigh33.03%$0.71
Claude Opus 5.5 with fallbacksxhigh35.77%$0.89
Claude Opus 5.5 with fallbacksmax42.47%$1.44
Claude Fable 5.1 with Opus 5 fallbackmax31.4%$2.45
  • Exact score and cost values from the chart data embedded in OpenAI's announcement.
  • Opus 5.5 results use fallback models as reported by OpenAI; these are not standalone Opus 5.5 results.
  • Fable 5.1 is reported at max effort only. Its cost omits fallbacks, which occurred on approximately 40% of tasks, so its actual cost is higher.
Source: Introducing GPT-6.1 Sol (OpenAI)
OSWorld 2.0 (offline set)

Score against cost · reported by OpenAI on 29 Sept 2026 · exact values

Partial reward on long-horizon computer-use workflows, using the offline set from release v2026.08.08.

  • GPT-6.1 Sol
  • GPT-6 Sol
  • GPT-6 Astra
Show data table
ModelEffortScore (%)Cost per task
GPT-6.1 Sollow58.96%$0.42
GPT-6.1 Solmedium66.84%$0.77
GPT-6.1 Solhigh69.56%$0.96
GPT-6.1 Solxhigh69.38%$1.05
GPT-6.1 Solmax71.42%$1.27
GPT-6 Sollow43.9%$1.01
GPT-6 Solmedium54%$1.38
GPT-6 Solhigh58.29%$1.71
GPT-6 Solxhigh60.54%$2.30
GPT-6 Solmax64.43%$3.37
GPT-6 Astralow62.17%$2.72
GPT-6 Astramedium69.25%$5.36
GPT-6 Astrahigh70.02%$6.91
GPT-6 Astraxhigh71.27%$7.49
GPT-6 Astramax73.49%$9.44
  • Exact score and cost values from the chart data embedded in OpenAI's announcement.
Source: Introducing GPT-6.1 Sol (OpenAI)
Terminal-Bench Science 0.1

Score against cost · reported by OpenAI on 29 Sept 2026 · exact values

Scientific research tasks completed through a terminal.

  • GPT-6.1 Sol
  • GPT-6 Sol
  • GPT-6 Astra
  • Claude Opus 5.5 with fallbacks
Show data table
ModelEffortScore (%)Cost per task
GPT-6.1 Sollow43.71%$1.79
GPT-6.1 Solmedium47.56%$2.34
GPT-6.1 Solhigh51.14%$2.76
GPT-6.1 Solxhigh53.71%$2.89
GPT-6.1 Solmax57.02%$5.47
GPT-6 Sollow9.17%$3.00
GPT-6 Solmedium14.49%$4.41
GPT-6 Solhigh14.61%$4.63
GPT-6 Solxhigh25.29%$6.77
GPT-6 Solmax27.59%$12.2
GPT-6 Astralow55.43%$11.4
GPT-6 Astramedium57.43%$12.3
GPT-6 Astrahigh62%$15
GPT-6 Astraxhigh60.86%$15.8
GPT-6 Astramax68.1%$23.8
Claude Opus 5.5 with fallbacksmax63.33%$23.2
  • Exact score and cost values from the chart data embedded in OpenAI's announcement.
  • Opus 5.5 results use fallback models as reported by OpenAI; these are not standalone Opus 5.5 results.
Source: Introducing GPT-6.1 Sol (OpenAI)
Factual error rate on difficult prompts

Error rate against cost · reported by OpenAI on 29 Sept 2026 · exact values

Answers containing at least one factual error on deliberately difficult prompts. Lower is better.

  • GPT-6.1 Sol
  • GPT-6 Sol
  • GPT-6 Astra
Show data table
ModelEffortAnswers with any factual error (%)Cost per task
GPT-6.1 Sollow7.72%$0.045
GPT-6.1 Solmedium6.29%$0.056
GPT-6.1 Solhigh4.52%$0.082
GPT-6.1 Solxhigh4.12%$0.10
GPT-6.1 Solmax4.61%$0.13
GPT-6 Sollow11.42%$0.05
GPT-6 Solmedium6.86%$0.069
GPT-6 Solhigh5.14%$0.099
GPT-6 Solxhigh4.52%$0.13
GPT-6 Solmax4.57%$0.18
GPT-6 Astralow6.26%$0.24
GPT-6 Astramedium4.41%$0.31
GPT-6 Astrahigh3.9%$0.48
GPT-6 Astraxhigh3.99%$0.60
GPT-6 Astramax3.91%$0.79
  • Exact score and cost values from the chart data embedded in OpenAI's announcement.
  • Prompts are de-identified conversations where users flagged an earlier model error. They are not representative of typical usage.
Source: Introducing GPT-6.1 Sol (OpenAI)