CheckNet.NETWORK DIAGNOSTICS / TOOLKIT
WorkspaceAI Model TrackerNETWORK TOOLKIT
All models

OpenAI · launched 22 Sept 2026

GPT-6 Sol

The mid-tier GPT-6 model, sitting between GPT-6 Luna and the flagship GPT-6 Astra.

Read the official announcement
Model details
Vendor
OpenAI
Launch date
22 Sept 2026
Input price
$2.00 / 1M tokens
Output price
$10 / 1M tokens
Cached input
$0.20 / 1M tokens
Batch discount
50%
Context window
1.05M tokens
Max output
128K tokens
Input types
Text, Image, File
API model ID
gpt-6-sol
OpenRouter ID
openai/gpt-6-sol
Compared with GPT-5.6 Sol

Input price

$2.00

was $4.00 · 50% cheaper

Output price

$10

was $20 · 50% cheaper

Blended price

$4.00

was $8.00 · 50% cheaper

Best published results
  • AutomationBench 1.0.6xhigh effort · $0.27 per task · reported by OpenAI33.2%
  • Agents' Last Exam V1max effort · $2.93 per task · reported by OpenAI56.4%
  • Factual error rate(lower is better)xhigh effort · $0.13 per task · reported by OpenAI4.5%
  • FrontierCode 1.1max effort · $2.14 per task · reported by OpenAI49.3%
  • DeepSWE 1.1max effort · $2.74 per task · reported by OpenAI68.8%
  • OSWorld 2.0max effort · $3.25 per task · reported by OpenAI64.4%
  • FrontierCode v1.1max effort · $2.07 per task · reported by Anthropic49.3%
  • AA-Briefcase v1.1max effort · $2.67 per task · reported by Anthropic1483 Elo

Price compared with other models

ModelLaunchedInput / 1MOutput / 1MCached / 1MBlended / 1Mvs GPT-6 Sol
GPT-6 LunaOpenAI22 Sept 2026$0.10$0.50$0.01$0.2095% cheaper
GPT-5.6 LunaOpenAI9 Jul 2026$0.20$1.20$0.02$0.4589% cheaper
Claude Sonnet 5Anthropic30 Jun 2026$2.00$10$0.20$4.00Same price
Claude Sonnet 5.5Anthropic28 Sept 2026$2.00$10$0.20$4.00Same price
GPT-6 SolOpenAI22 Sept 2026$2.00$10$0.20$4.00Baseline
Claude Opus 5.5Anthropic22 Sept 2026$4.00$20$0.20$8.002× the price
GPT-5.6 SolOpenAI · previous generationPrice before the GPT-6 launch, from OpenAI's announcement. OpenRouter now lists $2 / $10.9 Jul 2026$4.00$20–$8.002× the price
Claude Opus 5Anthropic24 Jul 2026$5.00$25$0.50$102.5× the price
Claude Fable 5Anthropic9 Jun 2026$10$50$1.00$205× the price
Claude Fable 5.1Anthropic1 Sept 2026$10$50$0.25$205× the price
GPT-6 AstraOpenAI4 Sept 2026$10$50$1.00$205× the price

USD per million tokens, checked 2 Oct 2026 via OpenRouter. Comparison uses the blended price (3 input tokens for every output token).

Official benchmark results

GPT-6 Sol and Luna alignment evaluation

Reported by OpenAI on 22 Sept 2026 · exact values

BenchmarkGPT-6 AstraGPT-6 SolGPT-6 LunaGPT-5.6 SolGPT-5.6 Luna
Coding deception rateAlignment · max effort · lower is better0.5%1.3%2.8%10.4%9.5%
  • Bold marks the best reported score in each row.
  • OpenAI's internal evaluation uses tasks deliberately chosen to provoke dishonesty, so deception is much rarer in typical use.
  • The announcement also covers broken search, reviewer bypass, warning circumvention and unauthorized interaction. See OpenAI's system card for those results.
Source: Introducing GPT-6 Sol and Luna (OpenAI)
Claude Sonnet 5.5 launch results

Reported by Anthropic on 28 Sept 2026 · exact values

BenchmarkClaude Sonnet 5.5Claude Sonnet 5Claude Opus 5.5GPT-6 Sol
Terminal-Bench 4.0Agentic coding70.6%10.3%66.4%–
FrontierCode 1.1 (main)Agentic coding · Sonnet 5.5 at max effort, 52.1% at xhigh46.2%42.4%54.4%49.3%
CursorBench 4.0Agentic coding55.5%34.1%57.8%–
GDPval-AA v2.1Knowledge work1844 Elo1449 Elo1846 Elo1487 Elo
AA-Briefcase v1.1Knowledge work1811 Elo1359 Elo1822 Elo1483 Elo
Humanity's Last ExamMultidisciplinary reasoning · with tools64.5%54.9%67.7%–
OSWorld 2.1Computer use · partial80.1%57%81.8%–
ChartographyVisual chart recognition · no tools61.6%15.6%64.4%53.6%
  • Bold marks the best reported score in each row.
  • Claude Opus 5.5's Terminal-Bench 4.0 score is at xhigh effort, its highest.
  • OpenAI recently fixed a bug affecting GPT-6 Sol's image understanding; its GDPval-AA, AA-Briefcase and Chartography scores may predate the fix.
  • A dash means the lab did not report a result for that model.
Source: Introducing Claude Sonnet 5.5 (Anthropic)
AutomationBench 1.0.6

Score against cost · reported by OpenAI on 22 Sept 2026 · exact values

AI agents are tested on end-to-end workflows using 47 tools across sales, marketing, operations, support, finance and HR.

  • GPT-6 Luna
  • GPT-5.6 Luna
  • GPT-6 Sol
  • GPT-5.6 Sol
  • GPT-6 Astra
  • Claude Opus 5
  • Claude Fable 5.1
Show data table
ModelEffortScoreCost per task
GPT-6 Lunalow1.2%$0.006
GPT-6 Lunamedium9.4%$0.016
GPT-6 Lunahigh14.5%$0.021
GPT-6 Lunaxhigh12.6%$0.025
GPT-6 Lunamax20.7%$0.037
GPT-5.6 Lunalow1.8%$0.01
GPT-5.6 Lunamedium4.3%$0.02
GPT-5.6 Lunahigh9.1%$0.05
GPT-5.6 Lunaxhigh12.9%$0.06
GPT-5.6 Lunamax17%$0.07
GPT-6 Sollow21.2%$0.19
GPT-6 Solmedium26.9%$0.21
GPT-6 Solhigh31.2%$0.24
GPT-6 Solxhigh33.2%$0.27
GPT-6 Solmax32%$0.34
GPT-5.6 Sollow11.7%$0.31
GPT-5.6 Solmedium19.6%$0.42
GPT-5.6 Solhigh24.8%$0.47
GPT-5.6 Solxhigh26.3%$0.54
GPT-5.6 Solmax28.8%$0.67
GPT-6 Astralow30.3%$1.08
GPT-6 Astramedium34.1%$1.27
GPT-6 Astrahigh37.1%$1.44
GPT-6 Astraxhigh39%$1.50
GPT-6 Astramax41.4%$1.73
Claude Opus 5low20.4%$1.64
Claude Opus 5medium23.9%$2.22
Claude Opus 5high20.5%$2.27
Claude Opus 5xhigh25.3%$2.71
Claude Opus 5max26.9%$3.05
Claude Fable 5.1max31.4%$2.45
  • Exact values from the data labels on OpenAI's charts. OpenAI took competitor scores from publicly available reports.
  • Claude Fable 5.1 ran with Claude Opus 5 as a fallback. OpenAI notes its cost omits the fallback runs (about 40% of tasks), so its real cost is higher.
Source: Introducing GPT-6 Sol and Luna (OpenAI)
Agents' Last Exam V1

Score against cost · reported by OpenAI on 22 Sept 2026 · exact values

AI agents are evaluated on long-horizon, economically valuable tasks spanning 55 sub-industries of professional computer work.

  • GPT-6 Luna
  • GPT-5.6 Luna
  • GPT-6 Sol
  • GPT-5.6 Sol
  • GPT-6 Astra
  • Claude Opus 5
  • Claude Fable 5
Show data table
ModelEffortScoreCost per task
GPT-6 Lunalow36.3%$0.025
GPT-6 Lunamedium46.8%$0.11
GPT-6 Lunahigh43.6%$0.11
GPT-6 Lunaxhigh47.9%$0.11
GPT-6 Lunamax50.9%$0.15
GPT-5.6 Lunalow31%$0.18
GPT-5.6 Lunamedium38.5%$0.38
GPT-5.6 Lunahigh46.1%$0.84
GPT-5.6 Lunaxhigh49.4%$1.55
GPT-5.6 Lunamax50.4%$2.57
GPT-6 Sollow48.7%$0.86
GPT-6 Solmedium53.1%$1.27
GPT-6 Solhigh52.6%$1.53
GPT-6 Solxhigh55.4%$1.67
GPT-6 Solmax56.4%$2.93
GPT-5.6 Sollow45.1%$1.63
GPT-5.6 Solmedium52.1%$3.41
GPT-5.6 Solhigh52.4%$3.79
GPT-5.6 Solxhigh53.6%$5.08
GPT-5.6 Solmax52.8%$7.13
GPT-6 Astralow53.4%$3.07
GPT-6 Astramedium57.6%$4.10
GPT-6 Astrahigh57.8%$4.64
GPT-6 Astraxhigh58.3%$5.40
GPT-6 Astramax59.3%$6.23
Claude Opus 5low51.9%$4.06
Claude Opus 5medium53%$5.29
Claude Opus 5high55.9%$7.29
Claude Opus 5xhigh55.5%$10
Claude Opus 5max52.7%$9.76
Claude Fable 5xhigh48.7%$28.6
Claude Fable 5adaptive41.3%$15.2
  • Exact values from the data labels on OpenAI's charts. OpenAI took competitor scores from publicly available reports.
  • Claude Fable 5 ran with Claude Opus 4.8 as a fallback.
Source: Introducing GPT-6 Sol and Luna (OpenAI)
Factual error rate on difficult prompts

Answers with any factual error against cost · lower is better · reported by OpenAI on 22 Sept 2026 · exact values

OpenAI's internal evaluation on de-identified ChatGPT conversations where users had flagged a factual error from a prior model.

  • GPT-6 Luna
  • GPT-5.6 Luna
  • GPT-6 Sol
  • GPT-5.6 Sol
  • GPT-6 Astra
Show data table
ModelEffortError rateCost per task
GPT-6 Lunalow27.7%$0.0024
GPT-6 Lunamedium17.5%$0.0033
GPT-6 Lunahigh12.5%$0.0045
GPT-6 Lunaxhigh10.2%$0.0062
GPT-6 Lunamax7.6%$0.012
GPT-5.6 Lunalow36.8%$0.0045
GPT-5.6 Lunamedium24.7%$0.0062
GPT-5.6 Lunahigh17.3%$0.011
GPT-5.6 Lunaxhigh12.2%$0.015
GPT-5.6 Lunamax12%$0.022
GPT-6 Sollow11.4%$0.05
GPT-6 Solmedium6.9%$0.069
GPT-6 Solhigh5.1%$0.099
GPT-6 Solxhigh4.5%$0.13
GPT-6 Solmax4.6%$0.18
GPT-5.6 Sollow19.4%$0.10
GPT-5.6 Solmedium15%$0.15
GPT-5.6 Solhigh10.8%$0.23
GPT-5.6 Solxhigh8.4%$0.39
GPT-5.6 Solmax8.5%$0.87
GPT-6 Astralow6.3%$0.24
GPT-6 Astramedium4.4%$0.31
GPT-6 Astrahigh3.9%$0.48
GPT-6 Astraxhigh4%$0.60
GPT-6 Astramax3.9%$0.79
  • Exact values from the data labels on OpenAI's charts. OpenAI took competitor scores from publicly available reports.
  • These prompts were chosen because they caused errors before, so error rates are far higher than in typical use.
Source: Introducing GPT-6 Sol and Luna (OpenAI)
FrontierCode 1.1, main set

Score against cost · reported by OpenAI on 22 Sept 2026 · exact values

AI agents write code graded on correctness and mergeability: test quality, scope discipline, code style and codebase standards.

  • GPT-6 Luna
  • GPT-5.6 Luna
  • GPT-6 Sol
  • GPT-5.6 Sol
  • GPT-6 Astra
  • Claude Opus 5
  • Claude Fable 5.1
Show data table
ModelEffortScoreCost per task
GPT-6 Lunalow25.7%$0.021
GPT-6 Lunamedium35.5%$0.053
GPT-6 Lunahigh37.3%$0.067
GPT-6 Lunaxhigh37.1%$0.073
GPT-6 Lunamax42.4%$0.11
GPT-5.6 Lunalow15.4%$0.06
GPT-5.6 Lunamedium25.7%$0.13
GPT-5.6 Lunahigh35.9%$0.23
GPT-5.6 Lunaxhigh38.9%$0.31
GPT-5.6 Lunamax39.8%$0.37
GPT-6 Sollow37.3%$0.45
GPT-6 Solmedium45.9%$0.80
GPT-6 Solhigh47.7%$1.08
GPT-6 Solxhigh48.4%$1.37
GPT-6 Solmax49.3%$2.14
GPT-5.6 Sollow35.4%$1.89
GPT-5.6 Solmedium39.9%$2.69
GPT-5.6 Solhigh45.1%$3.48
GPT-5.6 Solxhigh46.8%$4.15
GPT-5.6 Solmax47.5%$5.19
GPT-6 Astralow45.3%$1.70
GPT-6 Astramedium48.8%$2.43
GPT-6 Astrahigh50.9%$3.01
GPT-6 Astraxhigh50.6%$3.28
GPT-6 Astramax53.3%$4.59
Claude Opus 5low41.9%$2.68
Claude Opus 5medium53.4%$4.31
Claude Opus 5high48%$7.24
Claude Opus 5xhigh43.6%$9.14
Claude Opus 5max48%$11.4
Claude Fable 5.1low49.8%$2.38
Claude Fable 5.1medium50.9%$3.28
Claude Fable 5.1high50.3%$5.27
Claude Fable 5.1xhigh48.7%$9.27
Claude Fable 5.1max50.3%$12.8
  • Exact values from the data labels on OpenAI's charts. OpenAI took competitor scores from publicly available reports.
Source: Introducing GPT-6 Sol and Luna (OpenAI)
DeepSWE 1.1

Score against cost · reported by OpenAI on 22 Sept 2026 · exact values

AI agents solve original, long-horizon software engineering tasks in real codebases.

  • GPT-6 Luna
  • GPT-5.6 Luna
  • GPT-6 Sol
  • GPT-5.6 Sol
  • GPT-6 Astra
  • Claude Opus 5
  • Claude Fable 5
Show data table
ModelEffortScoreCost per task
GPT-6 Lunalow2.4%$0.0057
GPT-6 Lunamedium44.5%$0.052
GPT-6 Lunahigh59.3%$0.084
GPT-6 Lunaxhigh61.3%$0.11
GPT-6 Lunamax66.6%$0.22
GPT-5.6 Lunalow1.2%$0.011
GPT-5.6 Lunamedium9.3%$0.031
GPT-5.6 Lunahigh42.4%$0.13
GPT-5.6 Lunaxhigh56.2%$0.27
GPT-5.6 Lunamax62.2%$0.53
GPT-6 Sollow37.2%$0.16
GPT-6 Solmedium56.6%$0.38
GPT-6 Solhigh65.3%$0.64
GPT-6 Solxhigh66.6%$1.00
GPT-6 Solmax68.8%$2.74
GPT-5.6 Sollow45.4%$0.82
GPT-5.6 Solmedium61.1%$1.42
GPT-5.6 Solhigh69.4%$2.66
GPT-5.6 Solxhigh70.7%$3.60
GPT-5.6 Solmax72.7%$6.46
GPT-6 Astralow67%$1.60
GPT-6 Astramedium72.8%$3.08
GPT-6 Astrahigh73.2%$3.92
GPT-6 Astraxhigh74.1%$4.43
GPT-6 Astramax73.2%$7.50
Claude Opus 5low58.1%$1.66
Claude Opus 5medium68.9%$3.29
Claude Opus 5high72.8%$6.08
Claude Opus 5xhigh73.2%$9.07
Claude Opus 5max73.7%$11.8
Claude Fable 5low59.6%$3.76
Claude Fable 5medium65.4%$6.09
Claude Fable 5high68.6%$9.18
Claude Fable 5xhigh69.9%$13.4
Claude Fable 5max69.7%$21.6
  • Exact values from the data labels on OpenAI's charts. OpenAI took competitor scores from publicly available reports.
  • OpenAI used Claude Fable 5 scores because Fable 5.1 scores were unavailable.
Source: Introducing GPT-6 Sol and Luna (OpenAI)
OSWorld 2.0, offline set

Partial reward against cost · release v2026.08.08 · reported by OpenAI on 22 Sept 2026 · exact values

AI agents attempt long-horizon computer-use workflows covering everyday and professional tasks.

  • GPT-6 Luna
  • GPT-5.6 Luna
  • GPT-6 Sol
  • GPT-5.6 Sol
  • GPT-6 Astra
  • Claude Opus 5
Show data table
ModelEffortScoreCost per task
GPT-6 Lunalow8.3%$0.03
GPT-6 Lunamedium31.5%$0.062
GPT-6 Lunahigh41.4%$0.12
GPT-6 Lunaxhigh46.7%$0.16
GPT-6 Lunamax52.7%$0.27
GPT-5.6 Lunalow12%$0.019
GPT-5.6 Lunamedium22.7%$0.052
GPT-5.6 Lunahigh35.9%$0.16
GPT-5.6 Lunaxhigh48%$0.36
GPT-5.6 Lunamax52.7%$0.49
GPT-6 Sollow43.9%$0.97
GPT-6 Solmedium54%$1.32
GPT-6 Solhigh58.3%$1.64
GPT-6 Solxhigh60.5%$2.21
GPT-6 Solmax64.4%$3.25
GPT-5.6 Sollow29.8%$0.91
GPT-5.6 Solmedium49.7%$2.73
GPT-5.6 Solhigh56.5%$4.46
GPT-5.6 Solxhigh60.9%$5.93
GPT-5.6 Solmax66.2%$7.71
GPT-6 Astralow62.2%$2.55
GPT-6 Astramedium69.3%$5.10
GPT-6 Astrahigh70%$6.60
GPT-6 Astraxhigh71.3%$7.17
GPT-6 Astramax73.5%$9.07
Claude Opus 5low55.2%$9.88
Claude Opus 5medium60.3%$12.7
Claude Opus 5high65.9%$15.9
Claude Opus 5xhigh70.1%$23.9
Claude Opus 5max70.2%$24.1
  • Exact values from the data labels on OpenAI's charts. OpenAI took competitor scores from publicly available reports.
  • Claude Opus 5 values match the official OSWorld 2.0 leaderboard (osworld-v2.xlang.ai).
Source: Introducing GPT-6 Sol and Luna (OpenAI)
FrontierCode v1.1, main set

Accuracy against cost · reported by Anthropic on 28 Sept 2026 · exact values

Measures whether an agent's code changes would be merged.

  • Claude Sonnet 5.5
  • Claude Opus 5.5
  • Claude Sonnet 5
  • GPT-6 Sol
Show data table
ModelEffortScoreCost per task
Claude Sonnet 5.5low29.3%$0.19
Claude Sonnet 5.5medium36.5%$0.24
Claude Sonnet 5.5high49.4%$0.42
Claude Sonnet 5.5xhigh52.1%$1.59
Claude Sonnet 5.5max46.2%$20.8
Claude Opus 5.5low47.3%$0.40
Claude Opus 5.5medium54.6%$0.80
Claude Opus 5.5high54%$1.09
Claude Opus 5.5xhigh51.4%$2.25
Claude Opus 5.5max54.4%$6.19
Claude Sonnet 5low28.7%$2.39
Claude Sonnet 5medium35.2%$3.81
Claude Sonnet 5high39.4%$6.10
Claude Sonnet 5xhigh42.7%$10.1
Claude Sonnet 5max42.4%$17.1
GPT-6 Sollow37.3%$0.43
GPT-6 Solmedium45.9%$0.77
GPT-6 Solhigh47.7%$1.04
GPT-6 Solxhigh48.4%$1.32
GPT-6 Solmax49.3%$2.07
  • Exact values from the data published with Anthropic's announcement.
Source: Introducing Claude Sonnet 5.5 (Anthropic)
AA-Briefcase v1.1

Elo against cost · reported by Anthropic on 28 Sept 2026 · exact values

Artificial Analysis's benchmark of long-horizon knowledge work.

  • Claude Sonnet 5.5
  • Claude Opus 5.5
  • Claude Sonnet 5
  • GPT-6 Sol
Show data table
ModelEffortEloCost per task
Claude Sonnet 5.5low1264 Elo$0.87
Claude Sonnet 5.5medium1461 Elo$1.64
Claude Sonnet 5.5high1634 Elo$3.95
Claude Sonnet 5.5xhigh1746 Elo$9.63
Claude Sonnet 5.5max1811 Elo$29.2
Claude Opus 5.5low1285 Elo$1.15
Claude Opus 5.5medium1642 Elo$4.40
Claude Opus 5.5high1705 Elo$6.27
Claude Opus 5.5xhigh1780 Elo$12.3
Claude Opus 5.5max1822 Elo$21
Claude Sonnet 5low923 Elo$0.82
Claude Sonnet 5medium1056 Elo$1.73
Claude Sonnet 5high1177 Elo$3.76
Claude Sonnet 5xhigh1274 Elo$7.56
Claude Sonnet 5max1359 Elo$14.4
GPT-6 Sollow905 Elo$0.12
GPT-6 Solmedium1142 Elo$0.34
GPT-6 Solhigh1289 Elo$0.63
GPT-6 Solxhigh1364 Elo$1.19
GPT-6 Solmax1483 Elo$2.67
  • Exact values from the data published with Anthropic's announcement.
  • Artificial Analysis ran Sonnet 5.5 on a pre-release deployment with a since-fixed structured output bug, which may slightly understate its score.
Source: Introducing Claude Sonnet 5.5 (Anthropic)