GPT-6.1 Sol vs Claude Sonnet 5.5: OpenAI Bet on Pricing

Artificial Analysis benchmarks

BenchmarkGPT-6.1 Sol (Max)Claude Sonnet 5.5 (Max)Difference
Intelligence Index v4.3.251.856.0Sonnet +4.2
AA-Briefcase v1.11564.2 Elo1811 EloSonnet +246.8 Elo
GDPval-AA v2.11575.1 Elo1844 EloSonnet +268.9 Elo
AutomationBench-AA64.9%71%Sonnet +6.1 points
Terminal-Bench 4.056.1%64%Sonnet +7.9 points
SciCode54.2%61%Sonnet +6.8 points
Humanity's Last Exam52.9%55%Sonnet +2.1 points
GDP.pdf31.0%26%GPT-6.1 Sol +5.0 points
CritPt31.7%31%GPT-6.1 Sol +0.7 points
AA-Omniscience41.532GPT-6.1 Sol +9.5
AA-LCR v1.183%83%Tie

Cost and efficiency

MetricGPT-6.1 Sol (Max)Claude Sonnet 5.5 (Max)Difference
Intelligence Index51.856.0Sonnet +4.2
Cost per Intelligence Index task$0.72$7.60Sol about 90.5% cheaper
Output tokens per taskAbout 38KAbout 193KSol uses about 80% fewer
Output speedAbout 67 tokens/sAbout 138 tokens/sSonnet about 2.1x faster
Input, per 1M tokens$2$2Same
Output, per 1M tokens$10$10Same
Cached input, per 1M tokens$0.10$0.20Sol 50% cheaper

Pricing and specs

SpecificationGPT-6.1 SolClaude Sonnet 5.5Difference
Input, per 1M tokens$2$2Same
Cached input, per 1M tokens$0.10$0.20Sol 50% cheaper
Cache write, per 1M tokens$2.50$2.50Same
Output, per 1M tokens$10$10Same
Context window1.05M tokens1M tokensSol +50K tokens
Max output128K tokens128K tokensSame
Reasoning levelsLow to MaxLow to Max, adaptiveBoth adjustable
Long-context surchargeYes, above 272K input tokensNone statedDifferent pricing model