TokenRate

Claude Opus 5 vs GPT-5.6 Sol

The two flagship frontier models of July 2026, head to head. Same input price, different output price — compare SWE-bench, ARC-AGI-3, agentic coding, and context windows.

Comparing 2 models · Prices last verified

Verdict

Both charge $5 per million input tokens, but Opus 5 is 17% cheaper on output ($25 vs $30) — and output dominates most bills. Opus 5 leads on SWE-bench Pro (79.2% vs 64.6%), ARC-AGI-3 (30.2% vs 7.8%), Frontier-Bench (43.3% vs 34.4%), OSWorld 2.0 and GDPval-AA v2. GPT-5.6 Sol edges ahead on Terminal-Bench 2.1 and BrowseComp, and offers a slightly larger 1.05M context window plus much faster peak throughput. Opus 5 is the stronger default; Sol wins on raw speed and terminal work.

Pricing Comparison

ModelProviderInput / 1MOutput / 1MContextTier
Claude Opus 5Best valueAnthropic$5.00$25.001Mflagship
GPT-5.6 SolOpenAI$5.00$30.001Mflagship

Model Breakdown

Claude Opus 5

Anthropic

$5.00

per 1M input

Claude Opus 5 is Anthropic's flagship model, released July 24, 2026. It reaches close to Claude Fable 5's frontier intelligence at half the price, and holds the same $5/$25 per-million pricing as Opus 4.8 while more than doubling its agentic coding score. It tops Frontier-Bench v0.1 (43.3%), ARC-AGI-3 (30.2%, roughly 4x the next best model), and Artificial Analysis' GDPval-AA v2 knowledge-work leaderboard (1,861 Elo). It is also Anthropic's most aligned Opus model to date.

Near-Fable-5 frontier quality at half the price ($5/$25 vs $10/$50)
State of the art on agentic coding, computer use, and knowledge work

$5.00

per 1M input

GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series.

Massive 1M-token context window
Frontier-class quality on complex tasks

Our Verdict

Both charge $5 per million input tokens, but Opus 5 is 17% cheaper on output ($25 vs $30) — and output dominates most bills. Opus 5 leads on SWE-bench Pro (79.2% vs 64.6%), ARC-AGI-3 (30.2% vs 7.8%), Frontier-Bench (43.3% vs 34.4%), OSWorld 2.0 and GDPval-AA v2. GPT-5.6 Sol edges ahead on Terminal-Bench 2.1 and BrowseComp, and offers a slightly larger 1.05M context window plus much faster peak throughput. Opus 5 is the stronger default; Sol wins on raw speed and terminal work.

FAQ

Compare These Models Yourself

Use the TokenRate calculator to enter your budget or token count and see the exact cost for each model side by side.

Open Calculator →

More Comparisons