Multi-dimensional Model & Plan Comparison
Compare pricing, coding pass rates, context windows, and estimated monthly invoices side-by-side. Generate shareable posters with 1 click.
Popular Comparisons:
AnthropicVerified
Claude Opus 5.5
★ Save 20% ★ Score Leader (57.8%)
Input$4.00 / 1M
Output$20.00 / 1M
Context1M
CursorBench Pass Rate57.8%
Pick your workflow, then your budget
Estimates use our price snapshot, excluding cache, retries, taxes and tools. Check availability for preview and legacy models.
Workload cost · USD (lower costs less)
Claude Opus 5.5$80.00
GPT-5.6 Sol$110.00
Context capacity · tokens (not recall accuracy)
Claude Opus 5.51,000,000
GPT-5.6 Sol1,000,000
CursorBench 4.0 · %
Claude Opus 5.557.8%
GPT-5.6 Sol41.7%
Tool reliability, long-context recall and same-harness SWE-bench: untested, not zero.
CursorBench · 2026-09-22 ↗Patch review example: why does an empty array slip through?
Editorial example, not measured model output. Review boundary contracts and regression tests; this is not a model ranking.
export function mean(values: number[]) {
return values.reduce((a, b) => a + b, 0)
/ values.length;
}
// mean([]) => NaN
// Missing an explicit empty-input contract⚙️ Select ModelClaude Opus 5.5 vs GPT-5.6 Sol
Slot 1Anthropic
Slot 2OpenAI
Slot 3
💰 Monthly Cost Simulation
Simulate monthly cost differences based on your estimated token usage volume
Claude Opus 5.5
$80.00/ mo
(in: $4 + out: $20)
GPT-5.6 Sol
$110.00/ mo
(in: $5 + out: $30)
| Attribute | Claude Opus 5.5Anthropic | GPT-5.6 SolOpenAI |
|---|---|---|
| Input | $4.00 | $5.00 |
| Output | $20.00 | $30.00 |
| Context | 1M | 1M |
| CursorBench Pass Rate | 57.8% | 41.7% |
| Status | Verified | Verified |
| Tags & Capabilities | New, Flagship, Reasoning, Coding, Agentic, Adaptive thinking, SOTA | Flagship, Reasoning, Coding, Multimodal, Thinking |
| LMArena Code | #1 | #17 |
| LMArena Agent | #2 | #11 |
| CursorBench | #1 | #5 |
| Artificial Analysis | #1 | #11 |
| Vals AI | #3 | #10 |
| LiveBench | #2 | #8 |
| LMArena Text | #4 | #13 |