Multi-dimensional Model & Plan Comparison
Compare pricing, coding pass rates, context windows, and estimated monthly invoices side-by-side. Generate shareable posters with 1 click.
Popular Comparisons:
Pick your workflow, then your budget
Estimates use our price snapshot, excluding cache, retries, taxes and tools. Check availability for preview and legacy models.
Workload cost · USD (lower costs less)
Claude Sonnet 5.5$40.00
Claude 3.7 Sonnet (Hybrid Reasoning)$60.00
Context capacity · tokens (not recall accuracy)
Claude Sonnet 5.51,000,000
Claude 3.7 Sonnet (Hybrid Reasoning)200,000
CursorBench 4.0 · %
Claude Sonnet 5.5No verified data
Claude 3.7 Sonnet (Hybrid Reasoning)No verified data
Tool reliability, long-context recall and same-harness SWE-bench: untested, not zero.
CursorBench · 2026-09-22 ↗Patch review example: why does an empty array slip through?
Editorial example, not measured model output. Review boundary contracts and regression tests; this is not a model ranking.
export function mean(values: number[]) {
return values.reduce((a, b) => a + b, 0)
/ values.length;
}
// mean([]) => NaN
// Missing an explicit empty-input contractAnthropicverified
Claude Sonnet 5.5
★ Save 33% ★ Massive Context 1M
Input$2.00 / 1M
Output$10.00 / 1M
Context1M
CursorBench Pass Rate—
Anthropicverified
Claude 3.7 Sonnet (Hybrid Reasoning)
Input$3.00 / 1M
Output$15.00 / 1M
Context200K
CursorBench Pass Rate—
⚙️ Select ModelClaude Sonnet 5.5 vs Claude 3.7 Sonnet (Hybrid Reasoning)
SLOT 1Anthropic
SLOT 2Anthropic
SLOT 3
💰 Monthly Cost Simulation
Simulate monthly cost differences based on your estimated token usage volume
Claude Sonnet 5.5
$40.00/ mo
(in: $2 + out: $10)
Claude 3.7 Sonnet (Hybrid Reasoning)
$60.00/ mo
(in: $3 + out: $15)
| Attribute | Claude Sonnet 5.5Anthropic | Claude 3.7 Sonnet (Hybrid Reasoning)Anthropic |
|---|---|---|
| Input | $2.00 | $3.00 |
| Output | $10.00 | $15.00 |
| Context | 1M | 200K |
| CursorBench Pass Rate | — | — |
| Status | verified | verified |
| Tags & Capabilities | new, flagship, coding, agentic, reasoning, adaptive-thinking, vision, sota | Flagship, Hybrid Reasoning, SWE-bench 70.3%, 200k Context, SOTA Coding |
| LMArena Code | #3 | #— |
| LMArena Agent | #3 | #— |
| CursorBench | #— | #— |
| Artificial Analysis | #2 | #54 |
| Vals AI | #2 | #— |
| LiveBench | #17 | #— |
| LMArena Text | #27 | #62 |