Smart Stack

Side-by-side pricing, features and compliance for any tools in the directory.

Claude Opus 5 vs Claude Sonnet 5

Claude Opus 5

Anthropic

Pick it if: Claude Opus 5 launched July 24, 2026 with a 1 million token context window as the default tier and a new xhigh reasoning effort mode. It replaces Opus 4.8 as Anthropic's flagship Opus model for long-running agentic coding and research workloads.

Its edge: Leads SWE-bench Verified at 97.0%, the highest published score on the benchmark to date.

The catch: Output speed averages about 54 tokens per second, below the 73.7 t/s median for comparable reasoning models.

Claude Sonnet 5

Anthropic

Pick it if: Claude Sonnet 5 is Anthropic's mid-tier flagship (released June 30, 2026) with a 1M-token context window and 82.1% SWE-bench Verified, the first model to clear 80% on that benchmark. Priced at $3 input / $15 output per 1M tokens (40% below Opus 4.8), it defaults to adaptive thinking and posts 96.2% on GPQA Diamond and 81.2% on OSWorld-Verified computer use.

Its edge: First model to break 80% on SWE-bench Verified at 82.1%, ahead of Gemini 3.1 Pro (80.6%) and GPT-5.4 (~80%).

The catch: No native audio or video input/output; text and image only, requiring a separate ASR/TTS stack for voice apps.

Pricing

Input / 1M tokens$5$3
Output / 1M tokens$25$15
Cached input / 1M$0.5$0.3
Blended cost (3:1)$10$6
Free tierfalsetrue
Pricing modelper-tokenper-token

Verdict & fit

StrengthsLeads SWE-bench Verified at 97.0%, the highest published score on the benchmark to date.; Scores 30.16% on ARC-AGI-3 at high effort, about 20x Opus 4.8 and roughly 4x GPT-5.6 Sol Max.; 1M token context window is both the default and the ceiFirst model to break 80% on SWE-bench Verified at 82.1%, ahead of Gemini 3.1 Pro (80.6%) and GPT-5.4 (~80%).; Computer use jumped to 81.2% on OSWorld-Verified and 80.4% on Terminal-Bench 2.1, up from 78.5% and 67.0% on Sonnet 4.6.; Priced a
LimitationsOutput speed averages about 54 tokens per second, below the 73.7 t/s median for comparable reasoning models.; Accepts text, image, and PDF input only; no native audio or video input.; UK AI Security Institute testing found Opus 5 completed No native audio or video input/output; text and image only, requiring a separate ASR/TTS stack for voice apps.; New tokenizer produces roughly 30% more tokens for the same text than Sonnet 4.6, inflating raw token counts and requiring max_t

Capability envelope

Context window1M1M
Max output tokens128K128K
Input modalitiestext; image; pdftext; image; pdf
Output modalitiestext; tool-callstext; tool-calls
CapabilitiesVision; Tool use; Code executionVision; Tool use; Web browsing
Reasoning modeslow; medium; highadaptive-thinking
Long-context recallhighhigh
Opennessproprietaryproprietary

Compare up to four at a time, or run Smart Match to get a ranked shortlist. Browse the full AI directory.