Side-by-side pricing, features and compliance for any tools in the directory.
Anthropic
Pick it if: Claude Opus 5 launched July 24, 2026 with a 1 million token context window as the default tier and a new xhigh reasoning effort mode. It replaces Opus 4.8 as Anthropic's flagship Opus model for long-running agentic coding and research workloads.
Its edge: Leads SWE-bench Verified at 97.0%, the highest published score on the benchmark to date.
The catch: Output speed averages about 54 tokens per second, below the 73.7 t/s median for comparable reasoning models.
Anthropic
Pick it if: Claude Sonnet 5 is Anthropic's mid-tier flagship (released June 30, 2026) with a 1M-token context window and 82.1% SWE-bench Verified, the first model to clear 80% on that benchmark. Priced at $3 input / $15 output per 1M tokens (40% below Opus 4.8), it defaults to adaptive thinking and posts 96.2% on GPQA Diamond and 81.2% on OSWorld-Verified computer use.
Its edge: First model to break 80% on SWE-bench Verified at 82.1%, ahead of Gemini 3.1 Pro (80.6%) and GPT-5.4 (~80%).
The catch: No native audio or video input/output; text and image only, requiring a separate ASR/TTS stack for voice apps.
| Input / 1M tokens | $5 | $3 |
|---|---|---|
| Output / 1M tokens | $25 | $15 |
| Cached input / 1M | $0.5 | $0.3 |
| Blended cost (3:1) | $10 | $6 |
| Free tier | false | true |
| Pricing model | per-token | per-token |
| Strengths | Leads SWE-bench Verified at 97.0%, the highest published score on the benchmark to date.; Scores 30.16% on ARC-AGI-3 at high effort, about 20x Opus 4.8 and roughly 4x GPT-5.6 Sol Max.; 1M token context window is both the default and the cei | First model to break 80% on SWE-bench Verified at 82.1%, ahead of Gemini 3.1 Pro (80.6%) and GPT-5.4 (~80%).; Computer use jumped to 81.2% on OSWorld-Verified and 80.4% on Terminal-Bench 2.1, up from 78.5% and 67.0% on Sonnet 4.6.; Priced a |
|---|---|---|
| Limitations | Output speed averages about 54 tokens per second, below the 73.7 t/s median for comparable reasoning models.; Accepts text, image, and PDF input only; no native audio or video input.; UK AI Security Institute testing found Opus 5 completed | No native audio or video input/output; text and image only, requiring a separate ASR/TTS stack for voice apps.; New tokenizer produces roughly 30% more tokens for the same text than Sonnet 4.6, inflating raw token counts and requiring max_t |
| Context window | 1M | 1M |
|---|---|---|
| Max output tokens | 128K | 128K |
| Input modalities | text; image; pdf | text; image; pdf |
| Output modalities | text; tool-calls | text; tool-calls |
| Capabilities | Vision; Tool use; Code execution | Vision; Tool use; Web browsing |
| Reasoning modes | low; medium; high | adaptive-thinking |
| Long-context recall | high | high |
| Openness | proprietary | proprietary |
Compare up to four at a time, or run Smart Match to get a ranked shortlist. Browse the full AI directory.