Side-by-side pricing, features and compliance for any tools in the directory.
Side-by-side comparison of Claude Sonnet 5, GPT-5.6 Luna: pricing, capabilities, integrations and compliance — from verified HokAI records.
Anthropic
Pick it if: Sonnet 5 fits engineering teams running high-volume agentic coding and computer-use work who want near-Opus quality at Sonnet pricing. It's the wrong pick for voice or video-first products, since it has no native audio or video I/O, and for configs still pinned to non-default sampling parameters. Max output runs to 128,000 tokens on the synchronous API.
Its edge: First model to break 80% on SWE-bench Verified at 82.1%, ahead of Gemini 3.1 Pro (80.6%) and GPT-5.4 (~80%).
The catch: No native audio or video input/output; text and image only, requiring a separate ASR/TTS stack for voice apps.
OpenAI
Pick it if: Luna replaces expensive per-token inference for high-volume, latency-sensitive workloads like classification and routing, not hard reasoning tasks. It's the only model in the family on ChatGPT's free tier, so most casual users meet it first, while teams needing multi-agent orchestration should upgrade to a pricier sibling.
Its edge: Best cost-efficiency in GPT-5.6 family — $1/$6 per 1M tokens is 20% of Sol, 40% of Terra.
The catch: Reasoning capped at high effort — no xhigh, max, or pro mode; ceiling on complex multi-step reasoning.
| Input / 1M tokens | $3 | $1 |
|---|---|---|
| Output / 1M tokens | $15 | $6 |
| Cached input / 1M | $0.3 | $0.125 |
| Blended cost (3:1) | $6 | $2.25 |
| Summarise a 20-page PDF | $0.105 | $0.036 |
| Support reply | $0.010 | $0.0038 |
| One coding agent run | $0.900 | $0.320 |
| Free tier | true | true |
| Strengths | First model to break 80% on SWE-bench Verified at 82.1%, ahead of Gemini 3.1 Pro (80.6%) and GPT-5.4 (~80%).; Computer use jumped to 81.2% on OSWorld-Verified and 80.4% on Terminal-Bench 2.1, up from 78.5% and 67.0% on Sonnet 4.6.; Priced a | Best cost-efficiency in GPT-5.6 family — $1/$6 per 1M tokens is 20% of Sol, 40% of Terra.; Fastest latency in family (~800ms p50, ~90 tokens/sec) — ideal for latency-sensitive production paths.; Full vision + text capability with function c |
|---|---|---|
| Limitations | No native audio or video input/output; text and image only, requiring a separate ASR/TTS stack for voice apps.; New tokenizer produces roughly 30% more tokens for the same text than Sonnet 4.6, inflating raw token counts and requiring max_t | Reasoning capped at high effort — no xhigh, max, or pro mode; ceiling on complex multi-step reasoning.; No ultra multi-agent, no programmatic tool calling, no explicit cache breakpoints (auto caching only).; SWE-bench ~60%, GPQA ~55% — sign |
| Context window | 1M | 200K |
|---|---|---|
| Max output tokens | 128K | 64K |
| Input modalities | text; image; pdf | text; image |
| Output modalities | text; tool-calls | text; tool-calls; code |
| Capabilities | Vision; Tool use; Web browsing | Vision; Tool use; Code execution |
| Reasoning modes | adaptive-thinking | none; low; medium |
| Long-context recall | high | medium |
| Openness | proprietary | proprietary |
Compare up to four at a time, or run Smart Match to get a ranked shortlist. Browse the full AI directory.