
| Model | Context | Best for |
|---|---|---|
| GPT-5.5 | 1.05M | Frontier reasoning for complex professional workloads. Highest reliability on hard problems; most expensive GPT option. |
| GPT-5.6 Terra | 1.05M | Best everyday GPT. Balances intelligence and cost — half the price of the flagship tier with strong general capability. |
| GPT-5.6 Luna | 1.05M | Fast and cheap for high-volume simple work. ⚠️ Noticeably weaker at long-context recall — don't feed it huge documents. |
| GPT-5.4 Mini | 400K | Budget tier. Text + image input, tool use, web/file search. Good for bulk, low-complexity tasks. |
| Model | Context | Best for |
|---|---|---|
| Claude Opus 5 | 1M | Highest capability. Deep reasoning, long analysis, hardest coding. No long-context price penalty — best choice for very large documents. |
| Claude Sonnet 5 | 1M | Daily driver. Most agentic Sonnet yet — sustained coding, multi-step automation, debugging, research. Near-Opus quality at much lower cost. |
| Claude Haiku 4.5 | 200K | Fastest and lightest. Quick lookups, short drafts, summaries, high-volume simple tasks. |
| Claude Opus 4.8 | 1M | Previous premium tier. Still excellent for accuracy-critical work; superseded by Opus 5 at the same price. |
| Claude Sonnet 4.6 (1M) | 1M | Previous general-purpose default. Superseded by Sonnet 5 — use Sonnet 5 unless you need output consistency with older work. |
| Model | Context | Best for |
|---|---|---|
| Gemini 3.1 Pro | 1M | Google flagship. Advanced reasoning, agentic workflows, and coding; strongest multimodal handling (text, image, audio, video, PDF). |
| Gemini 3.6 Flash | 1M | Best everyday Gemini. Fast and cheap with a current knowledge cutoff (March 2026); beats 3.5 Flash on coding and long-context. |
| Gemini 3.5 Flash | 1M | Previous Flash generation. Superseded by 3.6 Flash, which is both cheaper and better. |