AI model comparison
Compare the major AI models side by side — Claude Opus 5, GPT-5.6, Gemini 3.1 Pro, Grok, Kimi K3 and more — by context window, cost, modalities, and what each is genuinely best for. Calm, sourced, and free.
Cost is shown as a tier, not a live price — model pricing moves often, so always confirm current rates on the provider's page before you commit. Last updated 2026-08-14.
| Model | Type | Context | Cost | Modalities | Open | Best for |
|---|---|---|---|---|---|---|
| Claude FableAnthropic · Fable 5 | Frontier | 1M | $$$$ | text, vision | — | The ceiling — problems where being right matters more than the bill |
| Claude OpusAnthropic · Opus 5 | Frontier | 1M | $$$ | text, vision | — | Agentic coding and multi-step engineering work — near the ceiling, at half the price |
| GPT-5.6 SolOpenAI | Frontier | 1M | $$$$ | text, vision, audio | — | The strongest generalist — complex agentic and scientific work across modalities |
| Gemini ProGoogle · 3.1 Pro | Frontier | 1M | $$ | text, vision, audio | — | Reasoning over very long inputs — whole codebases, long documents, video |
| GrokxAI · 4.6 | Frontier | 500K | $$ | text, vision | — | Agentic coding with tool use, especially inside xAI's own tooling |
| Claude SonnetAnthropic · Sonnet 5 | Balanced | 1M | $$ | text, vision | — | The default for most work — reads code like a senior engineer at a fair price |
| GPT-5.6 TerraOpenAI | Balanced | 1M | $$ | text, vision, audio | — | Production workloads that need frontier-family quality without the flagship bill |
| Gemini FlashGoogle · 3.7 Flash | Fast & cheap | 1M | $ | text, vision, audio | — | High-volume agent loops and prototyping — cheap enough to run constantly |
| GPT-5.6 LunaOpenAI | Fast & cheap | 1M | $ | text, vision | — | Classification, routing, and extraction at volume |
| Claude HaikuAnthropic · Haiku 4.5 | Fast & cheap | 200K | $ | text, vision | — | High-volume, low-latency work where speed and cost matter most |
| DeepSeekDeepSeek | Reasoning | 128K | $ | text | Open reasoning at a fraction of the cost of closed reasoning models | |
| KimiMoonshot AI · K3 | Open weights | 1M | $ | text, vision | Frontier-adjacent quality you can self-host — the open model closest to the closed leaders | |
| QwenAlibaba | Open weights | 128K | Free tier | text, vision | A strong open model for local use, especially good at coding for its size | |
| LlamaMeta | Open weights | 128K | Free tier | text, vision | Running locally or self-hosting — private, free to run, no vendor lock-in | |
| MistralMistral AI | Open weights | 256K | $ | text, vision | Efficient European open models for building without heavy lock-in |
The ceiling — problems where being right matters more than the bill
Agentic coding and multi-step engineering work — near the ceiling, at half the price
The strongest generalist — complex agentic and scientific work across modalities
Reasoning over very long inputs — whole codebases, long documents, video
Agentic coding with tool use, especially inside xAI's own tooling
The default for most work — reads code like a senior engineer at a fair price
Production workloads that need frontier-family quality without the flagship bill
High-volume agent loops and prototyping — cheap enough to run constantly
Classification, routing, and extraction at volume
High-volume, low-latency work where speed and cost matter most
Open reasoning at a fraction of the cost of closed reasoning models
Frontier-adjacent quality you can self-host — the open model closest to the closed leaders
A strong open model for local use, especially good at coding for its size
Running locally or self-hosting — private, free to run, no vendor lock-in
Efficient European open models for building without heavy lock-in
Head to head
The matrix answers “what exists”. These answer “which one for me” — each with the honest tradeoff, not a leaderboard.
- Claude Opus vs GPT-5.6 Sol
- Claude Opus vs Gemini Pro
- GPT-5.6 Sol vs Gemini Pro
- Claude Opus vs Grok
- GPT-5.6 Sol vs Grok
- Gemini Pro vs Grok
- Claude Sonnet vs GPT-5.6 Terra
- Claude Sonnet vs Gemini Pro
- Claude Fable vs Claude Opus
- Claude Opus vs Claude Sonnet
- GPT-5.6 Sol vs GPT-5.6 Terra
- Claude Haiku vs Gemini Flash
- GPT-5.6 Luna vs Gemini Flash
- Claude Haiku vs GPT-5.6 Luna
- Claude Opus vs Kimi
- GPT-5.6 Sol vs Kimi
- DeepSeek vs Kimi
- DeepSeek vs Qwen
- Llama vs Qwen
- Llama vs Mistral
Go deeper
The matrix is the quick answer. For the reasoning behind a pick, read Claude vs GPT-5.6, Claude vs Gemini, or running open models locally. Every category above — Frontier, Balanced, Fast & cheap, Reasoning, Open weights — maps to a use case, not a leaderboard rank.
Embed this comparison
Free to use on your own site — paste this snippet where you want the live, auto-updating matrix to appear.
<iframe src="https://kapyn.app/embed/compare" width="100%" height="640" style="border:1px solid #222;border-radius:14px" title="AI Model Comparison by Kapyn" loading="lazy"></iframe>