kapyn
Compare

AI model comparison

Compare the major AI models side by side, Claude Opus 5, GPT-5.6, Gemini 3.1 Pro, Grok, Kimi K3 and more , by context window, cost, modalities, and what each is genuinely best for. Calm, sourced, and free.

Cost is shown as a tier, not a live price, model pricing moves often, so always confirm current rates on the provider's page before you commit. Last updated 2026-09-04.

ModelTypeContextCostModalitiesOpenBest for
Claude FableAnthropic · Fable 5.1Frontier1M$$$$text, vision, Demanding reasoning and long-horizon agentic work, when Opus at high effort still falls short
Claude OpusAnthropic · Opus 5Frontier1M$$$text, vision, Agentic coding and multi-step engineering work, near the ceiling, at half the price
GPT AstraOpenAI · GPT-6 AstraFrontier1M$$$$text, vision, The hardest end-to-end work, computer use, and long agentic runs that span applications
GPT SolOpenAI · GPT-5.6 SolFrontier1M$$$text, vision, Frontier-grade generalist work now that Astra has taken the top of the line
Gemini ProGoogle · 3.1 ProFrontier1M$$text, vision, audio, Reasoning over very long inputs, whole codebases, long documents, video
GrokxAI · 4.6Frontier500K$$text, vision, Agentic coding with tool use, especially inside xAI's own tooling
Claude SonnetAnthropic · Sonnet 5Balanced1M$$text, vision, The default for most work, reads code like a senior engineer at a fair price
GPT TerraOpenAI · GPT-5.6 TerraBalanced1M$$text, vision, Production workloads that need frontier-family quality without the flagship bill
Gemini FlashGoogle · 3.8 FlashFast & cheap1M$text, vision, audio, High-volume agent loops and prototyping, cheap enough to run constantly
GPT LunaOpenAI · GPT-5.6 LunaFast & cheap1M$text, vision, Classification, routing, and extraction at volume
Claude HaikuAnthropic · Haiku 4.5Fast & cheap200K$text, vision, High-volume, low-latency work where speed and cost matter most
DeepSeekDeepSeek · V4Reasoning1M$textOpen reasoning at a fraction of the cost of closed reasoning models
KimiMoonshot AI · K3Open weights1M$text, visionFrontier-adjacent quality you can self-host, the open model closest to the closed leaders
QwenAlibaba · Qwen3.8Open weights256KFree tiertext, visionA strong open model for local use, especially good at coding for its size
LlamaMeta · Llama 4Open weights10MFree tiertext, visionThe widest open ecosystem, tooling and fine-tunes, though Meta's newer open work ships as Muse
GLMZ.ai · GLM-5.3Open weights1M$text, visionOpen coding and agentic work at close to closed-model quality
Muse GlimmerMeta · Muse Glimmer 30BOpen weights128KFree tiertext, visionAlways-on local agents, 30B distilled from Muse Spark to fit one consumer GPU
MistralMistral AI · Medium 3.5Open weights256K$text, visionEfficient European open models for building without heavy lock-in
Claude FableFrontier
Anthropic · Fable 5.1

Demanding reasoning and long-horizon agentic work, when Opus at high effort still falls short

Context 1MCost $$$$text · vision
Claude OpusFrontier
Anthropic · Opus 5

Agentic coding and multi-step engineering work, near the ceiling, at half the price

Context 1MCost $$$text · vision
GPT AstraFrontier
OpenAI · GPT-6 Astra

The hardest end-to-end work, computer use, and long agentic runs that span applications

Context 1MCost $$$$text · vision
GPT SolFrontier
OpenAI · GPT-5.6 Sol

Frontier-grade generalist work now that Astra has taken the top of the line

Context 1MCost $$$text · vision
Gemini ProFrontier
Google · 3.1 Pro

Reasoning over very long inputs, whole codebases, long documents, video

Context 1MCost $$text · vision · audio
GrokFrontier
xAI · 4.6

Agentic coding with tool use, especially inside xAI's own tooling

Context 500KCost $$text · vision
Anthropic · Sonnet 5

The default for most work, reads code like a senior engineer at a fair price

Context 1MCost $$text · vision
GPT TerraBalanced
OpenAI · GPT-5.6 Terra

Production workloads that need frontier-family quality without the flagship bill

Context 1MCost $$text · vision
Gemini FlashFast & cheap
Google · 3.8 Flash

High-volume agent loops and prototyping, cheap enough to run constantly

Context 1MCost $text · vision · audio
GPT LunaFast & cheap
OpenAI · GPT-5.6 Luna

Classification, routing, and extraction at volume

Context 1MCost $text · vision
Claude HaikuFast & cheap
Anthropic · Haiku 4.5

High-volume, low-latency work where speed and cost matter most

Context 200KCost $text · vision
DeepSeekReasoning
DeepSeek · V4

Open reasoning at a fraction of the cost of closed reasoning models

Context 1MCost $textOpen weights
KimiOpen weights
Moonshot AI · K3

Frontier-adjacent quality you can self-host, the open model closest to the closed leaders

Context 1MCost $text · visionOpen weights
QwenOpen weights
Alibaba · Qwen3.8

A strong open model for local use, especially good at coding for its size

Context 256KFree tiertext · visionOpen weights
LlamaOpen weights
Meta · Llama 4

The widest open ecosystem, tooling and fine-tunes, though Meta's newer open work ships as Muse

Context 10MFree tiertext · visionOpen weights
GLMOpen weights
Z.ai · GLM-5.3

Open coding and agentic work at close to closed-model quality

Context 1MCost $text · visionOpen weights
Muse GlimmerOpen weights
Meta · Muse Glimmer 30B

Always-on local agents, 30B distilled from Muse Spark to fit one consumer GPU

Context 128KFree tiertext · visionOpen weights
MistralOpen weights
Mistral AI · Medium 3.5

Efficient European open models for building without heavy lock-in

Context 256KCost $text · visionOpen weights

Head to head

The matrix answers “what exists”. These answer “which one for me” , each with the honest tradeoff, not a leaderboard.

Go deeper

The matrix is the quick answer. For the reasoning behind a pick, read Claude vs GPT, Claude vs Gemini, or running open models locally. Every category above , Frontier, Balanced, Fast & cheap, Reasoning, Open weights , maps to a use case, not a leaderboard rank.

Embed this comparison

Free to use on your own site , paste this snippet where you want the live, auto-updating matrix to appear.

<iframe src="https://kapyn.app/embed/compare" width="100%" height="640" style="border:1px solid #222;border-radius:14px" title="AI Model Comparison by Kapyn" loading="lazy"></iframe>