27 reviewed model records

AI model comparison finder

Find models that meet your actual constraints, then compare their context, capabilities, pricing, distribution, and lifecycle side by side. Every model card links back to a first-party source.

No sponsored rankingsNo login or API keySources checked July 23, 2026

Finder

Narrow the catalog

Results use reviewed first-party model facts, not benchmark rankings.

Side by side

Compare selected models

3 of 4 selected

Side-by-side comparison of GPT-5.6 Terra, Claude Sonnet 5, Gemini 3.6 Flash
Attribute

OpenAI

GPT-5.6 Terra

Anthropic

Claude Sonnet 5

Google

Gemini 3.6 Flash

Lifecycle
Active

No retirement announced

Active

Available until at least Jun 30, 2027

Active

No retirement announced

API / checkpoint IDgpt-5.6-terraclaude-sonnet-5gemini-3.6-flash
DistributionHosted APIHosted APIHosted API
Context window1.05M tokens1M tokens1.05M tokens
Maximum output128K tokens128K tokens65.5K tokens
Modalitiestext, imagetexttext, imagetexttext, image, video, audio, pdftext
Capabilities
reasoningfunction callingstructured outputsstreamingtool use
adaptive thinkingvisiontool useprompt cachingbatch processing
thinkingfunction callingstructured outputscode executionsearch groundingcontext caching
Standard text price

$2.50 input / $15 output

USD per 1M tokens

$2 input / $10 output

USD per 1M tokens

$1.50 input / $7.50 output

USD per 1M tokens

Open-weight detailsHosted API onlyHosted API onlyHosted API only
Official sourceGPT-5.6 Terra model cardClaude models overviewGemini 3.6 Flash model card

Matching models

17 of 27 reviewed records

Anthropic

Claude Fable 5

claude-fable-5
Active
Context
1M
Max output
128K
Distribution
Hosted API
Input / output price
$10 / $50

Available until at least Jun 9, 2027

text inputimage inputadaptive thinkingvisiontool use
Official source ↗

Anthropic

Claude Haiku 4.5

claude-haiku-4-5-20251001
Active
Context
200K
Max output
64K
Distribution
Hosted API
Input / output price
$1 / $5

Available until at least Oct 15, 2026

text inputimage inputextended thinkingvisiontool use
Official source ↗

Anthropic

Claude Opus 4.8

claude-opus-4-8
Active
Context
1M
Max output
128K
Distribution
Hosted API
Input / output price
$5 / $25

Available until at least May 28, 2027

text inputimage inputadaptive thinkingvisiontool use
Official source ↗

Anthropic

Claude Sonnet 5

claude-sonnet-5
Active
Context
1M
Max output
128K
Distribution
Hosted API
Input / output price
$2 / $10

Available until at least Jun 30, 2027

text inputimage inputadaptive thinkingvisiontool use

Google

Gemini 3.5 Flash

gemini-3.5-flash
Active
Context
1.05M
Max output
65.5K
Distribution
Hosted API
Input / output price
$1.50 / $9

No retirement announced

text inputimage inputvideo inputaudio inputpdf inputthinkingfunction callingstructured outputs
Official source ↗

Google

Gemini 3.5 Flash-Lite

gemini-3.5-flash-lite
Active
Context
1.05M
Max output
65.5K
Distribution
Hosted API
Input / output price
$0.30 / $2.50

No retirement announced

text inputimage inputvideo inputaudio inputpdf inputthinkingfunction callingstructured outputs
Official source ↗

Google

Gemini 3.6 Flash

gemini-3.6-flash
Active
Context
1.05M
Max output
65.5K
Distribution
Hosted API
Input / output price
$1.50 / $7.50

No retirement announced

text inputimage inputvideo inputaudio inputpdf inputthinkingfunction callingstructured outputs

Google

Gemma 4 31B

gemma-4-31b-it
Active
Context
256K
Max output
Not published
Distribution
API + open weights
Input / output price
Varies by host

No retirement announced

text inputimage inputreasoningfunction callingcoding
Official source ↗

OpenAI

GPT-5.6 Luna

gpt-5.6-luna
Active
Context
1.05M
Max output
128K
Distribution
Hosted API
Input / output price
$1 / $6

No retirement announced

text inputimage inputreasoningfunction callingstructured outputs
Official source ↗

OpenAI

GPT-5.6 Sol

gpt-5.6-sol
Active
Context
1.05M
Max output
128K
Distribution
Hosted API
Input / output price
$5 / $30

No retirement announced

text inputimage inputreasoningfunction callingstructured outputs
Official source ↗

OpenAI

GPT-5.6 Terra

gpt-5.6-terra
Active
Context
1.05M
Max output
128K
Distribution
Hosted API
Input / output price
$2.50 / $15

No retirement announced

text inputimage inputreasoningfunction callingstructured outputs

Mistral AI

Mistral Small 4

mistral-small-2603+1
Active
Context
256K
Max output
Not published
Distribution
API + open weights
Input / output price
$0.15 / $0.60

No retirement announced

text inputimage inputreasoningfunction callingstructured outputs
Official source ↗

OpenAI

gpt-oss-120b

gpt-oss-120b
Active
Context
131.1K
Max output
131.1K
Distribution
Open weights
Input / output price
Varies by host

No retirement announced

text inputreasoningfunction callingstructured outputs
Official source ↗

OpenAI

gpt-oss-20b

gpt-oss-20b
Active
Context
131.1K
Max output
131.1K
Distribution
Open weights
Input / output price
Varies by host

No retirement announced

text inputreasoningfunction callingstructured outputs
Official source ↗

Meta

Llama 4 Maverick

llama-4-maverick
Active
Context
1M
Max output
Not published
Distribution
Open weights
Input / output price
Varies by host

No retirement announced

text inputimage inputmixture of expertsmultilingualself-hosting
Official source ↗

Meta

Llama 4 Scout

llama-4-scout
Active
Context
10M
Max output
Not published
Distribution
Open weights
Input / output price
Varies by host

No retirement announced

text inputimage inputmixture of expertsmultilingualself-hosting
Official source ↗

Google

Gemini 3.1 Pro Preview

gemini-3.1-pro-preview
Preview
Context
1.05M
Max output
65.5K
Distribution
Hosted API
Input / output price
$2 / $12

No retirement date published

text inputimage inputvideo inputaudio inputpdf inputthinkingfunction callingstructured outputs
Official source ↗

Specifications narrow a shortlist; they do not rank quality.

Benchmarks, latency, rate limits, regional availability, safety behavior, and real workload quality still require testing. For workload-specific token costs, use the API cost and context planner.

Method and limits

Facts are versioned; aliases and prices can move

The catalog is a curated snapshot verified on July 23, 2026. It preserves model IDs, effective dates, caveats, and source links rather than silently treating a rolling alias as a permanent checkpoint.

Prices shown are standard base text-token rates where the provider publishes one. Long-context tiers, regions, caching, batches, media, tools, fine-tuning, and hosted open-model prices may differ. Confirm the linked source before committing a budget or migration.

Questions about comparing AI models

What does the AI model comparison include?

The finder compares first-party published facts: lifecycle state, model or deployment ID, context window, maximum output, input and output modalities, selected capabilities, standard text-token pricing, distribution, parameter counts, and licenses where applicable.

Does the cheapest model always cost less for a real workload?

No. Token counts, prompt caching, batch discounts, long-context tiers, tool calls, images, audio, reasoning tokens, retries, and provider-specific fees can change the bill. Use the linked cost planner with a realistic workload.

Does open-weight mean open-source?

Not necessarily. Open-weight means downloadable model weights are available. Each model still has its own license, and some community licenses impose conditions that differ from standard open-source licenses.

Which model is best?

There is no universal winner. This tool narrows models by documented constraints; it does not invent one composite quality score. Test the finalists on representative prompts and measure quality, latency, reliability, and total cost.