Model directory · updated 10 July 2026

Every AI model, compared without the vendor fog.

Specs, pricing, context windows and plain-English verdicts for the frontier models people are actually choosing between.

1M+token-class models tracked
$0.14+lowest 1M-token input cost
Jul 2026fresh benchmark cycle

Latest AI models

A compact index for model selection: maker, fit, price posture and launch window.

Jump to comparisons
SpaceXAI
New

Grok 4.5

The "Opus-class" workhorse, released 9 July — $2/$6 per 1M tokens, 500K context, ~4x fewer output tokens per task than Opus 4.8.

500K contextToken efficiency
Workhorse · Jul 2026
OpenAI
API TBD

GPT-Live (1 & 1 mini)

Full-duplex voice models that listen while they speak — now powering ChatGPT Voice, delegating hard questions to GPT-5.5. API "soon".

Voice · full-duplexIn ChatGPT now
Voice · Jul 2026
Anthropic
Frontier

Claude Fable 5

Anthropic's most capable released model — pulled under US export controls, back globally 1 July. $10/$50, 1M context.

1M contextHardest tasks
Frontier · Jun 2026
Anthropic
New

Claude Sonnet 5

Mid-tier with 1M context by default and a real agentic step up. Intro pricing $2/$10 through 31 Aug, then $3/$15.

1M contextPrice-performance
Mid-tier · Jun 2026
OpenAI
New

GPT-5.6 (Sol · Terra · Luna)

OpenAI's new flagship family, GA 9 July — 1.05M context, $5/$30 (Sol), split benchmark record vs Claude Fable 5.

1.05M context3 variants
Flagship · Jul 2026
Anthropic
Flagship

Claude Opus 4.8

Current Opus flagship — same $5/$25 rate card as 4.7, better benchmarks, fast mode now 3x cheaper at $10/$50.

1M contextAgentic work
Flagship · May 2026
Google
New

Gemini 3.5 Flash

Beats Gemini 3.1 Pro on coding at 25% lower cost — $1.50/$9, 1M context, 76.2% Terminal-Bench 2.1.

1M contextAgentic coding
Production · May 2026
Alibaba
Flagship

Qwen 3.7-Max

Alibaba's proprietary top model — ~1T-param MoE, 1M context, built for hours-long agent runs at $2.50/$7.50.

1M contextLong agents
Flagship · May 2026
ByteDance
Agents

Doubao Seed 2.1

Pro & Turbo variants behind China's #1 AI chatbot (155M WAU). Agent-focused, ¥6/¥30 per 1M tokens.

Agent focusLow cost
Agents · Jun 2026
Mistral AI
Efficient

Mistral Small 4

Reasoning, vision and coding unified in one 119B MoE with just 6B active params per token.

6B active3-in-1
Efficient · Mar 2026
Anthropic
Budget

Claude Haiku 4.5

Sonnet 4-level coding at one-third the cost and 2x+ the speed. $1/$5, 200K context, the subagent workhorse.

$1/$5 pricingSpeed tier
Budget · Oct 2025
OpenAI
Prev gen

GPT-5.5 "Spud"

First full retrain since GPT-4.5. 1M context, $5/$30 pricing, 82.7% on Terminal-Bench 2.0. Superseded by GPT-5.6 at the same price.

1M contextAgentic coding
Prev gen · Apr 2026
Google
Multimodal

Gemini 3.1 Ultra

2M-token context — the largest available. Native video, audio and text in one pass, no transcription step.

2M contextMedia work
Multimodal · Apr 2026
DeepSeek
Open weights

DeepSeek V4

Largest open-weights model of 2026 — 1.6T params, MIT-licensed, 80.6% SWE-bench, from $0.14/1M tokens.

MIT licenseLow cost
Open weights · Apr 2026
Google
Open

Gemma 4

Open Apache 2.0 family — four sizes from phone to workstation, 89.2% on AIME, runs on one GPU.

Apache 2.0On-device
Open source · Apr 2026
Moonshot AI
Coding

Kimi K2.6

Open-weight 1T-param coding model — ties GPT-5.5 on SWE-bench Pro at ~80% lower cost.

1T paramsCode tasks
Coding · Apr 2026
Anysphere
Tool

Cursor 3

AI code editor rebuilt around agent orchestration — the new Agents Window. Shipped 2 Apr 2026.

Agent IDEDev workflow
Dev tool · Apr 2026
Anthropic
Prev gen

Claude Opus 4.7

87.6% on SWE-bench Verified. 1M context, $5/$25 pricing. Superseded by Opus 4.8 at the same price.

1M contextCode quality
Prev gen · Apr 2026
SpaceXAI
Unreleased

Grok 5

Reportedly 6T parameters, trained on Colossus 2. Q1 and Q2 2026 targets both missed — Grok 4.5 shipped instead; still no date.

WatchlistFrontier scale
Unreleased · TBD
Ad slot — Google AdSense responsive unit goes here once approved

Free tools

Calculators built directly on the model index — same data, zero guesswork.

Head-to-head comparisons

The pairings that decide real architecture choices: quality margin, context, cost and deployment fit.

Comparison
New

GPT-5.6 vs Claude Fable 5

Half the price and a genuine benchmark split — Sol leads Terminal-Bench, Fable 5 leads SWE-bench Pro.

Updated Jul 2026
Comparison
New

GPT-5.6 vs GPT-5.5

Same $5/$30 price, higher published scores — the straightforward upgrade path.

Updated Jul 2026
Comparison
New

GPT-5.6 vs Grok 4.5

A near-exact SWE-bench Pro tie (64.6% vs 64.7%) next to a 2.5-5x price gap.

Updated Jul 2026
Comparison
New

Grok 4.5 vs Claude Opus 4.8

The "Opus-class" claim, tested — a 2-2 benchmark split at a 2.5-4x price gap.

Updated Jul 2026
Comparison
New

Claude Sonnet 5 vs GPT-5.5

Mid-tier challenger vs shipping flagship — a 2.5-3x price gap and matching 1M contexts.

Updated Jul 2026
Comparison
Frontier

Claude Fable 5 vs Opus 4.8

Anthropic's two top models at a clean 2x price gap — when does the Fable premium pay off?

Updated Jul 2026
Comparison
Upgrade

Claude Opus 4.8 vs 4.7

Same price, better benchmarks, 3x cheaper fast mode — the upgrade math, explained.

Updated Jul 2026
Comparison
Google

Gemini 3.5 Flash vs 3.1

The new Flash beats 3.1 Pro on coding at 25% lower cost. Where 3.1 Ultra still wins.

Updated Jul 2026
Comparison
Flagship

GPT-5.5 vs Gemini 3.1

Which flagship wins on coding, context window, multimodal ability and price?

Updated May 2026
Comparison
Cost

DeepSeek V4 vs GPT-5.5

Open MIT weights vs closed flagship — an 8-10x price gap and who each suits.

Updated May 2026
Comparison
Coding

Claude Opus 4.7 vs GPT-5.5

Code quality vs agentic reliability — Opus leads SWE-bench Pro, GPT-5.5 owns Codex.

Updated May 2026
Comparison
Open

Kimi K2.6 vs DeepSeek V4

Two open-weight coding models head to head — context, price and licensing.

Updated May 2026
Comparison
Multimodal

Claude Opus 4.7 vs Gemini 3.1

Code quality vs multimodal scale — 1M vs 2M context, and a 2x price gap.

Updated May 2026

Why AI Model Hub

New AI models ship almost every week, each with its own benchmarks, pricing tiers and marketing language. AI Model Hub cuts through it: every page gives you the specs, the pricing, the benchmark numbers and a plain-English read on who the model is actually for.

  • Independent model pages, no provider affiliation.
  • Comparison pages built around actual selection tradeoffs.
  • Refreshed whenever a material model release lands.