§ AI Model Tracker

The current state of AI models.

A live catalog of 347 models from OpenRouter, release timeline from Epoch AI, and geography of origin. Pricing shown in your currency of choice — real-time USD from OpenRouter, CAD converted at the Bank of Canada noon rate.

Refreshed 2026-07-10
FX: 1 USD = 1.4169 CAD (2026-07-09)
347
Models tracked
Live from OpenRouter
47%
Open-weights share
163 of 347
44.1%
From United States
Leading country of origin
2026-06-09
Latest notable release
Claude Fable 5
§ Release timeline

Release cadence, month by month

166 notable language models tracked by Epoch AI over the last 24 months. Chart = total per month; notable releases listed below with dates.

Notable releases in this window
David’s take

The line is fairly flat — 6-10 notable language models per month over the last two years, with occasional spikes when a lab clusters releases (Anthropic and DeepSeek both did this in mid-2026). The steady pace matters more than the peaks: the frontier isn't sprinting, it's compounding.

§ Where they come from

Geography of the AI model landscape

Country of origin for the 347 models in the OpenRouter catalog. Editorial signal, not a scientific measure — some labs release many variants of one architecture.

🇺🇸 United States
44.1%
153 models
🇨🇳 China
26.2%
91 models
🌐 Other
22.8%
79 models
🇫🇷 France
5.5%
19 models
🇨🇦 Canada
1.4%
5 models
David’s take

The US still supplies the plurality of models, but China is now solidly #2 at over a quarter of the catalog — three years ago that was under 5%. The share is driven by open-weights releases from DeepSeek, Alibaba (Qwen), Tencent, and MiniMax. That's the story worth tracking: not who leads on any given benchmark, but who's shipping models people can actually deploy.

§ The catalog

347 of 347 models

Filter to what you're trying to compare. CAD pricing per million tokens. Sorted by input price, cheapest first.

Tier
Weights
ModelOriginTierWeightsContextCAD /M inputCAD /M outputToolsVision
OpenRouter: Fusion
Openrouter
🌐BudgetClosed1,000,000$-1416900.0000$-1416900.0000
Pareto Code Router
Openrouter
🌐BudgetClosed2,000,000$-1416900.0000$-1416900.0000
Body Builder (beta)
Openrouter
🌐BudgetClosed128,000$-1416900.0000$-1416900.0000
Auto Router
Openrouter
🌐BudgetClosed2,000,000$-1416900.0000$-1416900.0000
Tencent: Hy3 (free)
Tencent
🇨🇳BudgetOpen262,144$0.0000$0.0000
Poolside: Laguna XS 2.1 (free)
Poolside
🌐BudgetOpen262,144$0.0000$0.0000
Cohere: North Mini Code (free)
Cohere
🇨🇦BudgetOpen256,000$0.0000$0.0000
NVIDIA: Nemotron 3.5 Content Safety (free)
NVIDIA
🇺🇸BudgetOpen128,000$0.0000$0.0000
NVIDIA: Nemotron 3 Ultra (free)
NVIDIA
🇺🇸BudgetOpen1,000,000$0.0000$0.0000
NVIDIA: Nemotron 3 Nano Omni (free)
NVIDIA
🇺🇸BudgetOpen256,000$0.0000$0.0000
Poolside: Laguna M.1 (free)
Poolside
🌐BudgetOpen262,144$0.0000$0.0000
Google: Gemma 4 26B A4B (free)
Google
🇺🇸BudgetOpen262,144$0.0000$0.0000
Google: Gemma 4 31B (free)
Google
🇺🇸BudgetOpen262,144$0.0000$0.0000
Google: Lyria 3 Pro Preview
Google
🇺🇸BudgetClosed1,048,576$0.0000$0.0000
Google: Lyria 3 Clip Preview
Google
🇺🇸BudgetClosed1,048,576$0.0000$0.0000
NVIDIA: Nemotron 3 Super (free)
NVIDIA
🇺🇸BudgetOpen1,000,000$0.0000$0.0000
Free Models Router
Openrouter
🌐BudgetClosed200,000$0.0000$0.0000
LiquidAI: LFM2.5-1.2B-Thinking (free)
Liquid
🌐BudgetOpen32,768$0.0000$0.0000
LiquidAI: LFM2.5-1.2B-Instruct (free)
Liquid
🌐BudgetOpen32,768$0.0000$0.0000
NVIDIA: Nemotron 3 Nano 30B A3B (free)
NVIDIA
🇺🇸BudgetOpen256,000$0.0000$0.0000
NVIDIA: Nemotron Nano 12B 2 VL (free)
NVIDIA
🇺🇸BudgetOpen128,000$0.0000$0.0000
Qwen: Qwen3 Next 80B A3B Instruct (free)
Qwen (Alibaba)
🇨🇳BudgetOpen262,144$0.0000$0.0000
NVIDIA: Nemotron Nano 9B V2 (free)
NVIDIA
🇺🇸BudgetOpen128,000$0.0000$0.0000
OpenAI: gpt-oss-120b (free)
OpenAI
🇺🇸BudgetOpen131,072$0.0000$0.0000
OpenAI: gpt-oss-20b (free)
OpenAI
🇺🇸BudgetOpen131,072$0.0000$0.0000
Qwen: Qwen3 Coder 480B A35B (free)
Qwen (Alibaba)
🇨🇳BudgetOpen1,048,576$0.0000$0.0000
Venice: Uncensored (free)
Cognitivecomputations
🌐BudgetOpen32,768$0.0000$0.0000
Meta: Llama 3.3 70B Instruct (free)
Meta
🇺🇸BudgetOpen131,072$0.0000$0.0000
Meta: Llama 3.2 3B Instruct (free)
Meta
🇺🇸BudgetOpen131,072$0.0000$0.0000
Nous: Hermes 3 405B Instruct (free)
Nousresearch
🌐BudgetOpen131,072$0.0000$0.0000
inclusionAI: Ling-2.6-flash
Inclusionai
🌐BudgetClosed262,144$0.0142$0.0425
IBM: Granite 4.0 Micro
Ibm Granite
🌐BudgetOpen131,000$0.0241$0.1587
Meta: Llama 3.1 8B Instruct
Meta
🇺🇸BudgetOpen131,072$0.0283$0.0425
Mistral: Mistral Nemo
Mistral
🇫🇷BudgetOpen131,072$0.0283$0.0425
Nex AGI: Nex-N2-Mini
Nex Agi
🌐BudgetOpen262,144$0.0354$0.1417
Meta: Llama 3.2 1B Instruct
Meta
🇺🇸BudgetOpen131,072$0.0383$0.2848
OpenAI: gpt-oss-20b
OpenAI
🇺🇸BudgetOpen131,072$0.0411$0.1984
Amazon: Nova Micro 1.0
Amazon
🇺🇸BudgetClosed128,000$0.0496$0.1984
OpenAI: gpt-oss-120b
OpenAI
🇺🇸BudgetOpen131,072$0.0510$0.2550
Cohere: Command R7B (12-2024)
Cohere
🇨🇦BudgetClosed128,000$0.0531$0.2125
Qwen: Qwen2.5 7B Instruct
Qwen (Alibaba)
🇨🇳BudgetOpen131,072$0.0567$0.1417
Sao10K: Llama 3 8B Lunaris
Sao10k
🌐BudgetOpen8,192$0.0567$0.0708
Arcee AI: Trinity Mini
Arcee Ai
🌐BudgetOpen131,072$0.0638$0.2125
Qwen: Qwen3 30B A3B Instruct 2507
Qwen (Alibaba)
🇨🇳BudgetOpen131,072$0.0682$0.2735
IBM: Granite 4.1 8B
Ibm Granite
🌐BudgetOpen131,072$0.0708$0.1417
NVIDIA: Nemotron 3 Nano 30B A3B
NVIDIA
🇺🇸BudgetOpen262,144$0.0708$0.2834
OpenAI: GPT-5 Nano
OpenAI
🇺🇸BudgetClosed400,000$0.0708$0.5668
Google: Gemma 3 4B
Google
🇺🇸BudgetOpen131,072$0.0708$0.1417
Google: Gemma 3 12B
Google
🇺🇸BudgetOpen131,072$0.0708$0.2125
Mistral: Mistral Small 3
Mistral
🇫🇷BudgetOpen32,768$0.0708$0.1134
Meta: Llama 3.2 3B Instruct
Meta
🇺🇸BudgetOpen131,072$0.0708$0.4676
Poolside: Laguna XS 2.1
Poolside
🌐BudgetOpen262,144$0.0850$0.1700
Google: Gemma 4 26B A4B
Google
🇺🇸BudgetOpen262,144$0.0850$0.4676
Z.ai: GLM 4.7 Flash
Z.AI
🇨🇳BudgetOpen202,752$0.0850$0.5668
Google: Gemma 3n 4B
Google
🇺🇸BudgetOpen32,768$0.0850$0.1700
Amazon: Nova Lite 1.0
Amazon
🇺🇸BudgetClosed300,000$0.0850$0.3401
MythoMax 13B
Gryphe
🌐BudgetOpen4,096$0.0850$0.0850
Tencent: Hy3 preview
Tencent
🇨🇳BudgetOpen262,144$0.0893$0.2975
Qwen: Qwen3.5-Flash
Qwen (Alibaba)
🇨🇳BudgetClosed1,000,000$0.0921$0.3684
Qwen: Qwen3 Coder 30B A3B Instruct
Qwen (Alibaba)
🇨🇳BudgetOpen160,000$0.0992$0.3826
Microsoft: Phi 4
Microsoft
🇺🇸BudgetOpen16,384$0.0992$0.1984
inclusionAI: Ring-2.6-1T
Inclusionai
🌐BudgetClosed262,144$0.1063$0.8856
inclusionAI: Ling-2.6-1T
Inclusionai
🌐BudgetClosed262,144$0.1063$0.8856
ByteDance Seed: Seed 1.6 Flash
Bytedance Seed
🌐BudgetClosed262,144$0.1063$0.4251
OpenAI: gpt-oss-safeguard-20b
OpenAI
🇺🇸BudgetOpen131,072$0.1063$0.4251
Mistral: Mistral Small 3.2 24B
Mistral
🇫🇷BudgetOpen128,000$0.1063$0.2834
NVIDIA: Nemotron 3 Super
NVIDIA
🇺🇸BudgetOpen1,000,000$0.1134$0.6376
Qwen: Qwen3 32B
Qwen (Alibaba)
🇨🇳BudgetOpen131,072$0.1134$0.3967
Google: Gemma 3 27B
Google
🇺🇸BudgetOpen131,072$0.1134$0.2267
DeepSeek: DeepSeek V4 Flash
DeepSeek
🇨🇳BudgetOpen1,048,576$0.1275$0.2550
Qwen: Qwen3 Next 80B A3B Instruct
Qwen (Alibaba)
🇨🇳BudgetOpen262,144$0.1275$1.5586
Qwen: Qwen3 235B A22B Instruct 2507
Qwen (Alibaba)
🇨🇳BudgetOpen262,144$0.1275$0.1417
Qwen: Qwen3 Next 80B A3B Thinking
Qwen (Alibaba)
🇨🇳BudgetOpen262,144$0.1381$1.1052
Reka Edge
Rekaai
🌐BudgetOpen16,384$0.1417$0.1417
Qwen: Qwen3.5-9B
Qwen (Alibaba)
🇨🇳BudgetOpen262,144$0.1417$0.2125
ByteDance Seed: Seed-2.0-Mini
Bytedance Seed
🌐BudgetClosed262,144$0.1417$0.5668
StepFun: Step 3.5 Flash
Stepfun
🌐BudgetOpen262,144$0.1417$0.4251
Mistral: Ministral 3 3B 2512
Mistral
🇫🇷BudgetOpen131,072$0.1417$0.1417
Mistral: Voxtral Small 24B 2507
Mistral
🇫🇷BudgetOpen32,000$0.1417$0.4251
ByteDance: UI-TARS 7B
Bytedance
🌐BudgetOpen128,000$0.1417$0.2834
Google: Gemini 2.5 Flash Lite
Google
🇺🇸BudgetClosed1,048,576$0.1417$0.5668
Qwen: Qwen3 14B
Qwen (Alibaba)
🇨🇳BudgetOpen131,702$0.1417$0.3401
OpenAI: GPT-4.1 Nano
OpenAI
🇺🇸BudgetClosed1,047,576$0.1417$0.5668
Meta: Llama 4 Scout
Meta
🇺🇸BudgetOpen10,000,000$0.1417$0.4251
Reka Flash 3
Rekaai
🌐BudgetOpen65,536$0.1417$0.2834
Meta: Llama 3.3 70B Instruct
Meta
🇺🇸BudgetOpen131,072$0.1417$0.4534
Qwen: Qwen3 VL 32B Instruct
Qwen (Alibaba)
🇨🇳BudgetOpen262,144$0.1474$0.5894
Xiaomi: MiMo-V2.5
Xiaomi
🌐BudgetOpen1,048,576$0.1488$0.3967
Qwen: Qwen3 Coder Next
Qwen (Alibaba)
🇨🇳BudgetOpen262,144$0.1559$1.1335
Qwen: Qwen3 VL 8B Thinking
Qwen (Alibaba)
🇨🇳BudgetOpen256,000$0.1658$1.9341
Qwen: Qwen3 VL 8B Instruct
Qwen (Alibaba)
🇨🇳BudgetOpen256,000$0.1658$0.6447
Qwen: Qwen3 8B
Qwen (Alibaba)
🇨🇳BudgetOpen131,072$0.1658$0.6447
Google: Gemma 4 31B
Google
🇺🇸BudgetOpen262,144$0.1700$0.4959
Qwen: Qwen3 30B A3B
Qwen (Alibaba)
🇨🇳BudgetOpen131,072$0.1700$0.7085
Qwen: Qwen3 VL 30B A3B Thinking
Qwen (Alibaba)
🇨🇳BudgetOpen131,072$0.1842$2.2104
Qwen: Qwen3 VL 30B A3B Instruct
Qwen (Alibaba)
🇨🇳BudgetOpen262,144$0.1842$0.7368
Qwen: Qwen3 30B A3B Thinking 2507
Qwen (Alibaba)
🇨🇳BudgetOpen131,072$0.1842$2.2104
Nous: Hermes 4 70B
Nousresearch
🌐BudgetOpen131,072$0.1842$0.5668
Z.ai: GLM 4.5 Air
Z.AI
🇨🇳BudgetOpen131,072$0.1842$1.2044
Tencent: Hy3
Tencent
🇨🇳BudgetOpen262,144$0.1984$0.8218
Showing first 100 of 347 matches · refine filters to narrow
§ What it costs

Popular models: monthly cost side-by-side

Preset workload: 1M input + 500K output tokens per month. Values compute as (input price × 1M) + (output price × 0.5M), then converted to CAD at 1.4169. Popular list is manually curated; every price is real-time from OpenRouter.

Google: Gemini 2.5 Flash LiteGoogleClosed
$0.1417 in · $0.5668 out · 1,048,576 ctx
$0.43
CAD/month
Meta: Llama 4 MaverickMetaOpen
$0.2125 in · $0.8501 out · 1,048,576 ctx
$0.64
CAD/month
DeepSeek: DeepSeek V4 ProDeepSeekOpen
$0.6164 in · $1.2327 out · 1,048,576 ctx
$1.23
CAD/month
OpenAI: GPT-5 MiniOpenAIClosed
$0.3542 in · $2.8338 out · 400,000 ctx
$1.77
CAD/month
Anthropic: Claude Haiku 4.5AnthropicClosed
$1.4169 in · $7.0845 out · 200,000 ctx
$4.96
CAD/month
OpenAI: GPT-5.6 Luna ProOpenAIClosed
$1.4169 in · $8.5014 out · 1,050,000 ctx
$5.67
CAD/month
xAI: Grok 4.5xAIClosed
$2.8338 in · $8.5014 out · 500,000 ctx
$7.08
CAD/month
Google: Gemini 2.5 ProGoogleClosed
$1.7711 in · $14.1690 out · 1,048,576 ctx
$8.86
CAD/month
Anthropic: Claude Sonnet 5AnthropicClosed
$2.8338 in · $14.1690 out · 1,000,000 ctx
$9.92
CAD/month
Anthropic: Claude Opus 4.8AnthropicClosed
$7.0845 in · $35.4225 out · 1,000,000 ctx
$24.80
CAD/month
Anthropic: Claude Fable 5AnthropicClosed
$14.1690 in · $70.8450 out · 1,000,000 ctx
$49.59
CAD/month
David’s take

Same workload, huge spread — Fable 5 is roughly 115× the cost of Gemini Flash Lite for the exact same tokens. For most bulk workloads — content generation, summarization, categorization — the budget tier does the job, and reserving frontier models like Opus and Fable for the 5% of work that needs it saves 50-100× on run cost. Volume matters more than model choice for the routine layer.

§ Benchmarks

How they compare on capability

Artificial Analysis composite indices (0-100), pulled live from each model's OpenRouter listing. Intelligence = general reasoning & knowledge · Coding = programming benchmarks · Agentic = multi-step tool use. Some models don't publish scores.

Not shown: Google: Gemini 2.5 Flash Lite, OpenAI: GPT-5.6 Luna Pro — no Artificial Analysis composite index published yet.
Source: Artificial Analysis ↗ · Pulled through OpenRouter model metadata
David’s take

Fable 5 leads all three benchmarks — but so does its price tag ($49.59for our preset workload). The interesting cluster is the second tier: Opus, Grok, Sonnet all land within 3 points of each other on intelligence, but Sonnet costs a third of Opus and less than half of Grok. That's the tradeoff to actually reason about — a rounding-error capability gap for meaningful cost savings.

§ Value

Capability vs. cost

Every popular model plotted by Artificial Analysis intelligence score (right = smarter) against monthly cost for the preset workload (bottom = cheaper, log scale). The value frontier lives in the bottom-right corner.

Closed weightsOpen weights
What each zone means
Value frontier
High capability, low cost. The sweet spot for most workloads — bulk content, categorization, extraction, routine reasoning. Where the majority of production traffic should route.
Premium
High capability, high cost. Reserve for the specific tasks that actually need frontier reasoning — hard multi-step problems, complex code, tasks where a smart mistake costs more than the token bill.
Bulk floor
Modest capability, low cost. Good for high-volume repetitive tasks with clear structure — classification, simple summarization, formatting. Under-appreciated for the money.
Avoid
Modest capability, high cost. Usually means legacy pricing or a mismatched positioning. If you find a model here, there's almost always a better option in the other three quadrants.
Zone thresholds: Intelligence 40 · Cost $5 CAD/month. Editorial calls, not universal.
David’s take

Look at the bottom-right of the chart: DeepSeek V4 Pro and Sonnet 5anchor the value frontier — high intelligence scores paired with costs a fraction of the frontier tier. Fable and Opus live in the top-right premium quadrant where you're paying a real premium for the last few points of capability. GPT-5 Mini and Gemini 2.5 Pro cluster at similar cost with lower intelligence — the middle of the chart is crowded with reasonable-but-not-differentiated options. The story: for volume work, pick something on the value frontier; reserve the premium tier for the specific tasks that actually need frontier reasoning.

§ What to try

David’s picks

Curated as of July 2026. What I’d actually reach for, and why.

Opus 4.8AnthropicHardest tasks

My default for the most complex work. When a problem has real depth or needs sustained, careful reasoning, this is what I reach for.

Fable 5AnthropicPlanning

Where I start. I lean on it for early planning and the first passes of development, before the shape of the work is settled.

Sonnet 5AnthropicDay to day

The workhorse for execution and everyday tasks. When something doesn't need a ton of context or extensive thinking, Sonnet handles it fast.

Grok VoicexAIOff the clock

Not a work tool. Claude does the real work; for my non-work, day-to-day questions I mostly just talk to Grok's voice AI.

Sources & methodology
OpenRouter Models API · auth: none
OpenRouter Models — marketing category filter · auth: none
OpenRouter Datasets — rankings-daily · auth: free OpenRouter API key
OpenRouter Marketing rankings (author breakdown) · auth: none (page, not JSON — manually snapshotted)
Epoch AI Notable Models Dataset (CC-BY) · auth: none
Hugging Face Hub API · auth: none
Bank of Canada Valet API (USD/CAD noon rate) · auth: none
Notes:CAD pricing is computed at each refresh using the Bank of Canada USD/CAD noon rate. “Open weights” means the model has a Hugging Face ID in the OpenRouter catalog. Country of origin is manually mapped from the model's org; models from labs we haven't mapped roll up into “Other.”
David Zagury
David's Digital Twin
Online
David Zagury
Hi — I'm David's AI twin. I've read all his writing and know his professional background well. Ask me anything about his work in media or AI.
Powered by Claude · AI can make mistakes