15 models · real prices · updated daily

What each AI model costs, and what it's actually good at.

Real prices per million tokens, context limits, and concrete use cases for 15 models and 8 coding agents. No benchmark talk.

Daily digest

What changed since the last update.

Auto-generated on every nightly run. Last update September 13, 2026.

Price changes

No price changes.

New releases

None in the last day. Latest: DeepSeek V4 Flash Vision Exp (2026-09-10).

Pick changes

Picks unchanged.

Quick answers

The picks, analyzed nightly.

AI reads every price and spec in the catalog and refreshes these picks on each run. Last analyzed September 13, 2026.

Best for coding
Claude Sonnet 5
$2 in · $10 out

Offers terminal agents and whole-repo refactors at $2 per 1M input tokens.

Cheapest
GLM-5.3-Flash
$0.075 in · $0.25 out

Has the lowest combined price of $0.075 input and $0.25 output per 1M tokens.

Best value
DeepSeek V4 Pro
$0.435 in · $0.87 out

Provides daily coding agent reasoning on a budget for only $0.87 per 1M output tokens.

Best for reasoning
GPT-6 Astra
$10 in · $50 out

Delivers complex reasoning and research-heavy work with a 1.05M token context.

Best for long context
GPT-5.6 Terra
$2 in · $12 out

Supports 1.05M context tokens for general coding and mixed text at $2 per 1M input.

Best free option
Gemini 3.8 Flash
$0.75 in · $3.75 out

Features a dedicated free-tier learning use case along with a 1.04M token context.

The catalog

A model for every job.

15 models with real prices per 1M tokens. Every card shows what to use it for and when to skip it.

Claude Sonnet 5
Anthropic · 2026-06-29
best for code
Use for

Whole-repo refactors · Features with tests · Terminal agents

Skip for

Bulk batch jobs — output costs $10 per 1M

in$2out$10cached in$0.2
1M context · 128K max output
Claude Opus 5
Anthropic · 2026-07-24
hard problems
Use for

Hard debugging · Architecture calls · Long agent runs

Skip for

Everyday edits — 2.5x Sonnet's output price

in$5out$25cached in$0.5
1M context · 128K max output
GPT-6 Astra
OpenAI · 2026-09-04
priciest
Use for

Complex reasoning · Research-heavy work

Skip for

Budget work — $50 per 1M output

in$10out$50cached in$1
1.05M context · 128K max output
GPT-5.6 Terra
OpenAI · 2026-07-09
balanced
Use for

General coding · Mixed text and code

Skip for

High-volume loops — Luna is 10x cheaper

in$2out$12cached in$0.2
1.05M context · 128K max output
GPT-5.6 Luna
OpenAI · 2026-07-09
cheap + fast
Use for

Autocomplete · Draft code · Simple scripts

Skip for

Complex multi-file changes

in$0.2out$1.2cached in$0.02
1.05M context · 128K max output
Gemini 3.8 Flash
Google · 2026-09-02
free tier
Use for

Large file processing · Fast iteration · Free-tier learning

Skip for

Top-tier reasoning — use Gemini Pro instead

in$0.75out$3.75cached in$0.075
1.05M context · 66K max output
Grok 4.6
xAI · 2026-08-12
chat-style
Use for

Chat-style coding help · 500K-context tasks

Skip for

Code-heavy repos — half the context of 1M rivals

in$2out$6cached in$0.5
500K context · 500K max output
DeepSeek V4 Pro
DeepSeek · 2026-08-12
best value
Use for

Daily coding agent · Reasoning on a budget · Long files

Skip for

Only when you need the absolute best output

in$0.435out$0.87cached in$0.004
1M context · 384K max output
DeepSeek V4 Flash
DeepSeek · 2026-09-10
cheap 1M context
Use for

High-volume edits · Tests and boilerplate · Cheap agents

Skip for

Tricky logic — weaker than V4 Pro

in$0.15out$0.6cached in$0.003
1M context · 384K max output
GLM-5.3
Z.ai · 2026-08-14
agentic
Use for

Agentic workflows · Code review · 1M-context tasks

Skip for

Cost-sensitive volume — Flash is 4x cheaper

in$1.4out$4.4cached in$0.26
1M context · 131K max output
GLM-5.3-Flash
Z.ai · 2026-08-26
cheapest
Use for

Bulk edits · Renames and cleanup · Learning projects

Skip for

Complex reasoning

in$0.075out$0.25cached in$0.015
1M context · 131K max output
Kimi K3
Moonshot · 2026-07-16
open flagship
Use for

Frontier work with open weights · 1M-context tasks

Skip for

Budget work — $15 per 1M output

in$3out$15cached in$0.3
1.05M context · 131K max output
Kimi K2.7 Code
Moonshot · 2026-06-12
built for code
Use for

Code generation · Refactors · Agent loops

Skip for

Very large files — limited to 262K context

in$0.95out$4cached in$0.19
262K context · 262K max output
Qwen3.8 Max
Alibaba · 2026-08-03
open flagship
Use for

General coding · Multilingual work

Skip for

Cheapest tier — GLM Flash costs a fraction

in$2out$6cached in$0.25
1M context · 131K max output
Muse Spark 1.3
Meta · 2026-09-02
open weights
Use for

Open-weight experimentation · Cost-controlled scale

Skip for

Frontier-quality needs

in$1.25out$4.25cached in$0.15
1.05M context · 131K max output

Prices updated September 13, 2026 · Source: models.dev. "Use for" and "Skip for" are editorial; prices and limits are live.

Head to head

Compare models side by side.

Pick up to three models. The winner of every row gets a star.

The market

Every model, one table.

The 15 catalog models plus every release from the 9 tracked labs in the last 120 days. Newest first. Verified September 13, 2026.

ModelLabReleasedIn $/1MOut $/1MContextCard
DeepSeek V4 FlashDeepSeek2026-09-10$0.15$0.61M
DeepSeek V4 Flash Vision ExpDeepSeek2026-09-10$0.15$0.61M
DeepSeek V4.1 FlashDeepSeek2026-09-10$0.15$0.61M
GPT-6 AstraOpenAI2026-09-04$10$501.05M
Gemini 3.8 FlashGoogle2026-09-02$0.75$3.751.05M
Muse Spark 1.3Meta2026-09-02$1.25$4.251.05M
Muse Spark 1.3 ContributorMeta2026-09-02$0.1$0.21.05M
Claude Fable 5.1Anthropic2026-09-01$10$501M
GLM-5.3-FlashZ.ai2026-08-26$0.075$0.251M
Qwen3.8 FlashAlibaba2026-08-26$0.15$0.471M
GLM-5.3Z.ai2026-08-14$1.4$4.41M
Gemini 3.7 FlashGoogle2026-08-13$0.75$3.751.05M
Gemini Flash LatestGoogle2026-08-13$0.75$3.751.05M
Grok 4.6xAI2026-08-12$2$6500K
DeepSeek V4 ProDeepSeek2026-08-12$0.435$0.871M
Muse Spark 1.2Meta2026-08-05$1.25$4.251.05M
Muse Spark 1.2 ContributorMeta2026-08-05$0.1$0.21.05M
Qwen3.8 MaxAlibaba2026-08-03$2$61M
DeepSeek V4 Flash 0731Alibaba2026-07-31$0.2$0.41M
Claude Opus 5Anthropic2026-07-24$5$251M
Gemini 3.6 FlashGoogle2026-07-21$0.75$3.751.05M
Gemini 3.5 Flash LiteGoogle2026-07-21$0.3$2.51.05M
Gemini Flash-Lite LatestGoogle2026-07-21$0.3$2.51.05M
Kimi K3Moonshot2026-07-16$3$151.05M
GPT-5.6 TerraOpenAI2026-07-09$2$121.05M
GPT-5.6 LunaOpenAI2026-07-09$0.2$1.21.05M
Claude Sonnet 5Anthropic2026-06-29$2$101M
Kimi K2.7 CodeMoonshot2026-06-12$0.95$4262K

★ marks the models with a full card above. The rest came from the nightly lab scan.

By the numbers
15
models tracked
9
labs covered
8
agents compared
$0.325
cheapest 1M in + out (GLM-5.3-Flash)
Output price per 1M tokens
GLM-5.3-Flash
$0.25
DeepSeek V4 Flash
$0.6
DeepSeek V4 Pro
$0.87
GPT-5.6 Luna
$1.2
Gemini 3.8 Flash
$3.75
Kimi K2.7 Code
$4
Muse Spark 1.3
$4.25
GLM-5.3
$4.4
Grok 4.6
$6
Qwen3.8 Max
$6
Claude Sonnet 5
$10
GPT-5.6 Terra
$12
Kimi K3
$15
Claude Opus 5
$25
GPT-6 Astra
$50

A 200x gap between the cheapest output (GLM-5.3-Flash at $0.25) and the priciest (GPT-6 Astra at $50).

The agents

The eight agents worth knowing.

The agent is the tool you type into; you pay for the model behind it. Here is what each one costs and runs.

opencode
Any provider
Terminal
Price

Free, open source

Best for

You bring your own API key and own the setup

opencode.ai ↗
Claude Code
Claude only
Terminal
Price

Included with Claude Pro ($20/mo)

Best for

Running Sonnet 5 or Opus 5 in a terminal

code.claude.com ↗
Cursor
Claude, GPT, Gemini
IDE
Price

Free tier, Pro $20/mo

Best for

Beginners — autocomplete plus agent mode in one editor

cursor.com ↗
Windsurf
Claude, GPT
IDE
Price

Free tier, paid from $15/mo

Best for

Fast inline completions with a familiar VS Code feel

windsurf.com ↗
OpenAI Codex
GPT-5.6, GPT-6
Cloud + CLI
Price

Free tier with ChatGPT

Best for

Delegating tasks that run in your repo or on GitHub

github.com/openai/codex ↗
Cline
Any provider, ClinePass
VS Code
Price

Free, open source

Best for

Flat $9.99/mo open-weight models inside VS Code

cline.bot ↗

Tip: most agents are free or open source. You pay for the model, not the harness — except ClinePass at a flat $9.99.

Recommended for you

Dial in your setup. Get your stack.

Slide your budget, set your priorities, and the AI picks the best stack for you.

What are you building?
Monthly budget $0
$0$50$100+
Priority Balanced
CheapestBalancedBest quality
Your experience

Make it yours

Pick a theme, change everything.

10 themes restyle the whole site at once. Click one, or press T to flip through them.

Your choice is saved in this browser. Try T right now.