No price changes.
What each AI model costs, and what it's actually good at.
Real prices per million tokens, context limits, and concrete use cases for 15 models and 8 coding agents. No benchmark talk.
What changed since the last update.
Auto-generated on every nightly run. Last update September 13, 2026.
None in the last day. Latest: DeepSeek V4 Flash Vision Exp (2026-09-10).
Picks unchanged.
The picks, analyzed nightly.
AI reads every price and spec in the catalog and refreshes these picks on each run. Last analyzed September 13, 2026.
Offers terminal agents and whole-repo refactors at $2 per 1M input tokens.
Has the lowest combined price of $0.075 input and $0.25 output per 1M tokens.
Provides daily coding agent reasoning on a budget for only $0.87 per 1M output tokens.
Delivers complex reasoning and research-heavy work with a 1.05M token context.
Supports 1.05M context tokens for general coding and mixed text at $2 per 1M input.
Features a dedicated free-tier learning use case along with a 1.04M token context.
A model for every job.
15 models with real prices per 1M tokens. Every card shows what to use it for and when to skip it.
Whole-repo refactors · Features with tests · Terminal agents
Bulk batch jobs — output costs $10 per 1M
Hard debugging · Architecture calls · Long agent runs
Everyday edits — 2.5x Sonnet's output price
Complex reasoning · Research-heavy work
Budget work — $50 per 1M output
General coding · Mixed text and code
High-volume loops — Luna is 10x cheaper
Autocomplete · Draft code · Simple scripts
Complex multi-file changes
Large file processing · Fast iteration · Free-tier learning
Top-tier reasoning — use Gemini Pro instead
Chat-style coding help · 500K-context tasks
Code-heavy repos — half the context of 1M rivals
Daily coding agent · Reasoning on a budget · Long files
Only when you need the absolute best output
High-volume edits · Tests and boilerplate · Cheap agents
Tricky logic — weaker than V4 Pro
Agentic workflows · Code review · 1M-context tasks
Cost-sensitive volume — Flash is 4x cheaper
Bulk edits · Renames and cleanup · Learning projects
Complex reasoning
Frontier work with open weights · 1M-context tasks
Budget work — $15 per 1M output
Code generation · Refactors · Agent loops
Very large files — limited to 262K context
General coding · Multilingual work
Cheapest tier — GLM Flash costs a fraction
Open-weight experimentation · Cost-controlled scale
Frontier-quality needs
Prices updated September 13, 2026 · Source: models.dev. "Use for" and "Skip for" are editorial; prices and limits are live.
Compare models side by side.
Pick up to three models. The winner of every row gets a star.
Every model, one table.
The 15 catalog models plus every release from the 9 tracked labs in the last 120 days. Newest first. Verified September 13, 2026.
| Model | Lab | Released | In $/1M | Out $/1M | Context | Card |
|---|---|---|---|---|---|---|
| DeepSeek V4 Flash | DeepSeek | 2026-09-10 | $0.15 | $0.6 | 1M | ★ |
| DeepSeek V4 Flash Vision Exp | DeepSeek | 2026-09-10 | $0.15 | $0.6 | 1M | — |
| DeepSeek V4.1 Flash | DeepSeek | 2026-09-10 | $0.15 | $0.6 | 1M | — |
| GPT-6 Astra | OpenAI | 2026-09-04 | $10 | $50 | 1.05M | ★ |
| Gemini 3.8 Flash | 2026-09-02 | $0.75 | $3.75 | 1.05M | ★ | |
| Muse Spark 1.3 | Meta | 2026-09-02 | $1.25 | $4.25 | 1.05M | ★ |
| Muse Spark 1.3 Contributor | Meta | 2026-09-02 | $0.1 | $0.2 | 1.05M | — |
| Claude Fable 5.1 | Anthropic | 2026-09-01 | $10 | $50 | 1M | — |
| GLM-5.3-Flash | Z.ai | 2026-08-26 | $0.075 | $0.25 | 1M | ★ |
| Qwen3.8 Flash | Alibaba | 2026-08-26 | $0.15 | $0.47 | 1M | — |
| GLM-5.3 | Z.ai | 2026-08-14 | $1.4 | $4.4 | 1M | ★ |
| Gemini 3.7 Flash | 2026-08-13 | $0.75 | $3.75 | 1.05M | — | |
| Gemini Flash Latest | 2026-08-13 | $0.75 | $3.75 | 1.05M | — | |
| Grok 4.6 | xAI | 2026-08-12 | $2 | $6 | 500K | ★ |
| DeepSeek V4 Pro | DeepSeek | 2026-08-12 | $0.435 | $0.87 | 1M | ★ |
| Muse Spark 1.2 | Meta | 2026-08-05 | $1.25 | $4.25 | 1.05M | — |
| Muse Spark 1.2 Contributor | Meta | 2026-08-05 | $0.1 | $0.2 | 1.05M | — |
| Qwen3.8 Max | Alibaba | 2026-08-03 | $2 | $6 | 1M | ★ |
| DeepSeek V4 Flash 0731 | Alibaba | 2026-07-31 | $0.2 | $0.4 | 1M | — |
| Claude Opus 5 | Anthropic | 2026-07-24 | $5 | $25 | 1M | ★ |
| Gemini 3.6 Flash | 2026-07-21 | $0.75 | $3.75 | 1.05M | — | |
| Gemini 3.5 Flash Lite | 2026-07-21 | $0.3 | $2.5 | 1.05M | — | |
| Gemini Flash-Lite Latest | 2026-07-21 | $0.3 | $2.5 | 1.05M | — | |
| Kimi K3 | Moonshot | 2026-07-16 | $3 | $15 | 1.05M | ★ |
| GPT-5.6 Terra | OpenAI | 2026-07-09 | $2 | $12 | 1.05M | ★ |
| GPT-5.6 Luna | OpenAI | 2026-07-09 | $0.2 | $1.2 | 1.05M | ★ |
| Claude Sonnet 5 | Anthropic | 2026-06-29 | $2 | $10 | 1M | ★ |
| Kimi K2.7 Code | Moonshot | 2026-06-12 | $0.95 | $4 | 262K | ★ |
★ marks the models with a full card above. The rest came from the nightly lab scan.
By the numbersA 200x gap between the cheapest output (GLM-5.3-Flash at $0.25) and the priciest (GPT-6 Astra at $50).
The eight agents worth knowing.
The agent is the tool you type into; you pay for the model behind it. Here is what each one costs and runs.
Free, open source
You bring your own API key and own the setup
Included with Claude Pro ($20/mo)
Running Sonnet 5 or Opus 5 in a terminal
Free tier, Pro $20/mo
Beginners — autocomplete plus agent mode in one editor
Free tier, paid from $15/mo
Fast inline completions with a familiar VS Code feel
Free tier with ChatGPT
Delegating tasks that run in your repo or on GitHub
Free tier, generous daily quota
A real agent at $0 — best free starting point
Free, open source
Flat $9.99/mo open-weight models inside VS Code
Free tier, Pro $10/mo
Autocomplete in the editor you already use
Tip: most agents are free or open source. You pay for the model, not the harness — except ClinePass at a flat $9.99.
Dial in your setup. Get your stack.
Slide your budget, set your priorities, and the AI picks the best stack for you.
Pick a theme, change everything.
10 themes restyle the whole site at once. Click one, or press T to flip through them.
Your choice is saved in this browser. Try T right now.
Where the data comes from.
Prices change weekly. These are the pages we check first.