AI & Developer Tools

AI Model Cheat Sheet

Every major model's price, context window, and sweet spot — on one page you can bookmark. Last verified August 2026 · printable (Ctrl+P) · updated monthly alongside our price tracker.

⇅ Click a price or size column header to sort (click again to reverse).

ModelProviderInput $/1M Output $/1M Context Max output Best for
Gemini 2.5 Flash-LiteGoogle$0.10$0.401M64KBulk classification, tagging, cheapest possible calls
GPT-5.6 LunaOpenAI$0.20$1.20400K*128K*Cheapest OpenAI model after an 80% price cut
GPT-5.4 NanoOpenAI$0.20$1.25400K*128K*High-volume simple tasks in the OpenAI stack
Gemini 2.5 FlashGoogle$0.30$2.501M64KBudget chat & summarization with big context
Claude Haiku 4.5Anthropic$1.00$5.00200K64KFast production workloads needing Claude quality
Gemini 3.6 FlashGoogle$1.50$7.501M64KCheaper, more token-efficient successor to 3.5 Flash
Gemini 3.7 FlashGoogle$1.50†$7.50†1M64KStronger coding & agent scores than 3.6 Flash at the same list price
Gemini 3.5 FlashGoogle$1.50$9.001M65KSpeed/intelligence balance, agentic browsing
Claude Sonnet 5Anthropic$2.00$10.001M128KCoding & agents near-frontier quality at flagship price
Gemini 3.1 ProGoogle$2.00$12.00500K64KStrong reasoning at mid-tier price
GPT-5.6 TerraOpenAI$2.00$12.00400K*128K*Newest mainstream flagship, now priced with the mid tier
GPT-5.4OpenAI$2.50$15.00400K*128K*Mainstream OpenAI production default
Claude Opus 4.8Anthropic$5.00$25.001M128KLong-horizon agents, hard coding, knowledge work
Claude Opus 5Anthropic$5.00$25.001M128KSame price as Opus 4.8, generational reasoning jump — new top Anthropic pick
GPT-5.5OpenAI$5.00$30.00400K*128K*OpenAI's frontier reasoning
GPT-5.6 SolOpenAI$5.00$30.00400K*128K*Top of the GPT-5.6 family
Claude Fable 5Anthropic$10.00$50.001M128KThe hardest reasoning and longest autonomous runs

Standard-tier list prices, August 2026. †Gemini 3.7 Flash intro rate $0.75/$3.75 through Dec 31, 2026. *GPT-5.x figures per the GPT-5 family's published specs — confirm on OpenAI's model page for the exact variant. All listed models support vision input, prompt caching (~90% off repeated input), and batch (~50% off).

Quick picks — skip the analysis

Cheapest that's still good: Gemini 2.5 Flash-Lite ($0.10/$0.40) for classification and extraction; GPT-5.4 Nano if you're already on OpenAI.
Best value flagship: Claude Sonnet 5 — near-frontier coding/agents at $2/$10, now a permanent price rather than an expiring intro rate.
Biggest context window: Claude's 1M-token models (Fable 5, Opus 4.8, Sonnet 5) and Gemini's 1M Flash tier — roughly 1,500 pages of text in one request.
Longest single output: 128K-output models (Claude 1M-tier, GPT-5.x) — a whole report or codebase-sized diff in one call.
Hardest problems, cost no object: Claude Opus 5 ($5/$25) leads on reasoning benchmarks; Claude Fable 5 ($10/$50) and GPT-5.6 Sol ($5/$30) are the other top picks.
Chatbot at scale: route 80% of traffic to Haiku 4.5 / Gemini Flash, escalate the rest — model it in the chatbot cost simulator.

Reading the spec sheet

Anything on this page can be wrong within a week

Model pricing is not a stable specification. Providers cut prices, launch cheaper variants, deprecate older models and occasionally raise rates, and none of it comes with notice. A cheat sheet is therefore only as trustworthy as the date on it — which is why every figure here is generated from a single source file rather than typed into the page, and why a build check fails the deploy if any two pages on this site quote the same model differently.

What that does and does not buy you:

Use a cheat sheet for what it is good at: narrowing a field of a dozen models to two or three worth pricing properly. Do the final arithmetic against the provider's numbers on the day you need them.

Frequently asked questions

Which single model should I default to?

For most production work in mid-2026: a flagship-tier model (Claude Sonnet 5, GPT-5.6 Terra, or Gemini 3.1 Pro) as the default, a budget model for easy traffic, and a frontier model only where quality visibly pays for itself.

Are benchmark scores on this page?

Deliberately not — leaderboard positions shuffle monthly and rarely predict your task. Price, context, and output limits are the stable facts; test your top 2–3 candidates on your own data.

How do I estimate my cost with these numbers?

Tokens ÷ 1,000,000 × price. Or skip the arithmetic: the AI cost calculator and token calculator do it live.

Can I print or save this?

Yes — Ctrl+P gives a clean printable version (navigation is stripped automatically). Bookmark the page for the monthly-updated version.

The guide that goes deeper

You might also need

Last reviewed: · Who maintains this · How it is checked

Prices are read from each provider's own published pricing page, not from third-party summaries. A check that runs on every build (check-prices.js) fails the deploy if any two pages on this site quote a model differently.