AI Model Pricing Cheat Sheet

Every major AI model's API price in one place — input and output cost per million tokens, context window, and what each is best for. Ranked cheapest first. Prices are verified from each provider's official pricing page. Tap a column to re-sort.

14 models · 7 makers · list price, flagship models
ModelMakerInput$/1MOutput$/1MBlended$/1M *ContextBest for
Grok 4 FastGroksrc ✓$0.2$0.5$0.28128K tokensCheaper, faster tier for everyday queries and high-volume use.
DeepSeek-V3.2DeepSeeksrc ✓$0.28$0.42$0.32128K tokensNear-frontier chat and coding quality at the lowest API price of any major provider.
Mistral Medium 3Le Chat (Mistral)src ✓$0.4$2$0.8128K tokensAggressively priced mid-tier that punches above its weight for everyday work.
Gemini 2.5 FlashGeminisrc ✓$0.3$2.5$0.851M tokensVery fast and cheap for everyday and high-volume workloads.
SonarPerplexitysrc ✓$1$1$1Varies by modelSearch-grounded answers with citations at a very low price — built for live-web Q&A.
Claude Haiku 4.5Claudesrc ✓$1$5$2200K tokensFast and cheap for high-volume tasks while staying genuinely capable.
Mistral LargeLe Chat (Mistral)src ✓$2$6$3128K tokensEurope's strongest frontier model — solid reasoning and code with EU data residency.
Claude Sonnet 5Claudesrc ✓$2$10$4200K tokensNew (Jun 2026) agent-focused workhorse — near-flagship quality at a steep discount.Introductory pricing through Aug 31, 2026; then $3 in / $15 out.
Gemini 3 ProGeminisrc ✓$2$12$4.51M tokensFrontier reasoning with a 1M-token context — reads entire codebases or document sets in one pass.
GPT-5.4ChatGPTsrc ✓$2.5$15$5.63128K tokensStrong general-purpose frontier model at half the flagship price.
Sonar ProPerplexitysrc ✓$3$15$6Varies by modelDeeper multi-step research answers with larger context and more citations.
Grok 4Groksrc ✓$3$15$6128K tokensStrong frontier reasoning with native access to real-time X (Twitter) data.
Claude Opus 4.8Claudesrc ✓$5$25$10200K tokensThe frontier flagship (May 2026) — top-tier deep reasoning and coding for long, complex, multi-step work; no long-context surcharge.
GPT-5.5ChatGPTsrc ✓$5$30$11.25128K tokensCurrent flagship — hard reasoning, coding, and agentic tasks. (GPT-5.6 is in limited preview.)

* Blended assumes a 3:1 input:output token mix (typical for chat/RAG) so you can rank models with one number — your real ratio will differ.

These are headline list prices for each maker's flagship models. Theyexclude prompt caching (often up to ~90% off), batch-API discounts (~50%), volume/committed-use tiers, long-context surcharges, and image/audio token costs — so a real bill can differ. Always confirm on the provider's official pricing page (the link on each row) before budgeting. Prices are human-verified, never machine-written, consistent with how we rank & disclose.

Compare full tool features →Which models are actually shipping? →