Methodology

How we validate AI API pricing data

Model Cost Comparison is the canonical calculator for workload-based LLM pricing. Rates are last validated Wednesday 7th October 2026. This page documents sources, refresh cadence, and what the calculator does not model.

Data pipeline

  1. Provider pricing is reviewed for an exact API model, billing channel and service tier. A dated evidence record stores the rates and conditions.
  2. OpenRouter is a discovery source for routed offers. It cannot overwrite direct-provider prices, invent cache discounts or verify model identity through fuzzy names.
  3. The refresh reads each official pricing page it knows how to parse. When that reader misses an offer, a LiteLLM row may fill it if the provider, model id, and tier are the same. A matched rate is written into the catalogue and that offer’s check date is renewed for seven days. An offer that still cannot be read keeps its previous rate and date.
  4. Build checks reject missing evidence, price drift and unsupported trust claims. Evidence has a seven-day review window; stale and unreviewed offers are excluded from default recommendations.
  5. Download the dated catalogue to inspect per-offer review status and scope.

What each dataset is for

Workload estimates (input, cached input, and output $/MTok, including Azure, Bedrock, and Vertex host prices) come from the MCC validated dataset. Intelligence Index, output speed, and cost-per-task on the explorer come from Artificial Analysis and are attributed there. When a calculator model is mapped to an AA slug, the explorer shows the same token prices as the calculator. Benchmark fetch dates and price-review dates are independent. The score belongs to the named benchmark variant, including its reasoning effort; it is not an intrinsic guarantee for every API configuration. Missing scores remain unavailable.

Cost formula

Monthly estimate = (non-cached input tokens × input $/MTok) + (cached input tokens × cached $/MTok when published) + (output tokens × output $/MTok), scaled to messages/month × tokens/message. If a model has no cached rate, cached tokens bill at the normal input rate.

Confidence tiers

  • Source reviewed — exact offer checked against a dated primary source, within the stated billing scope.
  • Primary source review does not mean a second source confirmed the price or that every billing dimension is included.
  • Needs review — missing, stale or conflicting evidence; excluded from recommendations.

Why some offers are excluded

The September 23 review covered all 59 previously excluded offers: 26 restored, 27 archived as retired, deprecated or duplicate entries, 5 requiring other billing inputs, and 1 with unresolved primary pricing evidence. Archived rates are historical records, not current purchase options.

See each excluded offer and its source
  • GPT-6 Sol Pro — Archived. Duplicate execution-mode label: GPT-6 and GPT-5.6 pro is reasoning.mode on the base model. Compare the base offer using all billed work tokens. This is not a separately verified API model ID. Primary source
  • GPT-6 Luna Pro — Archived. Duplicate execution-mode label: GPT-6 and GPT-5.6 pro is reasoning.mode on the base model. Compare the base offer using all billed work tokens. This is not a separately verified API model ID. Primary source
  • GPT-6 Astra Pro — Archived. Duplicate execution-mode label: GPT-6 and GPT-5.6 pro is reasoning.mode on the base model. Compare the base offer using all billed work tokens. This is not a separately verified API model ID. Primary source
  • GPT-5.6 Sol Pro — Archived. Duplicate execution-mode label: GPT-6 and GPT-5.6 pro is reasoning.mode on the base model. Compare the base offer using all billed work tokens. This is not a separately verified API model ID. Primary source
  • GPT-5.6 Terra Pro — Archived. Duplicate execution-mode label: GPT-6 and GPT-5.6 pro is reasoning.mode on the base model. Compare the base offer using all billed work tokens. This is not a separately verified API model ID. Primary source
  • GPT-5.6 Luna Pro — Archived. Duplicate execution-mode label: GPT-6 and GPT-5.6 pro is reasoning.mode on the base model. Compare the base offer using all billed work tokens. This is not a separately verified API model ID. Primary source
  • Claude Opus 4.1 — Archived. Direct Claude API retired 2026-08-05. Partner-hosted offers have separate lifecycles. Primary source
  • Claude Opus 4 — Archived. Direct Claude API retired 2026-06-15. Partner-hosted offers have separate lifecycles. Primary source
  • Claude Sonnet 4 — Archived. Direct Claude API retired 2026-06-15. Partner-hosted offers have separate lifecycles. Primary source
  • Claude Sonnet 3.7 — Archived. Direct Claude API retired 2026-02-19. Partner-hosted offers have separate lifecycles. Primary source
  • Claude Haiku 3.5 — Archived. Direct Claude API retired 2026-02-19. Partner-hosted offers have separate lifecycles. Primary source
  • Gemini 2.0 Flash (standard) — Archived. Gemini Developer API model shut down 2026-06-01, including standard and batch offers. Primary source
  • Gemini 2.0 Flash (batch) — Archived. Gemini Developer API model shut down 2026-06-01, including standard and batch offers. Primary source
  • Gemini 2.0 Flash-Lite (standard) — Archived. Gemini Developer API model shut down 2026-06-01, including standard and batch offers. Primary source
  • Gemini 2.0 Flash-Lite (batch) — Archived. Gemini Developer API model shut down 2026-06-01, including standard and batch offers. Primary source
  • Mistral Medium 3.1 — Archived. Listed by Mistral under deprecated and retired models; excluded from current recommendations. This does not assert that every hosted deployment is shut down. Primary source
  • Mistral Large 2.1 — Archived. Listed by Mistral under deprecated and retired models; excluded from current recommendations. This does not assert that every hosted deployment is shut down. Primary source
  • Mistral Small 3.2 — Archived. Listed by Mistral under deprecated and retired models; excluded from current recommendations. This does not assert that every hosted deployment is shut down. Primary source
  • Mistral Small Creative — Archived. Listed by Mistral under deprecated and retired models; excluded from current recommendations. This does not assert that every hosted deployment is shut down. Primary source
  • Magistral Medium 1.2 — Archived. Listed by Mistral under deprecated and retired models; excluded from current recommendations. This does not assert that every hosted deployment is shut down. Primary source
  • Command A+ — Additional billing inputs required. Official access is free only within rate limits; production deployment uses Model Vault. The legacy routed token tariff is not a verified uncapped direct API offer. Primary source
  • Command A — Unresolved price. Model identity is documented, but the current primary pricing page does not establish this offer’s paid rates. Historical or routed rates are not sufficient to restore it. Primary source
  • Grok 4.20 — Archived. Ambiguous generic duplicate: use the separately reviewed, dated Grok 4.20 reasoning, non-reasoning or multi-agent API offer. No separate generic tariff has been established. Primary source
  • Grok 4.1 Fast Reasoning — Archived. Retired 2026-05-15; alias redirects to Grok 4.3 with different behavior and prices. Use the current Grok 4.3 offer. Primary source
  • Grok 4.1 Fast Non-Reasoning — Archived. Retired 2026-05-15; alias redirects to Grok 4.3 with different behavior and prices. Use the current Grok 4.3 offer. Primary source
  • Sonar — Additional billing inputs required. Token rates verified, but mandatory request fees depend on search context: $5/$8/$12 per 1K requests for Sonar, $6/$10/$14 for Sonar Pro and Reasoning Pro. Pro Search has separate higher fees. Excluded until the UI accounts for the selected search mode. Primary source
  • Sonar Pro — Additional billing inputs required. Token rates verified, but mandatory request fees depend on search context: $5/$8/$12 per 1K requests for Sonar, $6/$10/$14 for Sonar Pro and Reasoning Pro. Pro Search has separate higher fees. Excluded until the UI accounts for the selected search mode. Primary source
  • Sonar Reasoning Pro — Additional billing inputs required. Token rates verified, but mandatory request fees depend on search context: $5/$8/$12 per 1K requests for Sonar, $6/$10/$14 for Sonar Pro and Reasoning Pro. Pro Search has separate higher fees. Excluded until the UI accounts for the selected search mode. Primary source
  • Sonar Deep Research — Additional billing inputs required. Requires separate citation ($2/M), reasoning ($3/M) and search-query ($5/1K) usage in addition to $2/M input and $8/M output. Excluded from token-only comparisons until those billing dimensions are supported. Primary source
  • DeepSeek V4 Flash — Archived. Superseded duplicate catalogue row. Use the reviewed current DeepSeek V4.1 Flash or V4 Pro 0813 offers, with explicit peak/off-peak tier; legacy routed rates are not direct tariffs. Primary source
  • DeepSeek V4 Pro — Archived. Superseded duplicate catalogue row. Use the reviewed current DeepSeek V4.1 Flash or V4 Pro 0813 offers, with explicit peak/off-peak tier; legacy routed rates are not direct tariffs. Primary source
  • deepseek-chat — Archived. Legacy API name retired 2026-07-24. Its old price cannot describe the current DeepSeek API. Primary source
  • deepseek-reasoner — Archived. Legacy API name retired 2026-07-24. Its old price cannot describe the current DeepSeek API. Primary source

Not included in estimates

The UI estimates token charges. Cache writes and storage, tool-call fees, web search grounding, vector storage, image tokenization, regional multipliers, taxes, and enterprise discounts are not modeled. Add those from your provider console when they apply.

Canonical tool vs article

The interactive calculator lives at modelcostcomparison.com. A long-form article with additional context remains on lazige.agency and points here as the canonical tool URL.

How to use the calculator →