Comparison

Claude vs Gemini API pricing comparison

Claude and Gemini publish different cached-input tiers and batch discounts. This page filters the calculator to both providers so you compare fairly on your actual token mix.

Comparing Anthropic · Google Gemini · same workload baseline · last update: Wednesday 7th October 2026

Calculate my API cost

Related: Anthropic · Google Gemini

Calculator

Compare Anthropic vs Google Gemini for your workload

Step 1

Review your assumptions

These are estimates, not measured usage. Edit the inputs or choose a preset.

Synced Oct 7, 2026
Edit assumptions · 50,000 requests/month · 800 tokens/request · 82% input · 55% of input cached

50K requests / mo

800 tokens avg

Step 2

Lowest estimated cost

Top 5 models for this workload (unreviewed and conditional tariffs excluded)

General intelligence is not a task-specific quality guarantee. Test shortlisted models on your own examples. Source reviewed — exact offer checked on its provider page. Needs review — missing, stale or conflicting evidence; excluded from default recommendations. Estimates cover input/cache-read/output tokens only. Cache writes, storage, tools, search fees, taxes and unconfigured billing conditions are additional. Include billed reasoning tokens in your output estimate.

Head-to-head

Anthropic vs Google Gemini

Cheapest eligible model on each side for the current workload · Google Gemini wins by $48.03/mo

Anthropic

Claude Haiku 4.5 (standard)

$52.56/month

Google Gemini

Lower cost

Gemini 2.5 Flash-Lite (standard)

$4.54/month

Selected model: cost breakdown

Estimate

Claude Opus 5.5 (standard)

Anthropic · focus model for this workload

Estimated monthly cost (USD)

$206.65

Based on: 50K requests · 656 input tokens/request · 144 output tokens/request · 55% of input from cache

Input $62.65 Cached 18.04M tok Output $144.00
Total tokens40.00M
Per 1k requests$4.13
Rank#11 +$202.11 vs #1

Token costs only. Tool calls, search, storage, regional pricing, commitment discounts, and taxes may be excluded.

Explore all priced models

Step 3

Models in this comparison

17 models · Anthropic · Google Gemini

Estimated monthly cost

$206.65

Rank #11 of 17 · $202.11 more than cheapest · Source reviewed

How costs are calculated

Prices from ai-provider-pricing-validated.json, validated Wednesday 7th October 2026. Confirm on official provider pages before billing decisions. Dataset contains 95 offer records, including archived entries and offers that require additional billing inputs or price evidence. Only reviewed offers appear in this calculator. See excluded offers and reasons

Embed this calculator

Embed

Want this calculator on your site?

Copy the iframe snippet below and paste it into any page, doc, or WordPress Custom HTML block.

<iframe src="https://modelcostcomparison.com/embed/ai-api-pricing-calculator?ref=topic-claude-vs-gemini" width="100%" height="980" style="border:0;border-radius:12px;overflow:hidden" loading="lazy" referrerpolicy="strict-origin-when-cross-origin" title="AI API Pricing Calculator by Model Cost Comparison"></iframe>

Model Cost Comparison · Built by Lazige · Methodology

How we calculate cost

Monthly estimate = (input tokens × input $/MTok) + (cached tokens × cached $/MTok when published) + (output tokens × output $/MTok), scaled to your message volume. See the methodology for validation sources and update cadence.

Anthropic vs Google Gemini: the same workload, different token costs

These fixed examples compare tracked, cross-checked models using the snapshot validated 2026-10-07. They are separate from the editable calculator. No cache discounts, batch discounts, taxes or non-token fees are included. Models are grouped by provider, not ranked for quality. Equal token counts normalize the arithmetic; providers may tokenize the same text differently.

  • Short replies: 10,000 requests per month × 1,000 input and 250 output tokens per request, with 0% cached input.
  • Longer prompts: 10,000 requests per month × 8,000 input and 1,000 output tokens per request, with 0% cached input.

Formula: monthly input tokens ÷ 1,000,000 × input rate + monthly output tokens ÷ 1,000,000 × output rate. Output means billable output, including reasoning tokens where applicable.

Monthly token charges in USD; published tier and source notes still apply.
Model / providerShort repliesLonger promptsRate source and conditions
Claude Fable 5 (standard)Anthropic$225.00$1300.00Official provider rates

Standard global Claude API text rates. Cache writes (5 minute or 1 hour), tools, fast mode, regional premiums and long-context conditions require separate accounting.

Claude Fable 5.1 (standard)Anthropic$225.00$1300.00Official provider rates

Standard global Claude API text rates. Cache writes (5 minute or 1 hour), tools, fast mode, regional premiums and long-context conditions require separate accounting.

Claude Haiku 4.5 (standard)Anthropic$22.50$130.00Official provider rates

Standard global Claude API text rates. Cache writes (5 minute or 1 hour), tools, fast mode, regional premiums and long-context conditions require separate accounting.

Claude Opus 4.5 (standard)Anthropic$112.50$650.00Official provider rates

Standard global Claude API text rates. Cache writes (5 minute or 1 hour), tools, fast mode, regional premiums and long-context conditions require separate accounting.

Claude Opus 4.6 (standard)Anthropic$112.50$650.00Official provider rates

Standard global Claude API text rates. Cache writes (5 minute or 1 hour), tools, fast mode, regional premiums and long-context conditions require separate accounting.

Claude Opus 4.7 (standard)Anthropic$112.50$650.00Official provider rates

Standard global Claude API text rates. Cache writes (5 minute or 1 hour), tools, fast mode, regional premiums and long-context conditions require separate accounting.

Claude Opus 5 (standard)Anthropic$112.50$650.00Official provider rates

Standard global Claude API text rates. Cache writes (5 minute or 1 hour), tools, fast mode, regional premiums and long-context conditions require separate accounting.

Claude Opus 5.5 (standard)Anthropic$90.00$520.00Official provider rates

Standard global Claude API text rates. Cache writes (5 minute or 1 hour), tools, fast mode, regional premiums and long-context conditions require separate accounting.

Claude Sonnet 4.5 (standard)Anthropic$67.50$390.00Official provider rates

Standard global Claude API text rates. Cache writes (5 minute or 1 hour), tools, fast mode, regional premiums and long-context conditions require separate accounting.

Claude Sonnet 4.6 (standard)Anthropic$67.50$390.00Official provider rates

Standard global Claude API text rates. Cache writes (5 minute or 1 hour), tools, fast mode, regional premiums and long-context conditions require separate accounting.

Claude Sonnet 5 (standard)Anthropic$45.00$260.00Official provider rates

Standard global Claude API text rates. Cache writes (5 minute or 1 hour), tools, fast mode, regional premiums and long-context conditions require separate accounting.

Gemini 2.5 Flash (standard)Google Gemini$9.25$49.00Official provider rates

Gemini Developer API paid text standard, input at most 1000000 tokens/request. Output includes thinking. Cache storage and grounding/search charges excluded; Pro above 200K uses a higher tariff.

Gemini 2.5 Flash-Lite (standard)Google Gemini$2.00$12.00Official provider rates

Gemini Developer API paid text standard, input at most 1000000 tokens/request. Output includes thinking. Cache storage and grounding/search charges excluded; Pro above 200K uses a higher tariff.

Gemini 2.5 Pro (standard)Google Gemini$37.50$200.00Official provider rates

Gemini Developer API paid text standard, input at most 200000 tokens/request. Output includes thinking. Cache storage and grounding/search charges excluded; Pro above 200K uses a higher tariff.

Gemini 3.8 Flash (standard)Google Gemini$16.88$97.50Official provider rates

Paid standard text rates through 2026-12-31; rates double 2027-01-01. Cache storage ($0.50/M token-hours), grounding and other tools excluded.

Unverified or conflicting snapshot rows and Batch/Flex/Priority tiers are excluded from these examples. This is a comparison of tracked rates, not an exhaustive list of currently available models.

What can change the final bill?

Anthropic: Prompt-cache writes and reads have different billing rules. Tools, long-context tiers and batch processing can change the bill; do not treat a cache-read rate as the cost of creating a cache. Check current billing rules.

Google Gemini: Check Standard versus Batch mode, modality and context tier. Thinking tokens may contribute to output usage; cache storage and grounding/tool charges are additional. Check current billing rules.

Which should you choose?

Use costs to narrow your shortlist, then test both providers on the same representative tasks. Compare answer accuracy, required context and modalities, response time, retries and tool-call success. A lower token price or higher aggregate benchmark score does not prove that a model meets your quality requirements.

Measure total cost per successful task, including repeated calls and billable tools. Replace these illustrative assumptions with measured usage in the calculator before budgeting.

Comparison scenarios

When teams compare Anthropic vs Google Gemini

Illustrative workloads for this head-to-head — same request volume and token mix on both sides, so list prices do not dominate the decision.

“Support traffic often flips the winner: short replies + cache hits can make the cheaper Anthropic or Google Gemini tier beat a premium default.”
Anthropic vs Google Gemini · support botSupport preset · high volume · short replies
“RAG and retrieval apps punish output-heavy pricing. Normalize cache ratio first, then compare Anthropic against Google Gemini on the same corpus size.”
Anthropic vs Google Gemini · RAG / searchRAG preset · retrieval-heavy · high cache
“Code copilots skew long-context input. Use the coding preset so Anthropic vs Google Gemini reflects editor sessions, not chatbot averages.”
Anthropic vs Google Gemini · code assistantCode preset · long context
“Agent loops multiply tool turns. Rank both providers on the agent preset before locking a production model family.”
Anthropic vs Google Gemini · agentsAgent preset · multi-step I/O