Provider comparison · Last verified 2026-08-16

Claude vs Gemini API Pricing: Same Workload, Real Numbers

Short answer: Gemini is cheaper at every tier we track — 3.6 Flash's promotional rate beats Sonnet 5 by ~62% on a typical document workload, and Flash-Lite is the cheapest model on either roster. Claude's advantages are structural: 128K max output (double Gemini's), a 1M context mid-tier, and a budget model that becomes the cheaper answer when the Gemini promo ends on January 1, 2027.

TL;DR

The verdict in four lines

Mid tier: Gemini wins during the promoGemini 3.6 Flash $0.75/$3.75 vs Claude Sonnet 5 $2.00/$10.00 per 1M tokens — through Dec 31, 2026. After that: $1.50/$7.50, still cheaper, but by less.
Budget tier: Gemini wins today, flips Jan 1, 2027Flash-Lite $0.30/$2.50 beats Haiku 4.5 $1.00/$5.00. But once 3.6 Flash doubles to $1.50/$7.50, Haiku 4.5 becomes cheaper than Google's mid model.
Flagship: no direct Gemini counterpartClaude Opus 5 ($5.00/$25.00) has no comparable Gemini flagship in the lineups tracked here — Google's entries are budget and mid-tier Flash models.
Long outputs: Claude wins structurally128,000 max output tokens on Opus 5/Sonnet 5 vs 65,536 on both Gemini models. Long-form generation favors Claude regardless of price.
Rate card

Head-to-head pricing table

Standard first-party text rates per 1,000,000 tokens. Claude rates verified 2026-08-09 to 2026-08-16; Gemini rates verified 2026-08-16.

TierClaudeGeminiInput verdictOutput verdict
Flagship
Hardest reasoning
Claude Opus 5 — $5.00 / $25.00— no flagship tracked
Mid
Balanced default
Claude Sonnet 5 — $2.00 / $10.00Gemini 3.6 Flash — $0.75 / $3.75PromoGemini −62.5%Gemini −62.5%
Budget
High-volume tasks
Claude Haiku 4.5 — $1.00 / $5.00Gemini 3.5 Flash-Lite — $0.30 / $2.50LowestGemini −70%Gemini −50%

Rates in USD per 1M input/output tokens. Gemini 3.6 Flash's $0.75/$3.75 is promotional through Dec 31, 2026 and becomes $1.50/$7.50 from Jan 1, 2027. Sonnet 5's planned Sep 1, 2026 increase was cancelled. Caching, batch, tools and taxes excluded.

Same workload

Identical document workload, priced monthly

Fixed profile: 60K input + 1K output tokens per request, 1,000 requests per day — 1.8B input and 30M output tokens per 30-day month. No list-price comparisons; this is the same work billed two ways.

ModelInput cost / moOutput cost / mo30-day total
Claude Opus 5
Anthropic · flagship
$9,000.00$750.00$9,750.00
Claude Sonnet 5
Anthropic · mid
$3,600.00$300.00$3,900.00
Claude Haiku 4.5
Anthropic · budget · 200K context
$1,800.00$150.00$1,950.00
Gemini 3.6 Flash
Google · mid (promo rate)
$1,350.00$112.50$1,462.50
Gemini 3.5 Flash-LiteLowest
Google · budget
$540.00$75.00$615.00

Computed from published per-1M rates on the fixed profile above using Gemini 3.6 Flash's promotional rate; from Jan 1, 2027 its row becomes $2,925.00 ($2,700.00 input + $225.00 output). Your tokenization and caching behavior will differ — run your own volumes below.

Live calculator

Your workload on Claude vs Gemini

Pick a preset or enter per-request token volumes and daily requests. All five models from both providers are priced side by side at current (promotional) rates.

Presets:
Cheapest 30-day total
Most expensive 30-day total
ModelPer requestPer day30-day estimate
Standard rates: first-party Anthropic and Google API pricing (Gemini at promotional rates). Excluded: caching, batch, tools and taxes. Pricing data verified —
Break-even

The break-even is a date, not a volume.

Gemini 3.6 Flash is cheaper than Sonnet 5 on both axes, so there is no usage level where Sonnet 5 wins on arithmetic — the gap only scales: $1.25 per 1M input plus $6.25 per 1M output during the promo, worth $2,437.50/month on the workload above. The decision between them is a quality and capability call, not a volume break-even.

The real break-even is January 1, 2027, when 3.6 Flash doubles to $1.50/$7.50. Two things change that day: the Sonnet 5 gap shrinks from $2,437.50 to $975/month on the same workload, and Claude Haiku 4.5 ($1.00/$5.00 → $1,950/month here) becomes cheaper than Google's mid-tier Flash ($2,925/month). If you are signing an annual budget in 2026, price the second half at post-promo rates.

The one structural Claude win that no promo erases: max output. Opus 5 and Sonnet 5 generate up to 128,000 tokens per request; both Gemini models cap at 65,536. Long reports and full-file code generation fit Claude's single-request budget.

Promo gap (now)Gemini 3.6 Flash saves $1.25 per 1M input + $6.25 per 1M output vs Sonnet 5 — $2,437.50/mo on the reference workload.
Post-promo gap (Jan 1, 2027)Shrinks to $975/mo vs Sonnet 5; Haiku 4.5 becomes cheaper than 3.6 Flash for budget traffic.
Output ceiling128K (Claude) vs 65,536 (Gemini) max output tokens — check your longest generations before routing.
Context fit & caveats

Both take 1M tokens — with one 200K exception.

Neither provider surcharges long prompts in current published rates: Sonnet 5 and Opus 5 accept 1,000,000 input tokens at flat prices, and both Gemini models accept 1,048,576. The exception is Haiku 4.5 at 200K tokens — long documents cannot route to the cheapest Claude, while even the cheapest Gemini takes the full 1M.

Tokenizers differ between Anthropic and Google, so identical text can meter several percent apart — measure real documents with the Token Counter, and confirm window fit with the Context Window Checker. Caching and batch discounts exist on both platforms but are not modeled on this page.

Deep dive: ClaudeOpus 5, Sonnet 5, Haiku 4.5 rates and the cancelled price increase: Claude API Pricing.
Deep dive: Gemini3.6 Flash promo window and Flash-Lite budget rates: Gemini API Pricing.
Budget the winnerModel monthly and annual spend with the API Cost Calculator.
FAQ

Claude vs Gemini pricing questions

Is Claude or Gemini cheaper?

Gemini is cheaper at every tier tracked here. Gemini 3.6 Flash costs $0.75/$3.75 per 1M input/output tokens on promotion (through Dec 31, 2026) versus Claude Sonnet 5 at $2.00/$10.00 — about 62% cheaper on a mixed document workload. Gemini 3.5 Flash-Lite ($0.30/$2.50) is the cheapest model on this page, undercutting Claude Haiku 4.5 ($1.00/$5.00) by roughly 68% on the same workload.

What happens to Gemini pricing on January 1, 2027?

Gemini 3.6 Flash's promotional rate ends December 31, 2026 and doubles to $1.50/$7.50 per 1M input/output tokens. It stays cheaper than Claude Sonnet 5 ($2.00/$10.00), but Claude Haiku 4.5 ($1.00/$5.00) then undercuts 3.6 Flash — so the cheapest-answer between the two providers flips on that date for budget workloads.

Claude vs Gemini: which is better for long outputs?

Claude. Claude Opus 5 and Sonnet 5 support up to 128,000 output tokens per request, while Gemini 3.6 Flash and 3.5 Flash-Lite cap at 65,536. Long reports, full-file code generation and book-chapter drafting either fit Claude's single-request budget or require chunked generation on Gemini.

Is there a long-context surcharge on Claude or Gemini?

Neither provider applies a long-context surcharge in their current published rates. Claude Sonnet 5 and Opus 5 accept 1,000,000 input tokens at flat rates; Gemini models accept 1,048,576. The exception is Claude Haiku 4.5, whose context window is 200,000 tokens — long documents must route to Sonnet 5, Opus 5, or a Gemini model.

Which is cheapest for document processing: Claude or Gemini?

On cost alone, Gemini 3.6 Flash. On a workload of 1.8B input and 30M output tokens per month it costs about $1,462.50 on promotion versus $3,900 on Claude Sonnet 5 and $1,950 on Claude Haiku 4.5. If your outputs exceed 65,536 tokens or you need Claude-specific behavior, Sonnet 5 is the practical Claude default — benchmark quality on your own documents before routing.

Methodology

Where these numbers come from.

All rates are taken from the providers' official pricing documentation and re-verified weekly. Workload costs multiply fixed token volumes by published per-million-token rates; they do not estimate tokenization and do not apply caching or batch discounts. "Best fit" on this page is a cost result, not a quality ranking — benchmark both providers on your own prompts.

Official sources: Anthropic pricing and Gemini API pricing. Always confirm the invoice before making purchasing decisions.

All three providersAdd OpenAI to the picture on the LLM API Pricing Comparison page.
OpenAI vs ClaudeHow the two 1M-context incumbents compare: OpenAI vs Claude Pricing.
OpenAI vs GeminiWhere OpenAI's budget model stops being cheapest: OpenAI vs Gemini Pricing.
Words to tokensConvert real content into tokens and price it: How Many Tokens in 1,000 Words?.
What 100K tokens is~75,000 words, 150 pages, and the chunk size that stays under OpenAI's 272K surcharge: How Many Words Is 100K Tokens?.
Context windows explainedInput and output share one budget — window sizes, output caps and what it costs to fill one: What Is a Context Window?.
What a token isDefinition, conversion by content type, and why output costs 5–8x input: What Is an AI Token?
HTML token savingsCleaning scraped markup removes 47–68% of input tokens: HTML Token Savings.
1M tokens in words750,000 words, 1,500 pages — and why only 5 of 8 models accept it in one request: How Many Words Is 1M Tokens?
Trim the promptFind repeated instructions, markup and filler: Prompt Weight Analyzer
How we verifySources, weekly cadence and what our estimates exclude: Methodology · Pricing changelog.