Gemini API pricing: every model, every rate

Gemini's headline input price looks cheap, but output is billed at a multiple of it. The table below shows both numbers, because the output rate is what dominates real bills.

In short: Google bills 19 text models here, from $0.1875 to $27 per 1M tokens blended (input + output). no Google list price could be verified for any of them, so every figure here is a gateway rate and no discount is claimed. A working workload — 50M input and 10M output tokens per month — costs about $37.50 on gemini-3.7-flash. Rates checked 2026-09-16.
Check current rates → Free to sign up · $1 minimum top-up · No prepayment

Full rate card

80 billable models in total, 19 of them text models. On those text models: 0 carry a rate verified against the vendor's own published pricing, the rest are gateway rates only. USD per 1M tokens, checked 2026-09-16.

ModelInputOutputOut/InOfficial (in / out)
gemini-2.5-flash-deepsearch-async$3$248.0×not verified
gemini-3-pro-image-preview-l$1$6060.0×not verified
gemini-3-pro-preview$1$66.0×not verified
gemini-3-pro-preview-11-2025$1$66.0×not verified
gemini-3-pro-preview-11-2025-thinking$1$66.0×not verified
gemini-3-pro-preview-thinking$1$66.0×not verified
gemini-3.1-pro-preview$1$66.0×not verified
gemini-3.1-pro-preview-customtools$1$66.0×not verified
gemma-2b-it$1$11.0×not verified
gemma-7b-it$1$11.0×not verified
gemini-3.5-flash$0.75$4.56.0×not verified
gemma2-27b-it$0.63$0.631.0×not verified
gemini-1.5-pro$0.625$2.54.0×not verified
gemini-1.5-pro-001$0.625$2.54.0×not verified
gemini-1.5-pro-002$0.625$2.54.0×not verified
gemini-2.5-pro$0.625$58.0×not verified
gemini-2.5-pro-nothinking$0.625$58.0×not verified
gemini-2.5-pro-preview-03-25$0.625$58.0×not verified
gemini-2.5-pro-preview-03-25-nothinking$0.625$58.0×not verified
gemini-2.5-pro-preview-05-06$0.625$58.0×not verified
gemini-2.5-pro-preview-05-06-nothinking$0.625$58.0×not verified
gemini-2.5-pro-preview-05-06-thinking$0.625$58.0×not verified
gemini-2.5-pro-preview-06-05$0.625$58.0×not verified
gemini-2.5-pro-preview-06-05-thinking$0.625$58.0×not verified
gemini-2.5-pro-thinking$0.625$58.0×not verified
gemini-2.5-pro-thinking-*$0.625$58.0×not verified
gemini-2.5-pro-preview-03-25-thinking$0.526$4.2088.0×not verified
gemini-2.5-pro-preview-tts$0.5$1020.0×not verified
gemini-3.6-flash$0.375$1.8755.0×not verified
gemini-3.7-flash$0.375$1.8755.0×not verified
gemini-3.8-flash$0.375$1.8755.0×not verified
gemini-2.5-flash-preview-tts$0.25$520.0×not verified
gemini-3-flash-preview$0.25$1.56.0×not verified
gemini-3-flash-preview-thinking$0.25$1.56.0×not verified
gemini-3.1-flash-preview$0.25$1.56.0×not verified
gemini-2.5-flash$0.15$1.2518.3×not verified
gemini-2.5-flash-nothinking$0.15$1.2518.3×not verified
gemini-2.5-flash-preview-09-2025$0.15$1.25258.3×not verified
gemini-2.5-flash-preview-09-2025-nothinking$0.15$1.25258.3×not verified
gemini-2.5-flash-preview-09-2025-thinking$0.15$1.25258.3×not verified
gemini-2.5-flash-preview-09-2025-thinking-*$0.15$1.25258.3×not verified
gemini-2.5-flash-thinking$0.15$1.2518.3×not verified
gemini-2.5-flash-thinking-*$0.15$1.2518.3×not verified
gemini-3.5-flash-lite$0.15$1.258.3×not verified
gemini-3.1-flash-lite$0.125$0.756.0×not verified
gemini-3.1-flash-lite-preview$0.125$0.756.0×not verified
gemini-embedding-2-preview$0.1$0.44.0×not verified
gemini-2.0-flash-001$0.075$0.34.0×not verified
gemini-2.5-flash-preview-04-17$0.075$0.34.0×not verified
gemini-2.5-flash-preview-04-17-nothinking$0.075$0.34.0×not verified
gemini-2.5-flash-preview-04-17-thinking$0.075$0.34.0×not verified
gemini-2.5-flash-preview-05-20$0.075$0.34.0×not verified
gemini-2.5-flash-preview-05-20-nothinking$0.075$0.34.0×not verified
gemini-2.5-flash-preview-05-20-thinking$0.075$0.34.0×not verified
gemini-embedding-001$0.075$0.34.0×not verified
gemma-3-27b-it$0.06$0.183.0×not verified
gemini-2.0-flash$0.05$0.24.0×not verified
gemini-2.5-flash-lite$0.05$0.24.0×not verified
gemini-2.5-flash-lite-nothinking$0.05$0.24.0×not verified
gemini-2.5-flash-lite-preview-06-17$0.05$0.24.0×not verified
gemini-2.5-flash-lite-preview-06-17-nothinking$0.05$0.24.0×not verified
gemini-2.5-flash-lite-preview-06-17-thinking$0.05$0.24.0×not verified
gemini-2.5-flash-lite-preview-09-2025$0.05$0.24.0×not verified
gemini-2.5-flash-lite-preview-09-2025-nothinking$0.05$0.24.0×not verified
gemini-2.5-flash-lite-preview-09-2025-thinking$0.05$0.24.0×not verified
gemini-2.5-flash-lite-preview-09-2025-thinking-*$0.05$0.24.0×not verified
gemini-2.5-flash-lite-thinking$0.05$0.24.0×not verified
gemma-2-27b-it$0.05$0.153.0×not verified
gemini-1.5-flash$0.0375$0.154.0×not verified
gemini-1.5-flash-002$0.0375$0.154.0×not verified
gemini-2.0-flash-lite$0.0375$0.154.0×not verified
gemini-2.0-flash-lite-001$0.0375$0.154.0×not verified
gemini-2.0-flash-lite-preview-02-05$0.0375$0.154.0×not verified
gemma-3-12b-it$0.035$0.1053.0×not verified
gemma-2-9b-it$0.025$0.0753.0×not verified
gemini-1.5-flash-8b$0.01875$0.0754.0×not verified
gemma-3-4b-it$0.0175$0.05253.0×not verified
gemma-2-2b-it$0.0125$0.03753.0×not verified
gemma-3-1b-it$0.01$0.033.0×not verified
gemma2-9b-it$0.01$0.011.0×not verified

Output is billed at a multiple of input on most models. The Out/In column is that multiple — it matters more than the input price once output dominates your bill.

Across 19 priced text models the blended rate spans $0.1875 (gemini-2.0-flash-lite) to $27 (gemini-2.5-flash-deepsearch-async) per 1M tokens — a 144× spread. Picking the wrong tier is usually the single most expensive mistake here.

What Google costs at three usage levels

Priced on gemini-3.7-flash at $0.375 in / $1.875 out per 1M tokens, with a 5:1 input-to-output ratio — roughly what an interactive workload produces.

Usage levelTokens per month (in / out)Monthly cost
gemini-3.7-flash
Light
pilot or side project
5M / 1M$3.75
Working
one product in production
50M / 10M$37.50
Heavy
high-volume pipeline
500M / 100M$375.00

This is the model we cover in depth for this vendor. Swap in your own token split in the calculator.

Same budget, other vendors

3 models from other vendors priced within 35% of gemini-3.7-flash on a blended rate — the cheapest, the dearest and one in between. This is the horizontal check a single-vendor pricing page cannot give you.

ModelVendorIn / out per 1Mvs this
grok-build-0.1xAI$0.5 / $1−33.3%
qwen3.5-397b-a17bAlibaba$0.3 / $1.8−6.7%
claude-haiku-4-5-20251001Anthropic$0.5 / $2.5+33.3%

Blended comparison (input + output per 1M tokens). Price is one axis — these are the models worth testing side by side at this budget, not interchangeable substitutes.

Why the output rate decides the bill

Across the 19 priced Google text models, the output rate runs from 4.0× to 8.3× the input rate, with a median of 6.0×. Output tokens are what a chat or agent turn produces — on a typical workload they are the smaller half of the token count but the larger half of the bill.

Worked out: at a 6.0× multiple, a job whose token count is 80% input still spends 60% of its cost on the output it generates. Comparing vendors on the input rate alone hides that.

What Google is good at

Gemini wins on long-context and multimodal work. Flash tiers are among the cheapest capable models available anywhere.

Before you compare: Gemini output rates are a multiple of input. Compare on your own in/out split, not on the input price alone.

Detailed pages for Google models

The models we cover in depth, each with worked cost at three usage levels and a comparison against similarly priced alternatives.

Other vendors

Every rate card here is built the same way, from the same price snapshot (2026-09-16), so the numbers are comparable across vendors:

Anthropic · OpenAI · DeepSeek · Alibaba · xAI · Zhipu · Moonshot · Meta · MiniMax · ByteDance

FAQ

How much does Google charge per million tokens?

Across the 19 priced text models in this catalogue, the blended rate (input plus output per 1M tokens) runs from $0.1875 on gemini-2.0-flash-lite to $27 on gemini-2.5-flash-deepsearch-async, as of 2026-09-16. The spread is what matters: choosing the wrong tier inside one vendor usually costs more than switching vendor.

What is the cheapest Google model for high-volume work?

gemini-2.0-flash-lite is the lowest blended rate here at $0.0375 input and $0.15 output per 1M tokens. On a Heavy workload — 500M input and 100M output tokens a month — that bills about $33.75. At the other end, the same workload on gemini-2.5-flash-deepsearch-async costs about $3,900.00.

Is Google cheaper through a gateway than buying direct?

We do not claim a discount on Google. Its list pricing is not published in a stable machine-readable form, so every figure on this page is a gateway rate with no verified official price to compare against. Check the number in your dashboard before committing to a budget.

How does output pricing change a Google bill?

Output is billed at a multiple of input on every model priced here, so the input rate alone never predicts the bill. A workload that is mostly input tokens still lands most of its cost on the output column once that multiple is applied. Price your own in/out split in the calculator rather than extrapolating from the headline input rate.

Gemini's input price looks cheap — is the bill actually cheap?

Not necessarily. Across the priced Gemini text models the output rate is a multiple of the input rate, and image-capable tiers push that multiple into the double digits. A workload that generates a lot of output will be dominated by the output column, not the headline input number.

Create free account

See all 827 models → or use the cost calculator with your own token split.