OpenAI API pricing: every model, every rate

OpenAI's catalogue has grown faster than its pricing page. Below is every model currently billable, with the rate you actually pay per million tokens.

In short: OpenAI bills 80 text models here, from $0.225 to $105 per 1M tokens blended (input + output). 4 of them carry a rate verified against OpenAI's own published pricing; the rest are gateway rates only. A working workload — 50M input and 10M output tokens per month — costs about $110.00 on gpt-5.6-terra. Rates checked 2026-09-16.
Check current rates → Free to sign up · $1 minimum top-up · No prepayment

Full rate card

150 billable models in total, 80 of them text models. On those text models: 4 carry a rate verified against the vendor's own published pricing, the rest are gateway rates only. USD per 1M tokens, checked 2026-09-16.

ModelInputOutputOut/InOfficial (in / out)
gpt-4.5-preview$37.5$752.0×not verified
gpt-4-32k$30not verified
gpt-4-32k-0613$30not verified
gpt-4$15not verified
gpt-4-0613$15not verified
gpt-4-gizmo-*$15$755.0×not verified
gpt-5.4-pro$15$906.0×not verified
gpt-5.5-pro$15$906.0×not verified
whisper-1$15not verified
gpt-5.2-pro$10.5$848.0×not verified
o3-pro$10$404.0×not verified
gpt-5-pro$7.5$608.0×not verified
o1$7.5not verified
o1-preview$7.5$304.0×not verified
gpt-4-0125-preview$5not verified
gpt-4-1106-preview$5$153.0×not verified
gpt-4-1106-vision-preview$5not verified
gpt-4-turbo$5not verified
gpt-4-turbo-preview$5not verified
gpt-4-vision-preview$5not verified
gpt-6$5$255.0×not verified
gpt-6-astra$5$255.0×$10 / $50 -50.0%
o3-deep-research$5$204.0×not verified
GPT-4o-TL$2.5$7.53.0×not verified
gpt-4o-alle$2.5$7.53.0×not verified
gpt-4o-realtime-preview$2.5$104.0×not verified
gpt-5.5$2.5$156.0×not verified
gpt-5.5-high$2.5$156.0×not verified
gpt-5.5-low$2.5$156.0×not verified
gpt-5.5-medium$2.5$156.0×not verified
gpt-5.5-openai-compact$2.5$156.0×not verified
gpt-5.5-xhigh$2.5$156.0×not verified
gpt-5.6-sol$2.5$156.0×$4 / $20 -37.5%
gpt-5.6-sol-high$2.5$156.0×not verified
gpt-5.6-sol-low$2.5$156.0×not verified
gpt-5.6-sol-max$2.5$156.0×not verified
gpt-5.6-sol-medium$2.5$156.0×not verified
gpt-5.6-sol-ultra$2.5$156.0×not verified
gpt-5.6-sol-xhigh$2.5$156.0×not verified
gpt-image-1$2.5$208.0×not verified
gpt-image-1.5$2.5$166.4×not verified
gpt-image-2$2.5$156.0×not verified
gpt-image-2.5-flare$2.5$156.0×not verified
gpt-image-2.5-sunburst$2.5$156.0×not verified
net-gpt-4$2.25$94.0×not verified
gpt-realtime-1.5-2026-02-23$2$84.0×not verified
gpt-realtime-2.1$2$126.0×not verified
gpt-realtime-2025-08-28$2$84.0×not verified
gpt-3.5-turbo-16k$1.5not verified
gpt-3.5-turbo-16k-0613$1.5not verified
gpt-3.5-turbo-instruct$1.5not verified
gpt-4o$1.25$54.0×not verified
gpt-4o-audio-preview$1.25$54.0×not verified
gpt-4o-search-preview$1.25$54.0×not verified
gpt-4o-transcribe$1.25$54.0×not verified
gpt-5.4$1.25$7.56.0×not verified
gpt-5.4-high$1.25$7.56.0×not verified
gpt-5.4-low$1.25$7.56.0×not verified
gpt-5.4-medium$1.25$7.56.0×not verified
gpt-5.4-openai-compact$1.25$7.56.0×not verified
gpt-5.4-xhigh$1.25$7.56.0×not verified
gpt-audio-2025-08-28$1.25not verified
gpt-3$1$11.0×not verified
gpt-4.1$1$44.0×not verified
gpt-5.6-terra$1$66.0×$2 / $12 -50.0%
gpt-5.6-terra-high$1$66.0×not verified
gpt-5.6-terra-low$1$66.0×not verified
gpt-5.6-terra-max$1$66.0×not verified
gpt-5.6-terra-medium$1$66.0×not verified
gpt-5.6-terra-ultra$1$66.0×not verified
gpt-5.6-terra-xhigh$1$66.0×not verified
gpt-image-1-mini$1$44.0×not verified
o3$1$44.0×not verified
o4-mini-deep-research$1$44.0×not verified
gpt-5.2$0.875$78.0×not verified
gpt-5.2-chat$0.875$78.0×not verified
gpt-5.2-codex$0.875$78.0×not verified
gpt-5.3-chat$0.875$78.0×not verified
gpt-5.3-codex$0.875$78.0×not verified
gpt-5.3-codex-high$0.875$78.0×not verified
gpt-5.3-codex-low$0.875$78.0×not verified
gpt-5.3-codex-medium$0.875$78.0×not verified
gpt-5.3-codex-spark$0.875$78.0×not verified
gpt-5.3-codex-xhigh$0.875$78.0×not verified
360gpt-pro$0.8572not verified
360gpt-turbo-responsibility-8k$0.8572not verified
gpt-3.5-turbo-0301$0.75$2.253.0×not verified
gpt-3.5-turbo-0613$0.75not verified
gpt-4o-mini-transcribe$0.75$34.0×not verified
gpt-5$0.625$58.0×not verified
gpt-5-codex$0.625$58.0×not verified
gpt-5-codex-high$0.625$58.0×not verified
gpt-5-codex-low$0.625$58.0×not verified
gpt-5-codex-medium$0.625$58.0×not verified
gpt-5-high$0.625$1.252.0×not verified
gpt-5-low$0.625$1.252.0×not verified
gpt-5-medium$0.625$1.252.0×not verified
gpt-5-minimal$0.625$1.252.0×not verified
gpt-5-search-api$0.625$58.0×not verified
gpt-5.1$0.625$58.0×not verified
gpt-5.1-chat$0.625$58.0×not verified
gpt-5.1-codex$0.625$58.0×not verified
gpt-5.1-codex-high$0.625not verified
gpt-5.1-codex-low$0.625not verified
gpt-5.1-codex-max$0.625$58.0×not verified
gpt-5.1-codex-medium$0.625not verified
gpt-5.1-high$0.625not verified
gpt-5.1-low$0.625not verified
gpt-5.1-medium$0.625not verified
gpt-5.1-thinking$0.625$58.0×not verified
gpt-oss-120b$0.55$2.24.0×not verified
o1-mini$0.55$2.24.0×not verified
o3-mini$0.55$2.24.0×not verified
o4-mini$0.55$2.24.0×not verified
gemini-3.1-flash-tts-preview$0.5$1020.0×not verified
gpt-3.5-turbo-1106$0.5not verified
qwen3-tts-flash$0.4$0not verified
gpt-5.4-mini$0.375$2.256.0×not verified
gpt-4o-mini-realtime-preview$0.3$1.24.0×not verified
gpt-4o-mini-tts$0.3$620.0×not verified
gpt-realtime-2.1-mini$0.3$1.24.0×not verified
gpt-3.5-turbo$0.25$0.753.0×not verified
gpt-3.5-turbo-0125$0.25$0.753.0×not verified
net-gpt-3.5-turbo$0.25not verified
tts-1$0.25$7.530.0×not verified
tts-1-1106$0.25$7.530.0×not verified
tts-1-hd$0.25$1560.0×not verified
tts-1-hd-1106$0.25$1560.0×not verified
gpt-4.1-mini$0.2$0.84.0×not verified
gpt-5-mini$0.125$18.0×not verified
gpt-5.1-codex-mini$0.125$0.21.6×not verified
gpt-5.4-nano$0.1$0.6256.2×not verified
gpt-5.6-luna$0.1$0.66.0×$0.2 / $1.2 -50.0%
gpt-5.6-luna-high$0.1$0.66.0×not verified
gpt-5.6-luna-low$0.1$0.66.0×not verified
gpt-5.6-luna-max$0.1$0.66.0×not verified
gpt-5.6-luna-medium$0.1$0.66.0×not verified
gpt-5.6-luna-ultra$0.1$0.66.0×not verified
gpt-5.6-luna-xhigh$0.1$0.66.0×not verified
gpt-oss-20b$0.1$0.44.0×not verified
360gpt-turbo$0.0858not verified
gpt-4o-mini$0.075$0.34.0×not verified
gpt-4o-mini-audio-preview$0.075$0.34.0×not verified
gpt-4o-mini-search-preview$0.075$0.34.0×not verified
text-embedding-3-large$0.065not verified
gpt-4.1-nano$0.05$0.24.0×not verified
text-embedding-ada-002$0.05not verified
gpt-5-nano$0.025$0.28.0×not verified
text-embedding-3-small$0.01not verified
text-embedding-v1$0.0075not verified

Output is billed at a multiple of input on most models. The Out/In column is that multiple — it matters more than the input price once output dominates your bill.

Across 80 priced text models the blended rate spans $0.225 (gpt-5-nano) to $105 (gpt-5.5-pro) per 1M tokens — a 467× spread. Picking the wrong tier is usually the single most expensive mistake here.

What OpenAI costs at three usage levels

Priced on gpt-5.6-luna at $0.1 in / $0.6 out; gpt-5.6-terra at $1 in / $6 out per 1M tokens, with a 5:1 input-to-output ratio — roughly what an interactive workload produces.

Usage levelTokens per month (in / out)Lowest-cost tier
gpt-5.6-luna
Median-rate model
gpt-5.6-terra
Light
pilot or side project
5M / 1M$1.10$11.00
Working
one product in production
50M / 10M$11.00$110.00
Heavy
high-volume pipeline
500M / 100M$110.00$1,100.00

The first column is this vendor's lowest-cost tier; the second is the model closest to its median rate, which is nearer what a real deployment bills. Swap in your own token split in the calculator.

Same budget, other vendors

3 models from other vendors priced within 35% of gpt-5.6-terra on a blended rate — the cheapest, the dearest and one in between. This is the horizontal check a single-vendor pricing page cannot give you.

ModelVendorIn / out per 1Mvs this
deepseek-coderDeepSeek$1 / $4−28.6%
doubao-1.5-pro-256kByteDance$2.5 / $4.5same
kimi-k3Moonshot$1.5 / $7.5+28.6%

Blended comparison (input + output per 1M tokens). Price is one axis — these are the models worth testing side by side at this budget, not interchangeable substitutes.

Why the output rate decides the bill

Across the 73 priced OpenAI text models, the output rate runs from 1.6× to 8.0× the input rate, with a median of 6.0×. Output tokens are what a chat or agent turn produces — on a typical workload they are the smaller half of the token count but the larger half of the bill.

Worked out: at a 6.0× multiple, a job whose token count is 80% input still spends 60% of its cost on the output it generates. Comparing vendors on the input rate alone hides that.

What OpenAI is good at

GPT-5.6 Luna is the volume workhorse — cheap enough for classification and routing, strong enough for most user-facing chat.

Before you compare: GPT-5.6 Sol is currently discounted by OpenAI itself (a promotional rate). Gateway Sol pricing is lower still, but by a smaller margin than the rest of the family.

Detailed pages for OpenAI models

The models we cover in depth, each with worked cost at three usage levels and a comparison against similarly priced alternatives.

Other vendors

Every rate card here is built the same way, from the same price snapshot (2026-09-16), so the numbers are comparable across vendors:

Anthropic · Google · DeepSeek · Alibaba · xAI · Zhipu · Moonshot · Meta · MiniMax · ByteDance

FAQ

How much does OpenAI charge per million tokens?

Across the 80 priced text models in this catalogue, the blended rate (input plus output per 1M tokens) runs from $0.225 on gpt-5-nano to $105 on gpt-5.5-pro, as of 2026-09-16. The spread is what matters: choosing the wrong tier inside one vendor usually costs more than switching vendor.

What is the cheapest OpenAI model for high-volume work?

gpt-5-nano is the lowest blended rate here at $0.025 input and $0.2 output per 1M tokens. On a Heavy workload — 500M input and 100M output tokens a month — that bills about $32.50. At the other end, the same workload on gpt-5.5-pro costs about $16,500.00.

Is OpenAI cheaper through a gateway than buying direct?

For the 4 models where a published OpenAI list price could be verified, the gateway rate is below list by between 38% and 50%, checked 2026-09-16. The rate is not fixed — it moves with upstream promotions — and several models in this catalogue have no published list price at all, so no comparison is claimed for them.

How does output pricing change a OpenAI bill?

Output is billed at a multiple of input on every model priced here, so the input rate alone never predicts the bill. A workload that is mostly input tokens still lands most of its cost on the output column once that multiple is applied. Price your own in/out split in the calculator rather than extrapolating from the headline input rate.

Why is GPT-5.6 Sol's discount smaller than the rest of the family?

Because OpenAI is already discounting Sol itself with a promotional rate, so there is less headroom left underneath it. The other models in the family carry no such promotion, which is why the gap between gateway rate and list price is wider on Terra, Luna and Astra than it is on Sol.

Create free account

See all 827 models → or use the cost calculator with your own token split.