API cost calculator
OpenAI API pricing calculator
Estimated cost
Total for this run
$75.00
$0.0075 per request
- Input
- $15.00
- Output
- $60.00
gpt-5.1 · $1.25 per 1M in · $10.00 per 1M out
Cheapest model for this workload: gpt-5-nano · $3.00
Costing an AI agent? We already built one.
Spur's agent answers WhatsApp, Instagram and Messenger for your customers, with no token maths to do yourself.
Trusted by 600+ businesses worldwide
How to calculate OpenAI API cost
- 01
Pick the model and tier
Batch is half price if the work can wait up to 24 hours.
- 02
Enter your tokens
Tokens per request, then how many requests. Roughly 4 characters is 1 token.
- 03
Read the total
Split by input and output, with the cost of a single request.
Three ways to pay less
Drop down a model size
The same job on gpt-5-nano costs a fraction of gpt-5.1. Try it before optimising anything else.
Reuse the same prompt prefix
Repeated input is billed as cached, often at a tenth of the normal rate.
Batch anything that can wait
Half price for work returned within 24 hours. Good for backfills and evals.
OpenAI API prices per 1M tokens
Prices last checked September 2, 2026. OpenAI's pricing page
| Model | Input | Cached | Output | Batch in |
|---|---|---|---|---|
| gpt-5.6-sol | $4.00 | $0.40 | $20.00 | $2.00 |
| gpt-5.6-terra | $2.00 | $0.20 | $12.00 | $1.00 |
| gpt-5.6-luna | $0.20 | $0.02 | $1.20 | $0.10 |
| gpt-5.5 | $5.00 | $0.50 | $30.00 | $2.50 |
| gpt-5.5-pro | $30.00 | – | $180.00 | $15.00 |
| gpt-5.4 | $2.50 | $0.25 | $15.00 | $1.25 |
| gpt-5.4-mini | $0.75 | $0.07 | $4.50 | $0.38 |
| gpt-5.4-nano | $0.20 | $0.02 | $1.25 | $0.10 |
| gpt-5.4-pro | $30.00 | – | $180.00 | $15.00 |
| gpt-5.2 | $1.75 | $0.17 | $14.00 | $0.88 |
| gpt-5.2-pro | $21.00 | – | $168.00 | $10.50 |
| gpt-5.1 | $1.25 | $0.13 | $10.00 | $0.63 |
| gpt-5 | $1.25 | $0.13 | $10.00 | $0.63 |
| gpt-5-mini | $0.25 | $0.03 | $2.00 | $0.13 |
| gpt-5-nano | $0.05 | $0.0050 | $0.40 | $0.03 |
| gpt-5-pro | $15.00 | – | $120.00 | $7.50 |
| gpt-4.1 | $2.00 | $0.50 | $8.00 | $1.00 |
| gpt-4.1-mini | $0.40 | $0.10 | $1.60 | $0.20 |
| gpt-4.1-nano | $0.10 | $0.03 | $0.40 | $0.05 |
| gpt-4o | $2.50 | $1.25 | $10.00 | $1.25 |
| gpt-4o-mini | $0.15 | $0.07 | $0.60 | $0.07 |
| o3 | $2.00 | $0.50 | $8.00 | $1.00 |
| o3-pro | $20.00 | – | $80.00 | $10.00 |
| o4-mini | $1.10 | $0.28 | $4.40 | $0.55 |
| o3-mini | $1.10 | $0.55 | $4.40 | $0.55 |
| o1 | $15.00 | $7.50 | $60.00 | $7.50 |
| o1-pro | $150.00 | – | $600.00 | $75.00 |
| gpt-4-turbo | $10.00 | – | $30.00 | $5.00 |
| gpt-3.5-turbo | $0.50 | – | $1.50 | $0.25 |
| text-embedding-3-small | $0.02 | – | – | – |
| text-embedding-3-large | $0.13 | – | – | – |
| text-embedding-ada-002 | $0.10 | – | – | – |
Standard tier unless marked. Prices in US dollars. OpenAI changes these without notice, so check the source before budgeting.
What this estimate leaves out
The token maths is exact. The things around it are not.
- 01
Reasoning models bill their thinking
Reasoning tokens are billed as output even though you never see them, so a real bill can be several times an estimate based on the visible reply.
- 02
System prompts count every time
Your instructions go with every request, not once per chat. A 2,000 token prompt across 10,000 calls is 20M input tokens before anyone speaks.
- 03
Prices change without warning
OpenAI moves these figures and retires models on its own schedule. Check the date above before you commit a budget.
Frequently asked questions
How OpenAI charges, what a token is, and where the bill comes from.
Cost is (input tokens ÷ 1,000,000 × input price) + (output tokens ÷ 1,000,000 × output price), multiplied by the number of requests. Input and output are priced separately and output is always dearer, usually four to eight times. The calculator above does this for you, including the cached and batch rates.
It depends entirely on the model. On current prices, gpt-5-nano is $0.05 per million input tokens while gpt-5.5-pro is $30. For a sense of scale, a support assistant sending 1,200 input and 600 output tokens over 10,000 conversations costs roughly $75 on gpt-5.1 and about $3 on gpt-5-nano.
That is the unit every OpenAI price is quoted in. One million input tokens is $1.25 on gpt-5, $0.25 on gpt-5-mini and $0.05 on gpt-5-nano. Output on the same models is $10, $2 and $0.40. A million tokens is roughly 750,000 English words.
A token is a chunk of text, usually part of a word. In English it averages about four characters, so 1,000 tokens is around 750 words. Both what you send and what the model returns are counted, which is why a chatty system prompt costs money on every single call.
Reading your prompt can be done in parallel; writing the reply has to happen one token at a time, so it occupies the hardware for longer. Output typically costs four to eight times input, which means the cheapest thing you can do is ask for shorter answers.
When consecutive requests start with the same long prefix, OpenAI can reuse its work on that prefix and bills those tokens at a much lower rate, often a tenth. It applies automatically. To benefit, put the parts that never change at the very start of the prompt and the parts that vary at the end.
Half. You submit a file of requests and get results back within 24 hours, at 50% of the standard rate for both input and output. It suits evaluations, backfills, bulk classification and anything overnight. It is not usable for anything a person is waiting on.
Among current text models, gpt-5-nano at $0.05 input and $0.40 output. Before reaching for a bigger one, try the smallest that passes your evals: the gap between the nano and flagship tiers is often more than twentyfold, which is a larger saving than any prompt optimisation will give you.
No. API usage is paid from the first request, separately from any ChatGPT subscription. New accounts sometimes get a small trial credit that expires, and OpenAI has at times offered free tokens in exchange for sharing data, but there is no standing free allowance.
In the order that pays best: move to a smaller model, cap output length, restructure prompts so the fixed part comes first and gets cached, batch anything that can wait 24 hours, and trim the system prompt. Switching model usually beats every prompt tweak combined.
They are billed differently and do not overlap. ChatGPT Plus is a flat monthly fee for a person using the app. The API is metered per token for software you build, and a subscription grants no API credit. Light API use costs far less than a subscription; heavy use costs far more.
The date above the table is when they were last read off OpenAI's own pricing page, and that page is linked next to it. OpenAI changes prices and retires models without notice, so treat this as an estimate and confirm against the source before you commit to a budget.
Before you build one. Build the agent, or just switch one on.
Spur's AI agent is trained on your brand and answers every customer message across WhatsApp, Instagram and Messenger, 24/7.
- Meta Business Partner
- Shopify Partner
- 7-day free trial
- Connects to Shopify in minutes
- Billed per conversation delivered





