Token prices(USD/1M tokens)

USD / 1M tokens
ModelInputOutputCached inputSource
grok-4.7xAI $2.00$6.00$0.50Source Oct 10
grok-build-0.1xAI $1.00$2.00$0.20Source Oct 10
grok-4.6xAI $2.00$6.00$0.50Source Oct 10
grok-4.5xAI $2.00$6.00$0.30Source Oct 10
grok-4.3xAI $1.25$2.50$0.20Source Oct 10
grok-4.20-multi-agent-0309xAI $1.25$2.50$0.20Source Oct 10
grok-4.20-0309-reasoningxAI $1.25$2.50$0.20Source Oct 10
grok-4.20-0309-non-reasoningxAI $1.25$2.50$0.20Source Oct 10

Standard API rates. Taxes, tools and subscriptions excluded.

Context

grok-4.7 · Global endpoint, below 200k prompt tokens.

grok-build-0.1 · Global endpoint, below 200k prompt tokens.

grok-4.6 · Global endpoint, below 200k prompt tokens.

grok-4.5 · Global endpoint, below 200k prompt tokens.

grok-4.3 · Global endpoint, below 200k prompt tokens.

grok-4.20-multi-agent-0309 · Global endpoint, below 200k prompt tokens.

grok-4.20-0309-reasoning · Global endpoint, below 200k prompt tokens.

grok-4.20-0309-non-reasoning · Global endpoint, below 200k prompt tokens.

Estimate token cost

Estimated cost—

Token-only estimate, not an actual task bill. Extra tool calls and retries are excluded.

Value for money →

Frequently asked questions

More
What unit are AI token prices quoted in?

Prices are US dollars per one million tokens. Input, output and cached input are separate rates. These are API prices, not ChatGPT, Claude or Grok subscription fees.

How is an API request cost calculated?

Multiply each token count by its matching rate, add the amounts, then divide by one million. Cached input replaces part of the input count; it is not counted twice. The calculator uses verified rates and shows unknown when a required rate is unavailable.

Does changing a model change the table order?

On the home comparison table, selecting a model updates only that provider row. Click an input or output column arrow to sort explicitly. A low input price does not necessarily mean a low total cost.

Can I assume every input token is billed at the cached price?

No. A listed cached rate does not guarantee a cache hit. Use cached billing only for input covered by the provider’s cache rules. The calculator requires cached input to be no greater than total input and does not assume a discount when its rate is unknown.

Can the token price change for a long context?

Some published pricing conditions distinguish context lengths or other request settings. Check the context notes and source for the selected model. The calculator uses a recorded long-context rate when available; it does not invent a missing rate.

Is the model with the cheapest input always cheapest overall?

No. Total cost depends on input, output and applicable cached tokens. A model with cheap input but expensive output may cost more for a long answer. Enter your expected token counts in the calculator and compare the rates separately.