Tools · Prompt Calculator

AI Token Calculator

725 models priced

A free token counter for every model in the index. Paste a prompt to count its tokens exactly and see what one request costs; to project a monthly bill from request volume, use the budget calculator.

As of October 2026, the models in this calculator blend to a median of $0.80 per million tokens (our 70/30 weighting of published rates, not a price any provider quotes), with half between $0.24 and $2.38.

Model
Anthropic
$4.000/M input$20.000/M output1,000,000 context◇ approximate · cl100k_base
0 tokens · 726 chars
Sample: Article summary (RAG)$0.00 per request
0 tokens · 405 chars
$0.00 per request
Per request
$0.00
0 in · 0 out
Per 1,000 requests
$0.00
at the same input + output size
Per 100K requests
$0.00
linear scaling
Per 1M requests
$0.00
at production scale
How this works, and common questionsExplainer · FAQ · Glossary

A free AI token calculator and token counter. OpenAI models use the exact OpenAI BPE encoder; Claude, Gemini and open-weights families use cl100k_base as a documented approximation (typically within ~5–10% for English). Browse the AI model prices for the full rate sheet, or use the side-by-side model comparison to evaluate models on benchmarks and capabilities.

A workload of 10 million tokens a month therefore costs about $8.00 on the median model and $0.20 on the cheapest paid one. 10 of the models priced here carry no per-token charge at all. The most expensive is OpenAI o1-pro at $285.00 per million tokens blended.

How this AI token calculator works

AI models bill per token, and a token is roughly four characters of English text. Every request is priced in two parts: the input tokens you send (system prompt, context, and message) and the output tokens the model generates. Providers publish these as per-million-token rates. To price a single request, multiply your input token count by the input rate and your output token count by the output rate, then add the two — which is exactly what this page does once you paste the text. Counting happens in your browser: nothing you paste is sent to a server. To turn a per-request figure into a monthly bill, take the same numbers to the AI budget calculator, which projects spend from your request volume.

What a typical request costs

A single request of 1,000 input tokens (roughly a page of context and a short question) and 300 output tokens (a few paragraphs of answer), priced on live rates at three points across the index:

Cost = (input tokens ÷ 1,000,000 × input rate) + (output tokens ÷ 1,000,000 × output rate). Rates are read from the live index; models are picked by position in the price distribution, not hand-chosen.
ModelInput, per 1MOutput, per 1MThis request
DeepSeek-R1-Distill-8BBudget tier · DeepSeek$0.070$0.200$0.00013
Mistral Large 3Median model · Mistral$0.500$1.50$0.00095
Kimi K3Premium tier · Moonshot AI$3.00$15.00$0.00750

Multiply by your request volume to project a month: at 100,000 requests the same workload runs $13.00 on DeepSeek-R1-Distill-8B and $750 on Kimi K3. That gap, not the per-request figure, is what makes model choice a budget decision.

Frequently Asked Questions

How many tokens are in a word?

For ordinary English prose, one token averages about four characters, which works out to roughly 0.75 words per token, so 1,000 tokens is around 750 words. The ratio shifts with the content: code, rare words, other languages, and heavy punctuation all tokenize into more tokens per word. Paste your actual text above for an exact count rather than relying on the average.

How do I count the tokens in a prompt?

Paste the prompt into the input box above and the calculator counts it in your browser. OpenAI models are counted with the exact BPE encoder those models use, so the figure matches what the API will bill. Other families are counted with cl100k_base as a documented approximation, typically within 5 to 10 per cent for English text.

What is the difference between input and output token pricing?

Input tokens are everything you send to the model (your system prompt, context, and user message) while output tokens are everything the model generates in its completion. Providers almost always charge more per output token than per input token, so a model with cheap input but expensive output can still be costly for long, generative workloads.

Is this AI token calculator free?

Yes. The calculator is free to use, needs no sign-up, and counts tokens in your browser, the text you paste is never sent to a server. Pricing comes from the same provider-published rates as the rest of Tokenando.

Cost terms explained