Count tokens and price a prompt before you send it.
Paste a prompt or a document to see roughly how many tokens it is, how much of a context window it fills and what it costs to send at scale, for GPT, Claude, Gemini or any model you add. Counted in your browser. Nothing is sent anywhere.
- Free, no sign-up
- Counted in your browser
- GPT, Claude and Gemini
What it costs to send
Enter the reply length and how often you send it. Edit any price: providers change them often.
Caching and batch pricing
| Model | Input $ / 1M | Output $ / 1M | Window (K) | Per request | Per month | Compare |
|---|
Standard list prices checked on 7 October 2026, for prompts under 200K tokens. They go out of date: check the provider’s pricing page and change any number above. Token counts are estimates and can differ from a provider’s own count by 10 to 15 percent.
How to count tokens and estimate API cost
Paste your prompt
Paste or open the text you will send. The estimate updates as you type.
Set the reply and the volume
Enter how long the reply will be and how many requests you send each month.
Compare the models
Read the cost per request and per month for each model. Change any price to match the provider’s page, or add a model of your own.
Language models read and bill text in tokens: pieces of words, numbers and symbols. A short English word is usually one token, a long word several, and code, numbers and non-English text use more per word. Knowing the token count tells you whether a prompt fits in a model’s context window and what it will cost you at scale. This page estimates both in your browser, without sending your text anywhere.
What it does
- Estimates the token count of any text and shows how it splits between words, symbols and code, numbers and non-Latin text.
- Shows how much of an 8K, 32K, 128K, 200K and 1M context window the prompt fills.
- Works out the cost per request and per month from your reply length and request volume, for several models side by side.
- Lets you edit every price, add your own models and keep them in your browser for next time.
- Handles prompt caching and the batch discount, which can cut the bill a lot for repeated prompts.
- Opens text, Markdown, JSON, CSV and code files on your device.
Limits, honestly
- The count is an estimate, usually within 10 to 15 percent. Each provider’s tokenizer is different, and their vocabularies are too large to run in a web page. For an exact count use the provider’s own tokenizer or token-counting API.
- The prices were checked on 7 October 2026. Providers change them often and add models: check the pricing page and edit the numbers.
- It counts text only. Images, audio and files sent as attachments are priced differently by each provider.
- Hidden costs are not included: reasoning tokens some models spend before answering, tool calls and retries.
- Context windows listed are common sizes. A model’s real limit and its rules for output length are on its documentation page.
LLM Token Counter & API Cost Calculator: questions and answers
What is a token in an LLM?
How accurate is this token counter?
How many words is 1,000 tokens?
How is LLM API cost calculated?
Why do the prices differ from the provider’s page?
Is my text uploaded or stored?
What happens if a prompt is longer than the context window?
Why does non-English text use more tokens?
More free tools
Need a custom tool, site or store?
I scope, design and build tools, websites and Shopify stores for teams in India and abroad. Send a short brief and I will reply within 48 hours.