Token Counter
Paste text (or upload a file) to count tokens and see exactly what it costs to send
to GPT, Claude, Gemini, DeepSeek and 550+ other models — using a real BPE tokenizer
(o200k_base), not a word-count estimate. Everything runs in your browser; nothing is uploaded.
Counts are exact for OpenAI models (o200k_base is GPT-4o/4.1/5's tokenizer). For Claude, Gemini and other vendors — whose tokenizers are not public — treat the count as a close estimate (typically within ±15%).
Cost of this text across popular models
| Model | Input $/1M | Output $/1M | This text (input) | +500 output tokens |
|---|---|---|---|---|
| Claude Opus 5.5 (Max, Default Fallback) | 4.00 | 20.00 | $0 | $0 |
| Claude Sonnet 5.5 (Max, Default Fallback) | 2.00 | 10.00 | $0 | $0 |
| Claude Fable 5.1 (Max, Default Fallback) | 10.00 | 50.00 | $0 | $0 |
| GPT-6 Astra (Max) | 10.00 | 50.00 | $0 | $0 |
| Gemini 4 Argon (High) | 2.00 | 10.00 | $0 | $0 |
| GPT-6.1 Sol (Max) | 2.00 | 10.00 | $0 | $0 |
| Claude Opus 5 (Max) | 5.00 | 25.00 | $0 | $0 |
| Claude Fable 5 (Max, Opus 4.8 Fallback) | 10.00 | 50.00 | $0 | $0 |
Need a different model? Every priced model has its own calculator.
What is a token?
LLM APIs don't bill by word or character — they bill by token, the unit the model actually reads. A token is a common chunk of text: whole short words ("the", "and"), pieces of longer words ("token" + "izer"), punctuation, or whitespace. In English, one token averages about 4 characters or ¾ of a word, but the exact split depends on the tokenizer — which is why a real BPE count beats any rule of thumb.
This tool runs o200k_base, the byte-pair-encoding scheme used by OpenAI's current
models, directly in your browser. Your text never leaves the page.
FAQ
How accurate is this token counter?
For OpenAI models (GPT-4o, GPT-4.1, GPT-5 family) it is exact: we run the same o200k_base byte-pair encoding the API uses. Anthropic, Google and most other vendors do not publish their tokenizers, so for Claude, Gemini and others the o200k count is a close estimate — typically within about 15% of the billed figure.
Is my text uploaded anywhere?
No. The tokenizer is a JavaScript bundle that runs entirely in your browser. Nothing you paste or upload leaves your device — you can verify in your browser's network tab.
How is API cost calculated?
Cost = (input tokens ÷ 1,000,000) × the model's official input price, plus (expected output tokens ÷ 1,000,000) × the output price. Prices are official list prices in USD per million tokens, refreshed daily from Artificial Analysis data.
Why do tokens matter more than words?
Every LLM API bills per token, and context windows are measured in tokens. Two prompts with the same word count can differ meaningfully in tokens — code, non-English text and unusual formatting all tokenize differently.
How many tokens is a typical page of text?
A full English page (~500 words) is roughly 650-700 tokens. A rule of thumb is 1 token per 4 characters, but the whole point of this tool is that you don't need rules of thumb — paste the actual text and get the actual number.
Which models can I estimate costs for?
Every model in our comparison database that publishes pricing — 550+ models from OpenAI, Anthropic, Google, Meta, DeepSeek, Alibaba, Mistral, xAI and more. See the full sortable table on the Model Prices page, or jump straight to a per-model counter from the directory at the bottom of this page.
Looking for full price and performance data on all models? See the model comparison table.
Token counters by model
Every priced model gets its own token counter, pre-filled with its price — jump straight to one instead of picking it from the list above.
AI21 Labs
Alibaba
- Qwen3.8 Max (0902)
- Qwen3.8 Max
- Qwen3.8 2.4T A95B
- Qwen3.8-Flash-Next
- Qwen3.8 27B (Xhigh)
- Qwen3.7 Max
- Qwen3.6 Max Preview
- Qwen3.6 Plus
- Qwen3.7 Plus
- Qwen3.5 27B (Reasoning)
- Qwen3.5 397B A17B (Non-reasoning)
- Qwen3.6 27B (Reasoning)
- Qwen3.5 Omni Plus
- Qwen3.5 35B A3B (Reasoning)
- Qwen3.6 35B A3B (Reasoning)
- Qwen3.5 122B A10B (Non-reasoning)
- Qwen3 Max Thinking (Preview)
- Qwen3 Max
- Qwen3 VL 235B A22B (Reasoning)
- Qwen3.5 9B (Non-reasoning)
- Qwen3.5 4B (Reasoning)
- Qwen3 235B A22B 2507 (Reasoning)
- Qwen3 Max (Preview)
- Qwen3.5 Omni Flash
- Qwen3 235B A22B 2507 Instruct
- Qwen3 Coder 480B A35B Instruct
- Qwen3 VL 32B (Reasoning)
- Qwen3 Next 80B A3B (Reasoning)
- Qwen3 VL 235B A22B Instruct
- Qwen3 30B A3B 2507 (Reasoning)
- Qwen3 Coder 30B A3B Instruct
- Qwen3 Next 80B A3B Instruct
- QwQ 32B
- Qwen3 235B A22B (Reasoning)
- Qwen3 VL 30B A3B (Reasoning)
- Qwen3 Coder Next
- Qwen3 32B (Reasoning)
- Qwen3 VL 32B Instruct
- Qwen3 235B A22B (Non-reasoning)
- Qwen3 14B (Reasoning)
- Qwen3 VL 8B (Reasoning)
- Qwen3 VL 30B A3B Instruct
- Qwen3 Omni 30B A3B (Reasoning)
- Qwen2.5 Instruct 72B
- Qwen3 30B A3B (Reasoning)
- Qwen3 30B A3B 2507 Instruct
- Qwen3 32B (Non-reasoning)
- Qwen3 8B (Reasoning)
- Qwen3 VL 8B Instruct
- Qwen3 14B (Non-reasoning)
- Qwen3 30B A3B (Non-reasoning)
- Qwen2.5 Turbo
- Qwen3 8B (Non-reasoning)
- Qwen3 Omni 30B A3B Instruct
Allen Institute for AI
Amazon
Anthropic
- Claude Opus 5.5 (Max, Default Fallback)
- Claude Sonnet 5.5 (Max, Default Fallback)
- Claude Fable 5.1 (Max, Default Fallback)
- Claude Opus 5 (Max)
- Claude Fable 5 (Max, Opus 4.8 Fallback)
- Claude Haiku 5.5 (Max)
- Claude Opus 4.8 (Max)
- Claude Opus 4.7 (Max)
- Claude Sonnet 5 (Max)
- Claude Opus 4.6 (Max)
- Claude Sonnet 4.6 (Max)
- Claude Opus 4.5 (Reasoning)
- Claude Opus 4.6 (Non-reasoning, High)
- Claude Sonnet 4.6 (Non-reasoning, High)
- Claude Opus 4.5 (Non-reasoning)
- Claude Sonnet 4.6 (Non-reasoning, Low)
- Claude 4.1 Opus (Reasoning)
- Claude 4.5 Sonnet (Reasoning)
- Claude 4 Opus (Reasoning)
- Claude 4.5 Sonnet (Non-reasoning)
- Claude 4.1 Opus (Non-reasoning)
- Claude 4.5 Haiku (Reasoning)
- Claude 4 Opus (Non-reasoning)
- Claude 4.5 Haiku (Non-reasoning)
- Claude 3.7 Sonnet (Non-reasoning)
- Claude 3 Opus
- Claude 3.5 Sonnet (Oct '24)
- Claude 3.5 Sonnet (June '24)
- Claude 3 Haiku
- Claude Fable 5 (Adaptive Reasoning, Max Effort, Opus 5 Fallback)
Apodex
Arcee AI
Baidu
ByteDance Seed
Celeris
Cohere
Deep Cogito
DeepSeek
- DeepSeek V4.1 Flash (Max)
- DeepSeek V4 Pro 0813 (Max)
- DeepSeek V4 Flash Vision (Max)
- DeepSeek V4 Flash 0731 (Max)
- DeepSeek V4 Pro 0424 (Max)
- DeepSeek V4 Flash 0420 (High)
- DeepSeek V3.2 (Reasoning)
- DeepSeek V3.2 Exp (Reasoning)
- DeepSeek V3.2 (Non-reasoning)
- DeepSeek V3.1 Terminus (Reasoning)
- DeepSeek V3.1 Terminus (Non-reasoning)
- DeepSeek V3.2 Exp (Non-reasoning)
- DeepSeek V3.1 (Non-reasoning)
- DeepSeek V3.1 (Reasoning)
- DeepSeek R1 0528 (May '25)
- DeepSeek R1 (Jan '25)
- DeepSeek V3 0324
- DeepSeek V3 (Dec '24)
- DeepSeek R1 Distill Llama 70B
- Gemini 4 Argon (High)
- Gemini 3.8 Flash (High)
- Gemini 3.7 Flash (Medium)
- Gemini 3.6 Flash (High)
- Gemini 3.5 Flash (Medium)
- Gemini 3.1 Pro Preview
- Gemini 3 Pro Preview (High)
- Gemini 3 Flash Preview (Reasoning)
- Gemini 3.5 Flash-Lite
- Gemini 3 Flash Preview (Non-reasoning)
- Gemma 4 26B A4B (Reasoning)
- Gemini 2.5 Pro
- Gemini 3.1 Flash-Lite
- Gemini 2.5 Pro Preview (May' 25)
- Gemma 4 12B (Reasoning)
- Gemini 2.5 Flash (Reasoning)
- Gemini 2.5 Flash-Lite Preview (Sep '25) (Reasoning)
- Gemini 2.5 Flash (Non-reasoning)
- Gemini 2.5 Flash-Lite Preview (Sep '25) (Non-reasoning)
- Gemma 4 E4B (Reasoning)
- Gemini 2.5 Flash-Lite (Reasoning)
- Gemini 2.5 Flash-Lite (Non-reasoning)
Inception
InclusionAI
Kimi
KwaiKAT
LongCat
Meta
Microsoft
Mistral
Multiverse Computing
NVIDIA
- Nemotron 3 Ultra 550B A55B (Reasoning)
- Nemotron 3.5 Lightning
- Nemotron 3 Super 120B A12B (Reasoning)
- Nemotron 3 Nano Omni 30B A3B Reasoning
- NVIDIA Nemotron 3 Nano 30B A3B (Reasoning)
- NVIDIA Nemotron Nano 9B V2 (Reasoning)
- NVIDIA Nemotron 3 Nano 30B A3B (Non-reasoning)
- NVIDIA Nemotron Nano 9B V2 (Non-reasoning)
- NVIDIA Nemotron Nano 12B v2 VL (Non-reasoning)
Nous Research
OpenAI
- GPT-6 Astra (Max)
- GPT-6.1 Sol (Max)
- GPT-6 Sol (Max)
- GPT-5.6 Sol (Max)
- GPT-5.6 Terra (Max)
- GPT-5.4 (Xhigh)
- GPT-5.5 (Xhigh)
- GPT-6 Luna (Max)
- GPT-5.6 Luna (Max)
- GPT-5.3 Codex (Xhigh)
- GPT-5.2 (Xhigh)
- GPT-5.2 Codex (Xhigh)
- GPT-5.5 Instant (June 2026)
- GPT-5 Codex (High)
- GPT-5.1 (High)
- GPT-5.4 mini (Xhigh)
- GPT-5.1 Codex (High)
- GPT-5 (High)
- GPT-5.5 Instant (May 2026)
- o3-pro
- GPT-5.4 nano (Xhigh)
- GPT-5 mini (Medium)
- GPT-5.1 Codex mini (High)
- o3
- o4-mini (High)
- o1
- GPT-5 nano (High)
- GPT-4.1
- o3-mini
- o1-pro
- gpt-oss-120b (High)
- o1-preview
- GPT-4.1 mini
- gpt-oss-20b (Low)
- GPT-4o (Nov '24)
- GPT-4.1 nano
- GPT-4o (Aug '24)
- GPT-4o (May '24)
- GPT-4 Turbo
- GPT-4
- GPT-4o mini
- GPT-3.5 Turbo
- GPT-5.4 Pro (Xhigh)