Tokens per word

How Many Tokens Per Word?

One of the most-asked questions about LLMs. Enter any word count to convert it to tokens instantly.

Short answer

About 1.3 tokens per English word — roughly 0.75 words per token.

Rule of thumb for ordinary English prose. Code, non-English text, and heavy punctuation tokenize less efficiently.

Words

Estimated tokens

≈ 1,300

1,000 words × 1.3 tokens/word

Words to tokens cheat sheet

WordsTokens (approx.)
1 word1.3
10 words13
50 words65
100 words130
500 words (1 page)650
1,000 words1,300
10,000 words13,000
100,000 words (a novel)130,000

Based on 1.3 tokens per English word. Actual counts vary with vocabulary, formatting, and the model's tokenizer — use the token calculator for exact numbers.

FAQ

How many tokens is 1 word?

About 1.3 tokens on average for English. Common words like “the” are often a single token; rare or long words split into several.

Why isn't it exactly 1 token per word?

Tokenizers split text into subword pieces based on frequency, not dictionary words. Frequent words get one token; uncommon ones are broken into fragments, which pushes the average above 1.

Does tokens per word differ by language?

Yes. English is the most token-efficient at ~1.3 per word. Many other languages cost noticeably more per word because tokenizers were trained mostly on English text.

How many tokens per word for code?

Code usually costs more than prose — roughly 1.5 to 2+ tokens per word equivalent. Braces, indentation, camelCase names, and long ID strings all split into extra pieces.

Is “4 characters = 1 token” the same rule?

It is the same idea from the other side. An average English word plus its following space is about 5 characters, and 5 ÷ 4 ≈ 1.25 — consistent with 1.3 tokens per word.

How do I get the exact token count?

Paste your text into the free token calculator on the homepage. It runs a real tokenizer in your browser and shows the exact count.