Word counter with a CEFR reading level

Two people at a desk with an abacus, a numbers key, a bar chart, a ruler and a magnifying glass — counting the words in a text and measuring how hard it is to read.

Words, characters, sentences, paragraphs and reading time, the way any counter gives them to you — and then the part no other counter can do: an estimate of which CEFR level a learner would need in order to read the text comfortably, worked out by checking every word in it against the levelled catalogue behind our vocabulary lessons. It runs as you type and nothing you paste leaves the page.

Words · characters · sentencesReading timeCEFR estimate9,278 levelled wordsNo sign-up

Count it, and see what level it is

Counts update as you type. The CEFR estimate is checked against 9,278 levelled words from our own lessons.

🔒 No sign-up  Nothing you paste is uploaded, stored or logged.

Words
96
Characters
501
with spaces
Characters
406
no spaces
Sentences
5
Paragraphs
1
Avg. sentence
19.2
words
Reading time
29 sec
fluent, silent
Learner time
48 sec
B1 reader
Read aloud
44 sec
out loud
B2CEFR

Estimated reading level

A B2 reader knows about 97% of these words, which is enough to read it with a dictionary now and then. To read it comfortably, without stopping takes more than C1.

The hardest words in it

B2 and C1 words, linked to the page that teaches each one.

  • bottomB2
  • documentB2
  • placesB2

Off our lists

Not in the words this site teaches. Often names, numbers or ordinary words we simply have not used yet — not necessarily hard.

  • atlas
  • librarian
  • historical

How to use the level estimate

The number at the top is only useful if you know what question it answers. It answers this one: at what level does a learner know enough of these words to get through this text?

If you are a teacher, this is the fastest way to check that a text you found is at the level of the class you found it for. Paste it in, look at the coverage bars, and read the off-list words. Those off-list words are your pre-teaching list — they are the ones your learners will not have met on any of our lists at any level, which usually means a handful of names, a few technical terms, and two or three words genuinely worth putting on the board.

If you are a learner, paste something you have written. If it comes back below your own level, you are writing under your ceiling — you know more words than you are using, which is extremely common and easy to fix. If it comes back above your level, look at which words did it. Half the time it is one unnecessary word where a shorter one was available, which is the definition of writing that is harder to read than it needs to be.

Either way, read the sentence length too. It is reported separately for a reason, and that reason is the whole limitation of the method, explained below.

What the CEFR estimate actually measures

A line chart of cumulative vocabulary coverage rising from A1 to C1 for a sample paragraph, crossing the 95 per cent threshold at B2, which is the level the tool reports.
The estimate is not a verdict — it is the point where the coverage line crosses 95%. Drawn from the worked example further down this page.

This is a lexical coverage measure, and it is worth being precise about that, because "readability" is a word that gets used for a dozen different things.

Coverage is the share of the running words in a text that a reader already knows. Vocabulary research has converged on two useful figures: at roughly 95% coverage a reader can get through a text with occasional help, and at roughly 98% they can read it comfortably and unaided. Both numbers are reported here rather than one, because they answer different questions and because presenting a single threshold as if it were a law of nature would be overstating what is a well-supported convention.

So the tool takes every word in your text, finds the lowest level at which this site teaches it, and adds up the running total. The level it reports is the first one at which the total passes 95%. It is a count, not a model, and you can check it by hand if you disagree with it.

  • Function words count as A1. The, of, and, was, which — the grammar words. They appear on no topic vocabulary list anywhere because they are grammar rather than vocabulary, and a learner meets all of them in their first term. Roughly 40% of any English text is made of them, which is why the chart on this page starts at 40% rather than zero.
  • Endings are undone before giving up. Walked is looked up, missed, stripped to walk, and found. Plurals, -ing forms, comparatives, adverbs in -ly, the common contractions and the irregular verb table are all handled, so an inflected word is not counted as unknown.
  • Names and numbers are set aside. A capitalised word in the middle of a sentence is treated as a name and left out of the calculation entirely. It is not vocabulary in either direction, and counting Barcelona as an unknown C1 word would make every travel text unreadable.
  • Everything left over is "off-list". That means: not in the words this site teaches. It does not mean the word is hard.
One word travelling through the lookup: walked is not in the lists, its -ed ending is undone to walk, which is found at A1 and linked to the page that teaches it. Words that never match come out as a name, a number, or off-list.
Every word takes this route, on your own device. There are three ways out and the results panel tells you which one each word took.

The word list behind all of this is not a general dictionary. It is 5,290 words from our levelled vocabulary catalogue, plus 4,005 more taken from the levelled stories and exercises we publish — 9,278 words in total, each at the lowest level our own material uses it. A word in an A1 story is a word we hand to an A1 learner and expect them to read, which is as good a definition of "an A1 word" as this site can honestly offer.

How wrong this can be, and in which direction

It is an estimate. Not a measurement, not a score, and not something to put in a report. Here is what limits it, stated as plainly as we can manage, because a tool that hides its error bars is worse than no tool.

  • It is a picture of our syllabus, not of English. The word list behind it is what we teach, at the level we teach it. Another course would place many of the same words a level up or down, and an ordinary word we happen never to have used comes out as off-list however easy it is.
  • It cannot see grammar at all. A third conditional built entirely from A1 words scores A1. This is the biggest single limitation of any coverage measure and it is why sentence length is reported next to the level rather than folded into it — when the sentences are long and the vocabulary is easy, the tool says so out loud instead of quietly changing the number.
  • It takes the easiest sense of a word. Board is an A1 school word and a B2 business word; the lookup takes A1. Every estimate is therefore biased slightly downward — a text is a little more likely to be harder than the level shown than easier.
  • Short texts are noisy. Below about forty countable words, one unusual word moves the whole result by a level. The tool says when your sample is that small rather than reporting a number as if it were solid.
  • It says nothing about content. A text can be linguistically A2 and entirely unsuitable for the reader in front of you. That judgement is yours and no counter is going to make it.

Used within those limits it is genuinely useful, and it is checkable: every word it levelled is a word you can look up on the vocabulary pages, on the page it links to in the results. If you think it has a word at the wrong level, you can see exactly where that level came from — which is more than most readability scores will tell you.

About the reading times

Three figures are given, because one number would be misleading. Silent reading assumes about 200 words per minute, which is a normal adult reading in their first language. Learner reading assumes about 120, which is closer to a B1 reader working through a text in English. Reading aloud assumes about 130, which is the figure to use if you are timing a presentation or a speaking exam answer.

All three are averages over a lot of people and none of them is your speed. If you are planning a two-minute IELTS speaking turn, the aloud figure is close enough to be useful; if you are estimating how long a class will take over a text, add half again and you will be nearer.

A worked example

Here is the tool's complete output for a short paragraph, produced by the same code that runs in the box above and printed here exactly as it came out.

The example paragraph

The library opens at nine and closes at six. On Saturday morning it is usually quiet, so I take my books and sit near the window at the back. Last week I found an old atlas on the bottom shelf, printed before either of my parents was born. The borders were in the wrong places and half the countries had different names. I sat with it until the librarian turned the lights off, and then I walked home in the rain thinking about how quickly a map becomes a historical document rather than a useful one.
Words                96
Sentences            5
Average sentence     19.2 words (long)
Reading time         29 sec silent · 48 sec for a B1 learner

Known by A1            79%
Known by A2            88%
Known by B1            94%
Known by B2            97%   ← crosses 95% here
Known by C1            97%   ← crosses 95% here

Names and numbers    1
Off our lists        3   (atlas, librarian, historical)

ESTIMATE             B2

The coverage curve above is drawn from this same paragraph — the line is the running total, and the level it reports is simply where the line crosses the 95% rule. If you paste this text into the box yourself you will get these numbers back, because it is the same calculation.

It is worth reading that off-list line rather than skipping it, because it is the most honest part of the output. Those 3 words are not in anything we publish — which says something about our word lists, and nothing at all about how hard the word is. A tool that reported them as "difficult" would be making a claim it cannot support.

What to do with the answer

A level is only useful if there is something at that level to go and read.

Frequently asked questions

How does the CEFR level estimate work?

It checks every word in your text against our own levelled catalogue of about 9,600 words, adds up how much of the text a reader at each level would know, and reports the first level at which that total passes 95% of the running words — the point at which vocabulary research says a reader can get through a text with occasional help. It is a vocabulary count, not a grammar analysis.

Is the CEFR level accurate?

It is an estimate with known limits, and the page lists them. The main ones: it cannot see grammar, so a hard sentence made of easy words scores low; it reflects our syllabus rather than English as a whole; and it takes the easiest meaning of an ambiguous word, which biases it slightly downward. Read it as a strong hint, not as a measurement.

Does it work for texts I have written myself?

Yes, and that is the most useful thing to do with it. Paste your own writing and see whether it lands at your level, above it or below it. Writing consistently below your own level is the commonest pattern and the easiest to correct.

What does "off-list" mean?

It means the word is not in the vocabulary, stories or exercises this site publishes. It does not mean the word is difficult or wrong — proper nouns, technical terms and perfectly ordinary words we simply have not used yet all come out off-list.

How is reading time calculated?

Three ways, because readers differ. Silent reading at 200 words per minute for a fluent adult, learner reading at 120, and reading aloud at 130. All three are averages and none is anybody's exact speed.

Is my text saved anywhere?

No. Everything is calculated inside the page. There is no upload, no storage and no logging, and you can confirm it by opening your browser's network panel while you type.