Vocabulary-size benchmarks for CEFR levels are useful enough to be worth knowing, but precise enough to cause real confusion if treated as an exact, universally agreed number rather than a rough, source-dependent estimate that varies based on how "knowing" a word is actually defined and measured.
Why exact numbers vary by source
Different studies and frameworks estimate vocabulary size somewhat differently, since 'knowing' a word can be defined at different depths — recognition versus active use.
Different vocabulary-size studies define "knowing" a word differently — some count only active production ability, others include passive recognition, and this methodological difference alone can shift reported vocabulary-size estimates for the same CEFR level considerably, which is why comparing numbers from different sources without checking their underlying definition can be misleading.
Using benchmarks as a guide, not a gate
Vocabulary size is a useful rough indicator of level, but reading comprehension depends on more than word count alone — grammar and context matter too.
Vocabulary size functions best as a rough, directional guide rather than a strict gate for content selection, since reading comprehension depends on considerably more than raw word count alone — grammatical knowledge, contextual inference ability, and familiarity with common sentence structures all contribute meaningfully alongside vocabulary size itself.
Applying it to content choice
Matching a story's vocabulary load to a student's approximate known-word range keeps extensive reading in its effective, comfortable zone.
Matching a story's vocabulary load to a student's approximate known-word range, using these benchmarks as a starting estimate rather than a precise cutoff, keeps content selection reasonably grounded without pretending to a level of precision the underlying research doesn't actually support.
Using benchmarks practically despite their imprecision
Rather than seeking a single definitive vocabulary-size number per level, using a benchmark range from a reputable source as a rough starting point for content selection, then adjusting based on how a specific student or class actually performs against real material, produces more reliable outcomes than treating any single cited number as precisely authoritative.
Why vocabulary size alone doesn't fully predict reading success
Two students with similar vocabulary-size test scores can perform quite differently on the same reading passage if one has stronger grammatical intuition or more experience inferring meaning from context — vocabulary size is a strong contributing factor to reading ability, not a complete predictor on its own.
When direct vocabulary-size testing is worth the added effort
For most ongoing classroom purposes, reading comprehension performance and quiz scores provide a more practical, continuously updated signal of vocabulary adequacy than periodic direct vocabulary-size testing, though a dedicated vocabulary assessment can be useful for a specific diagnostic purpose, such as investigating why a student's reading comprehension seems to lag behind their apparent grammatical ability.
A caution against overreliance on any single benchmark source
Because vocabulary-size estimates vary meaningfully across different research sources and methodologies, a school citing a specific number to parents or in official reporting should ideally note which source that figure comes from, or better, frame progress in terms of CEFR level and observed reading performance rather than a precise vocabulary count that could easily be contradicted by an equally credible source using a different counting method.
A caveat on comparing students across different tests
Two students who took different vocabulary-size assessments shouldn't be compared directly on their raw scores, since the underlying test design affects the number as much as genuine ability does — comparisons are only meaningful when the same instrument was used for both students.
Where iRead fits: Story vocabulary is pre-mapped to CEFR level directly, so content selection relies on that existing mapping rather than requiring a school to calculate vocabulary-size targets manually from published benchmark tables.
See it in iRead: iRead's four word games train recall, connection, construction, and recognition on every story's vocabulary.
See the four gamesKey takeaway
Vocabulary-size benchmarks per CEFR level are a useful rough guide, not a precise, universally agreed figure — reading ability depends on more than word count alone, so these numbers work best as a starting point for content selection, refined by how students actually perform against real material.
Frequently asked questions
Is vocabulary size the best single predictor of reading ability?
It's one strong predictor among several — grammar knowledge and reading strategy also matter significantly.
Should schools test vocabulary size directly?
It can be useful periodically, though ongoing reading and quiz performance often give a more practical, continuous signal.
Should vocabulary size be reported to parents as a specific number?
Given how much vocabulary-size figures vary by measurement method, reporting general progress in terms of CEFR level and specific words recently learned tends to be clearer and more meaningful for parents than a precise vocabulary-count figure that may not mean what it appears to at face value.