Base Difficulty

Base difficulty: Level 162 edit

Level distribution for noun (quantitative_concept): 20: 1 162: 24 276: 25 380: 25 390: 1 456: 1 1000: 1 1050: 2 1070: 2 1080: 10 1090: 1 1110: 1 1202: 2 1203: 1 1204: 2 1206: 1 1210: 2 1211: 1 1213: 7 1218: 1 1225: 2 1226: 1

Per-Language Difficulty Overrides

No difficulty overrides set. The base difficulty level will be used for all languages.

Add Difficulty Override
-1 to exclude, 1-1299 for level
Frequency and Tier Signals
Combined Frequency Rank: — (run python -m storage.admin --calc-ranks to populate)
Tier Signals
Source Tier Rank used
basic_english Basic (1 of 2) 600
cefr A2 (2 of 6) 2100

Rank used = synthetic rank this tier contributes to the combined frequency rank (lower = more common).

Corpus Frequency
Corpus Freq (per M) Best form Share Zipf rank Forms
19th_books 113.82 820 100% 820 1
20th_books 98.50 973 100% 973 1
wiki_math 75.61 1476 100% 1476 1
wiki_geography 183.78 673 100% 673 1
wiki_biology 81.83 1685 100% 1685 1
wiki_modern_life 333.58 317 100% 317 1
wiki_arts 80.28 1656 100% 1656 1
wiki_society 308.02 354 100% 354 1
wiki_linguistics 47.16 2328 100% 2328 1
wiki_physical_science 932.36 99 100% 99 1
wiki_history 323.32 312 100% 312 1
early_modern_science 431.17 247 100% 247 1
religious_translated 132.13 794 100% 794 1
cooking 37.39 2115 100% 2115 1
legal_scotus 237.76 535 100% 535 1
eu_parliament_debates 271.68 457 100% 457 1
openstax_science 905.53 100 100% 100 1
openstax_society 256.32 524 100% 524 1
wpa_life_histories 49.16 1826 100% 1826 1

Best form = the single best-ranked form in the corpus, share-blind — how the spelling itself ranked. Share = how much of that spelling belongs to this lemma, when several senses compete for it. Zipf rank = all the forms combined (variant spellings included), each scaled by its share, and what feeds the combined frequency rank. Combining forms improves on the best form, but a share below 100% pushes the other way and usually wins: a sense holding a small slice of a common spelling is a rare word, so its rank lands downrated below the best form — and once it falls past what the corpus can support it is capped at the rank an unlisted word would get.