Base Difficulty

Base difficulty: Level 20 edit

Level distribution for adjective (quality): 15: 5 16: 5 18: 23 19: 43 20: 35 21: 1 36: 5 37: 5 38: 5 39: 5

Per-Language Difficulty Overrides

No difficulty overrides set. The base difficulty level will be used for all languages.

Add Difficulty Override
-1 to exclude, 1-100 for level
Frequency and Tier Signals
Combined Frequency Rank: — (run python -m storage.admin --calc-ranks to populate)
Tier Signals
Source Tier Rank used
cambridge_yle Flyers (3 of 3) 1200

Rank used = synthetic rank this tier contributes to the combined frequency rank (lower = more common).

Corpus Frequency
Corpus Freq (per M) Best form Share Zipf rank Forms
19th_books 92.98 996 100% 996 1
20th_books 117.09 822 100% 822 1
wiki_math 419.22 342 100% 342 1
wiki_geography 0.00 0
wiki_biology 0.00 0
wiki_modern_life 36.43 3758 100% 3758 1
wiki_arts 29.09 4117 100% 4117 1
wiki_society 25.62 4652 100% 4652 1
wiki_linguistics 0.00 0
wiki_physical_science 35.43 3440 100% 3440 1
wiki_history 0.00 0
early_modern_science 40.66 2546 100% 2546 1
religious_translated 65.57 1520 100% 1520 1
cooking 0.00 0
legal_scotus 11.66 5979 100% 5979 1

Best form = the single best-ranked form in the corpus, share-blind — how the spelling itself ranked. Share = how much of that spelling belongs to this lemma, when several senses compete for it. Zipf rank = all the forms combined (variant spellings included), each scaled by its share, and what feeds the combined frequency rank. Combining forms improves on the best form, but a share below 100% pushes the other way and usually wins: a sense holding a small slice of a common spelling is a rare word, so its rank lands downrated below the best form — and once it falls past what the corpus can support it is capped at the rank an unlisted word would get.