Base difficulty: Level 196 edit
Level distribution for noun (process_event): 19: 2 20: 1 196: 26 285: 26 289: 27 290: 32 330: 2 350: 2 360: 1 370: 1 380: 1 405: 9 476: 1 1000: 2 1010: 6 1020: 6 1030: 6 1040: 12 1050: 18 1060: 5 1070: 12 1080: 2 1110: 2 1120: 7 1130: 7 1202: 1 1204: 1 1210: 1 1211: 1 1216: 1
No difficulty overrides set. The base difficulty level will be used for all languages.
Add Difficulty Override
python -m storage.admin --calc-ranks to populate)
Tier Signals
| Source | Tier | Rank used |
|---|---|---|
basic_english |
Basic (1 of 2) | 600 |
cambridge_yle |
Starters (1 of 3) | 325 |
cefr |
A1 (1 of 6) | 800 |
Rank used = synthetic rank this tier contributes to the combined frequency rank (lower = more common).
Corpus Frequency
| Corpus | Freq (per M) | Best form | Share | Zipf rank | Forms |
|---|---|---|---|---|---|
19th_books |
339.65 | 258 | 80% | 322 downrated | 1 |
20th_books |
363.58 | 256 | 80% | 320 downrated | 1 |
wiki_math |
160.06 | 668 | 80% | 835 downrated | 1 |
wiki_geography |
424.18 | 195 | 80% | 244 downrated | 1 |
wiki_biology |
272.58 | 339 | 80% | 424 downrated | 1 |
wiki_modern_life |
470.14 | 163 | 80% | 204 downrated | 1 |
wiki_arts |
461.00 | 174 | 80% | 218 downrated | 1 |
wiki_society |
314.17 | 270 | 80% | 337 downrated | 1 |
wiki_linguistics |
416.33 | 231 | 80% | 289 downrated | 1 |
wiki_physical_science |
269.67 | 401 | 80% | 501 downrated | 1 |
wiki_history |
495.57 | 149 | 80% | 186 downrated | 1 |
early_modern_science |
368.57 | 224 | 80% | 280 downrated | 1 |
religious_translated |
459.67 | 215 | 80% | 269 downrated | 1 |
cooking |
200.32 | 557 | 80% | 696 downrated | 1 |
legal_scotus |
140.90 | 736 | 80% | 920 downrated | 1 |
eu_parliament_debates |
444.49 | 230 | 80% | 288 downrated | 1 |
openstax_science |
366.36 | 236 | 80% | 295 downrated | 1 |
openstax_society |
402.52 | 228 | 80% | 285 downrated | 1 |
wpa_life_histories |
1059.07 | 109 | 80% | 136 downrated | 1 |
Best form = the single best-ranked form in the corpus, share-blind — how the spelling itself ranked. Share = how much of that spelling belongs to this lemma, when several senses compete for it. Zipf rank = all the forms combined (variant spellings included), each scaled by its share, and what feeds the combined frequency rank. Combining forms improves on the best form, but a share below 100% pushes the other way and usually wins: a sense holding a small slice of a common spelling is a rare word, so its rank lands downrated below the best form — and once it falls past what the corpus can support it is capped at the rank an unlisted word would get.