Base Difficulty

Base difficulty: Level 460 edit

Level distribution for noun (collection_things): 350: 2 405: 1 460: 13 1000: 5 1010: 1 1030: 1 1050: 3 1060: 1 1080: 1 1205: 3 1211: 1 1214: 1 1220: 3 1223: 2

Per-Language Difficulty Overrides

No difficulty overrides set. The base difficulty level will be used for all languages.

Add Difficulty Override
-1 to exclude, 1-1299 for level
Frequency and Tier Signals
Combined Frequency Rank: — (run python -m storage.admin --calc-ranks to populate)
Tier Signals
Source Tier Rank used
cefr B1 (3 of 6) 4200

Rank used = synthetic rank this tier contributes to the combined frequency rank (lower = more common).

Corpus Frequency
Corpus Freq (per M) Best form Share Zipf rank Forms
19th_books 0.00 — — — 0
20th_books 0.00 — — — 0
wiki_math 0.00 — — — 0
wiki_geography 72.16 1804 100% 1804 1
wiki_biology 0.00 — — — 0
wiki_modern_life 400.95 251 100% 251 1
wiki_arts 39.12 3232 100% 3232 1
wiki_society 61.75 2184 100% 2184 1
wiki_linguistics 0.00 — — — 0
wiki_physical_science 75.64 1864 100% 1864 1
wiki_history 45.65 2774 100% 2774 1
early_modern_science 0.00 — — — 0
religious_translated 0.00 — — — 0
cooking 0.00 — — — 0
legal_scotus 43.93 2542 100% 2542 1
eu_parliament_debates 59.60 1812 100% 1812 1
openstax_science 52.27 2458 100% 2458 1
openstax_society 56.46 2375 100% 2375 1
wpa_life_histories 39.33 2178 100% 2178 1

Best form = the single best-ranked form in the corpus, share-blind — how the spelling itself ranked. Share = how much of that spelling belongs to this lemma, when several senses compete for it. Zipf rank = all the forms combined (variant spellings included), each scaled by its share, and what feeds the combined frequency rank. Combining forms improves on the best form, but a share below 100% pushes the other way and usually wins: a sense holding a small slice of a common spelling is a rare word, so its rank lands downrated below the best form — and once it falls past what the corpus can support it is capped at the rank an unlisted word would get.