Base Difficulty

Base difficulty: Level 230 edit

Level distribution for noun (nationality): 230: 17

Per-Language Difficulty Overrides

No difficulty overrides set. The base difficulty level will be used for all languages.

Add Difficulty Override
-1 to exclude, 1-1299 for level
Frequency and Tier Signals
Combined Frequency Rank: — (run python -m storage.admin --calc-ranks to populate)
Tier Signals

No lemma-tier assignment found yet.

Corpus Frequency
Corpus Freq (per M) Best form Share Zipf rank Forms
19th_books 6.28 8574 100% 8574 1
20th_books 16.45 5823 100% 3425 2
wiki_math 46.53 2089 100% 2089 1
wiki_geography 0.00 — — — 0
wiki_biology 0.00 — — — 0
wiki_modern_life 49.39 2915 100% 2915 1
wiki_arts 0.00 — — — 0
wiki_society 0.00 — — — 0
wiki_linguistics 0.00 — — — 0
wiki_physical_science 0.00 — — — 0
wiki_history 0.00 — — — 0
early_modern_science 128.12 869 100% 869 1
religious_translated 0.00 — — — 0
cooking 0.00 — — — 0
legal_scotus 0.00 — — — 0
eu_parliament_debates 0.00 — — — 0
openstax_science 62.64 2119 100% 2119 1
openstax_society 0.00 — — — 0
wpa_life_histories 23.09 3237 100% 3237 1

Best form = the single best-ranked form in the corpus, share-blind — how the spelling itself ranked. Share = how much of that spelling belongs to this lemma, when several senses compete for it. Zipf rank = all the forms combined (variant spellings included), each scaled by its share, and what feeds the combined frequency rank. Combining forms improves on the best form, but a share below 100% pushes the other way and usually wins: a sense holding a small slice of a common spelling is a rare word, so its rank lands downrated below the best form — and once it falls past what the corpus can support it is capped at the rank an unlisted word would get.