Base difficulty: Level 406 edit
Level distribution for noun (small_movable_object): 2: 13 19: 5 20: 4 27: 1 29: 12 194: 16 195: 29 199: 29 200: 27 405: 4 406: 8 416: 1 1000: 1 1040: 8 1090: 2 1201: 3 1202: 7 1205: 8 1206: 3 1208: 4 1209: 1 1210: 1 1211: 3 1214: 1 1215: 4 1216: 2 1217: 6 1218: 1 1219: 1 1220: 3 1221: 2 1223: 3 1225: 1 1226: 1
No difficulty overrides set. The base difficulty level will be used for all languages.
Add Difficulty Override
python -m storage.admin --calc-ranks to populate)
Tier Signals
| Source | Tier | Rank used |
|---|---|---|
cambridge_yle |
Flyers (3 of 3) | 1200 |
cefr |
A1 (1 of 6) | 800 |
Rank used = synthetic rank this tier contributes to the combined frequency rank (lower = more common).
Corpus Frequency
| Corpus | Freq (per M) | Best form | Share | Zipf rank | Forms |
|---|---|---|---|---|---|
19th_books |
6.73 | 648 | 3% | 13948 downrated | 2 |
20th_books |
5.69 | 773 | 3% | 16653 downrated | 2 |
wiki_math |
3.39 | 1497 | 3% | 8000 capped | 2 |
wiki_geography |
1.00 | 3407 | 3% | 9000 capped | 1 |
wiki_biology |
1.62 | 2431 | 3% | 9000 capped | 1 |
wiki_modern_life |
12.37 | 577 | 3% | 9931 downrated | 2 |
wiki_arts |
14.38 | 448 | 3% | 7867 downrated | 2 |
wiki_society |
1.88 | 3629 | 3% | 11000 capped | 2 |
wiki_linguistics |
1.46 | 3185 | 3% | 9000 capped | 2 |
wiki_physical_science |
2.23 | 3254 | 3% | 8000 capped | 2 |
wiki_history |
5.24 | 1371 | 3% | 9000 capped | 2 |
early_modern_science |
11.75 | 374 | 3% | 9243 downrated | 2 |
religious_translated |
8.00 | 697 | 3% | 10400 capped | 2 |
cooking |
54.78 | 132 | 3% | 2954 downrated | 2 |
legal_scotus |
0.66 | 5334 | 3% | 20000 capped | 2 |
eu_parliament_debates |
1.64 | 2144 | 3% | 11000 capped | 2 |
openstax_science |
3.36 | 1970 | 3% | 11000 capped | 2 |
openstax_society |
3.07 | 2149 | 3% | 11000 capped | 2 |
wpa_life_histories |
8.06 | 557 | 3% | 12843 downrated | 2 |
Best form = the single best-ranked form in the corpus, share-blind — how the spelling itself ranked. Share = how much of that spelling belongs to this lemma, when several senses compete for it. Zipf rank = all the forms combined (variant spellings included), each scaled by its share, and what feeds the combined frequency rank. Combining forms improves on the best form, but a share below 100% pushes the other way and usually wins: a sense holding a small slice of a common spelling is a rare word, so its rank lands downrated below the best form — and once it falls past what the corpus can support it is capped at the rank an unlisted word would get.