Base difficulty: Level 380 edit
Level distribution for noun (quantitative_concept): 20: 1 162: 24 276: 25 380: 25 390: 1 456: 1 1000: 1 1050: 2 1070: 2 1080: 10 1090: 1 1110: 1 1202: 2 1203: 1 1204: 2 1206: 1 1210: 2 1211: 1 1213: 7 1218: 1 1225: 2 1226: 1
No difficulty overrides set. The base difficulty level will be used for all languages.
Add Difficulty Override
python -m storage.admin --calc-ranks to populate)
Tier Signals
| Source | Tier | Rank used |
|---|---|---|
cambridge_yle |
Flyers (3 of 3) | 1200 |
cefr |
B1 (3 of 6) | 4200 |
Rank used = synthetic rank this tier contributes to the combined frequency rank (lower = more common).
Corpus Frequency
| Corpus | Freq (per M) | Best form | Share | Zipf rank | Forms |
|---|---|---|---|---|---|
19th_books |
13.43 | 3895 | 50% | 5138 downrated | 2 |
20th_books |
14.21 | 3683 | 50% | 4751 downrated | 2 |
wiki_math |
12.80 | 3054 | 50% | 6108 downrated | 1 |
wiki_geography |
14.54 | 3856 | 50% | 7712 downrated | 1 |
wiki_biology |
0.00 | — | — | — | 0 |
wiki_modern_life |
68.23 | 1546 | 50% | 2111 downrated | 2 |
wiki_arts |
86.74 | 894 | 50% | 1460 downrated | 2 |
wiki_society |
0.00 | — | — | — | 0 |
wiki_linguistics |
9.82 | 4316 | 50% | 8632 downrated | 1 |
wiki_physical_science |
0.00 | — | — | — | 0 |
wiki_history |
28.26 | 2259 | 50% | 4518 downrated | 1 |
early_modern_science |
16.11 | 5056 | 50% | 5056 | 2 |
religious_translated |
0.00 | — | — | — | 0 |
cooking |
27.19 | 1679 | 50% | 3358 downrated | 1 |
legal_scotus |
19.15 | 3605 | 50% | 4505 downrated | 2 |
eu_parliament_debates |
4.96 | 5181 | 50% | 10362 downrated | 1 |
openstax_science |
30.28 | 3392 | 50% | 3788 downrated | 2 |
openstax_society |
38.53 | 3105 | 50% | 3193 downrated | 2 |
wpa_life_histories |
8.00 | 6019 | 50% | 7123 downrated | 2 |
Best form = the single best-ranked form in the corpus, share-blind — how the spelling itself ranked. Share = how much of that spelling belongs to this lemma, when several senses compete for it. Zipf rank = all the forms combined (variant spellings included), each scaled by its share, and what feeds the combined frequency rank. Combining forms improves on the best form, but a share below 100% pushes the other way and usually wins: a sense holding a small slice of a common spelling is a rare word, so its rank lands downrated below the best form — and once it falls past what the corpus can support it is capped at the rank an unlisted word would get.