Base difficulty: Level 162 edit
Level distribution for noun (quantitative_concept): 20: 1 162: 24 276: 25 380: 25 390: 1 456: 1 1000: 1 1050: 2 1070: 2 1080: 10 1090: 1 1110: 1 1202: 2 1203: 1 1204: 2 1206: 1 1210: 2 1211: 1 1213: 7 1218: 1 1225: 2 1226: 1
No difficulty overrides set. The base difficulty level will be used for all languages.
Add Difficulty Override
python -m storage.admin --calc-ranks to populate)
Tier Signals
| Source | Tier | Rank used |
|---|---|---|
cefr |
A2 (2 of 6) | 2100 |
Rank used = synthetic rank this tier contributes to the combined frequency rank (lower = more common).
Corpus Frequency
| Corpus | Freq (per M) | Best form | Share | Zipf rank | Forms |
|---|---|---|---|---|---|
19th_books |
61.19 | 1464 | 100% | 1464 | 1 |
20th_books |
36.54 | 2205 | 100% | 2205 | 1 |
wiki_math |
72.12 | 1532 | 100% | 1532 | 1 |
wiki_geography |
43.78 | 2766 | 100% | 2766 | 1 |
wiki_biology |
101.02 | 1382 | 100% | 1382 | 1 |
wiki_modern_life |
440.37 | 222 | 100% | 222 | 1 |
wiki_arts |
60.55 | 2190 | 100% | 2190 | 1 |
wiki_society |
28.05 | 4379 | 100% | 4379 | 1 |
wiki_linguistics |
43.79 | 2476 | 100% | 2476 | 1 |
wiki_physical_science |
523.92 | 223 | 100% | 223 | 1 |
wiki_history |
36.52 | 3353 | 100% | 3353 | 1 |
early_modern_science |
95.13 | 1206 | 100% | 1206 | 1 |
religious_translated |
47.65 | 1991 | 100% | 1991 | 1 |
cooking |
0.00 | — | — | — | 0 |
legal_scotus |
33.73 | 3078 | 100% | 3078 | 1 |
eu_parliament_debates |
60.03 | 1804 | 100% | 1804 | 1 |
openstax_science |
412.32 | 290 | 100% | 290 | 1 |
openstax_society |
44.09 | 2883 | 100% | 2883 | 1 |
wpa_life_histories |
48.24 | 1861 | 100% | 1861 | 1 |
Best form = the single best-ranked form in the corpus, share-blind — how the spelling itself ranked. Share = how much of that spelling belongs to this lemma, when several senses compete for it. Zipf rank = all the forms combined (variant spellings included), each scaled by its share, and what feeds the combined frequency rank. Combining forms improves on the best form, but a share below 100% pushes the other way and usually wins: a sense holding a small slice of a common spelling is a rare word, so its rank lands downrated below the best form — and once it falls past what the corpus can support it is capped at the rank an unlisted word would get.