Word Tokens
A word token is one spelling in one language — “top”, “will”, “ice cream” — regardless of which sense means it. Frequency is measured here, then split across the senses that share the string.
Top unlinked tokens
Frequent spellings that no lemma claims yet — the queue of vocabulary the corpora measured but the database has no sense for.
Most common unlinkedCorpus vocabulary
Words that lean toward one corpus more than the rest — the cooking words, the science words.
19th_books
20th_books
wiki_vital
wiki_math
wiki_geography
wiki_biology
wiki_modern_life
early_modern_science
religious_translated
cooking
legal_scotus
Register-neutral words
The complement: words that sit at the same frequency in every corpus.
Steadiest across corpora