Explore vocabulary
Your project vocabulary is counted from what you already have, in the browser. Top terms ranks the vocabulary in the selected part of your project. Terms found near each other finds single terms that repeatedly share a nearby passage. The Pinned button beside Top terms opens the vocabulary you follow as the library grows, and explains how to start if you have no pins yet.Choose what to count
Start with Count in. All project material combines the three named content streams, Sources counts source titles and abstracts, Notes counts note titles and bodies, and Author keywords counts only the keywords supplied on source records. Keywords never silently stand in for a missing abstract, so switching scopes really does separate what you are reading, what you are writing, and how authors index their own work. The task view, scope, term length and ranking method stay in the URL, which means a reload or shared link restores the same lens. If a scope has nothing in it, it stays empty rather than borrowing results from All. Choose Sources vs notes at the top of the page for the focused comparison. It ranks vocabulary that leans toward source prose, vocabulary shared by both sides, and vocabulary that leans toward your notes under the same term-size and ranking lens. Source and note occurrence counts are labelled separately. The percentages are within-side shares of positive ranking scores, so a larger pile of source text does not win merely by being larger. Treat the result as a difference in vocabulary, not automatic evidence of novelty or contribution. Two more controls change how the active scope gets counted. Term length offers Single words, Two-word phrases, and Three-word phrases, which matters because “identity” and “identity formation” are different claims about your vocabulary. Rank by offers Most frequent · Count, Distinctive within documents · TF-IDF, and Unusually frequent · Log-likelihood. Count tells you what occurs most; the statistical methods surface concentration within documents or unexpected frequency against a baseline. Open How these counts work below the table for the exact method. Export CSV takes the current scoped table out.Read the ranked table
The columns answer different questions:- # is the term’s position under the current ranking method.
- Occurrences is the exact number of occurrences in the selected scope, even when TF-IDF or log-likelihood is doing the ranking.
- Matching items is the number of matching items. An item can be a source, a note or a source’s author-keyword record.
- In count mode, the bar accompanies Occurrences. Statistical rankings show a separate TF-IDF score or Signed G² score; on narrow screens this appears under the term. Bars are relative to the strongest score magnitude in this view.