Statistics
Precomputed corpus statistics with direct links back to text.
Public statistics count ancient-language texts only: Greek and Latin. Modern translations stay visible elsewhere in Jerome but are excluded from counting.
Word forms are counted directly from the available text. Lemmas, parts of speech, and morphology appear only where CLTK parsing exists.
Ordinary statistics count one curated ancient representative per work. Legacy links with `version_scope=original` still load, but they resolve to the same counted universe.
Coverage
Overview
Distinctive vocabulary
Distinctive vocabulary is currently computed only for authors, and it always stays inside one ancient language at a time.
Distinctive vocabulary ranks terms used by the focus author. Terms absent from the focus are not included in this view.
Most frequent lemmas
These rows describe CLTK-covered text only. Exports return the current page only and remain limited by your access tier.
Most frequent word forms
Word forms are counted directly from available text. Statistics filters use exact or prefix matching only; use Search for full text or wildcard search.