Jerome

Statistics

Precomputed corpus statistics with direct links back to text.
Settings
Not signed in

Statistics

Precomputed corpus statistics with direct links back to text.

Public statistics count ancient-language texts only: Greek and Latin. Modern translations stay visible elsewhere in Jerome but are excluded from counting.

Word forms are counted directly from the available text. Lemmas, parts of speech, and morphology appear only where CLTK parsing exists.

Ordinary statistics count one curated ancient representative per work. Legacy links with `version_scope=original` still load, but they resolve to the same counted universe.

Coverage

Overview

Distinctive vocabulary

Distinctive vocabulary is currently computed only for authors, and it always stays inside one ancient language at a time.

Distinctive vocabulary ranks terms used by the focus author. Terms absent from the focus are not included in this view.

Most frequent lemmas

These rows describe CLTK-covered text only. Exports return the current page only and remain limited by your access tier.

Most frequent word forms

Word forms are counted directly from available text. Statistics filters use exact or prefix matching only; use Search for full text or wildcard search.

Parts of speech

Morphology summary

Current Limitations

KWIC, distribution/trends, collocations, witness comparison, and all-versions statistics are not implemented in this surface.
Evidence links stay inside the counted statistics universe. Exports remain current-page only and do not imply bulk full-corpus export.