datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
pseudepigrapha
NuBerea OCP Pseudepigrapha
Verse-level texts of 36 Old Testament pseudepigrapha from the Online Critical
Pseudepigrapha (OCP) project, spanning 6 languages: English, Greek, Aramaic,
Ethiopic, Syriac, and Latin. The corpus covers the major Second Temple
pseudepigraphal works — 1 Enoch, Jubilees, Psalms of Solomon, Sibylline
Oracles, Testaments of Job/Abraham/Solomon, 4 Ezra, 4 Maccabees, Joseph and
Aseneth, the Letter of Aristeas, and others — one row per verse (or section,
for… See the full description on the dataset page: https://huggingface.co/datasets/NuBerea/pseudepigrapha.pseudepigrapha-analysis
NuBerea Pseudepigrapha Analysis
Derived linguistic datasets over pseudepigraphal literature, part of the NuBerea curated corpus estate. Covers the Greek and Latin witnesses of these texts along with a multilingual view across the available witness languages.
License
CC BY 4.0.
Attribution
Source
Link
License
NuBerea project
https://huggingface.co/NuBerea
CC BY 4.0
