CoolFace
Datasetpublic

NHLOCAL/judaic-texts-corpus

Judaic Texts Corpus Dataset Summary Judaic Texts Corpus is a machine-readable Hebrew and Aramaic corpus of Judaic texts derived from the Otzaria library release archives. It is intended for language-model training, retrieval, search, digital humanities research, and other NLP workflows that need structured access to rabbinic and traditional Jewish texts. The current dataset build is produced from the official Otzaria/otzaria-library release assets, which package… See the full description on the dataset page: https://huggingface.co/datasets/NHLOCAL/judaic-texts-corpus.

sourceHugging Facecc-by-4.0updated 6d agoView on Hugging Face
1likes325downloads

NHLOCAL/judaic-texts-corpus · main · files are served by the source, never re-hosted here