CoolFace
Datasetpublic

NHLOCAL/project-ben-yehuda

Project Ben-Yehuda Corpus Dataset Summary This dataset contains Hebrew literary texts sourced from Project Ben-Yehuda and structured for NLP, language modeling, text analysis, historical linguistics, and computational Hebrew research. Each record represents one text/work and includes the full text, a fixed source label, and structured metadata from the Project Ben-Yehuda catalogue. Current Version Dataset version: pby-2026.03Source snapshot: Project… See the full description on the dataset page: https://huggingface.co/datasets/NHLOCAL/project-ben-yehuda.

sourceHugging Facecc0-1.0updated 5mo agoView on Hugging Face
2likes192downloads
settings

This repository belongs to NHLOCAL on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameproject-ben-yehuda
visibilitypublic
licencecc0-1.0
gatedno
ownerNHLOCAL
Account settings
NHLOCAL/project-ben-yehuda · CoolFace