CoolFace
Datasetpublicgated

shiima/vejin-Dataset-Normalization

Kurdish Books Dataset (Preprocessed) Dataset Description This dataset contains 18,565 Kurdish books with asosoft preprocessing applied to the content field. The dataset was created from an Excel file and includes book metadata along with preprocessed text content. Languages Central Kurdish (ckb) Kurdish (ku) Dataset Structure The dataset contains the following columns: author book title url content Data Processing… See the full description on the dataset page: https://huggingface.co/datasets/shiima/vejin-Dataset-Normalization.

sourceHugging Faceapache-2.0updated 8mo agoView on Hugging Face
0likes3downloads

shiima/vejin-Dataset-Normalization · main · files are served by the source, never re-hosted here

This repository is gated. The listing is public, but downloading a file means accepting the publisher’s terms at Hugging Face first — the links above take you there rather than around it.