CoolFace
Datasetpublicgated

shiima/vejin-Dataset-Normalization

Kurdish Books Dataset (Preprocessed) Dataset Description This dataset contains 18,565 Kurdish books with asosoft preprocessing applied to the content field. The dataset was created from an Excel file and includes book metadata along with preprocessed text content. Languages Central Kurdish (ckb) Kurdish (ku) Dataset Structure The dataset contains the following columns: author book title url content Data Processing… See the full description on the dataset page: https://huggingface.co/datasets/shiima/vejin-Dataset-Normalization.

sourceHugging Faceapache-2.0updated 8mo agoView on Hugging Face
0likes3downloads
Dataset Card

No card is published for this repository, or it could not be fetched from Hugging Face right now.