shiima/vejin-Dataset-Normalization
Kurdish Books Dataset (Preprocessed) Dataset Description This dataset contains 18,565 Kurdish books with asosoft preprocessing applied to the content field. The dataset was created from an Excel file and includes book metadata along with preprocessed text content. Languages Central Kurdish (ckb) Kurdish (ku) Dataset Structure The dataset contains the following columns: author book title url content Data Processing… See the full description on the dataset page: https://huggingface.co/datasets/shiima/vejin-Dataset-Normalization.
03
No card is published for this repository, or it could not be fetched from Hugging Face right now.
