nawabhussain/Kashmiri-Language-Corpus
Kashmiri Textual Data Corpus Introduction This repository contains a combined dataset of Kashmiri textual data collected from various sources. The data has been sourced from different locations and may contain non-Kashmiri text (e.g., Urdu, Persian). The goal of this corpus is to provide a wide variety of Kashmiri text data for research and language processing tasks. Sources of Data 1. mzmmoazam/kashmiri_dataset (HTML Data) Source:… See the full description on the dataset page: https://huggingface.co/datasets/nawabhussain/Kashmiri-Language-Corpus.
2184
