codemurt/uyghur_ner_dataset
Uyghur NER dataset Description This dataset is in WikiAnn format. The dataset is assembled from named entities parsed from Wikipedia, Wiktionary and Dbpedia. For some words, new case forms have been created using Apertium-uig. Some locations have been translated using the Google Translate API. The dataset is divided into two parts: train and extra. Train has full sentences, extra has only named entities. Tags: O (0), B-PER (1), I-PER (2), B-ORG (3), I-ORG (4)… See the full description on the dataset page: https://huggingface.co/datasets/codemurt/uyghur_ner_dataset.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face