kasunUdayanga/Sinhala_Annotation_Dataset
Sinhala Named Entity Recognition (NER) Dataset - 85,000 Annotations Dataset Description This is a high-quality Named Entity Recognition (NER) dataset for the Sinhala language, consisting of approximately 85,000 annotations. The dataset was manually curated and annotated by a team of three students to support NLP research for low-resource languages. The data is sourced from diverse domains, including social media comments, news articles, and public domain texts… See the full description on the dataset page: https://huggingface.co/datasets/kasunUdayanga/Sinhala_Annotation_Dataset.
Update README.md
Update README.md
Update README.md
Update README.md
Upload Sinhala_NER.jsonl
Delete Sinhala_NER_Dataset.conll
Upload sinhala_ner_dataset.csv
Update README.md
Update README.md
add description
Upload Sinhala_NER_Dataset.conll
initial commit
