CoolFace
Datasetpublic

kasunUdayanga/Sinhala_Annotation_Dataset

Sinhala Named Entity Recognition (NER) Dataset - 85,000 Annotations Dataset Description This is a high-quality Named Entity Recognition (NER) dataset for the Sinhala language, consisting of approximately 85,000 annotations. The dataset was manually curated and annotated by a team of three students to support NLP research for low-resource languages. The data is sourced from diverse domains, including social media comments, news articles, and public domain texts… See the full description on the dataset page: https://huggingface.co/datasets/kasunUdayanga/Sinhala_Annotation_Dataset.

sourceHugging Facecc-by-4.0updated 9mo agoView on Hugging Face
4likes40downloads
12 commits on main
8afb4a49mo ago

Update README.md

kasunUdayanga
d586a859mo ago

Update README.md

kasunUdayanga
462ba959mo ago

Update README.md

kasunUdayanga
e4bf1d09mo ago

Update README.md

kasunUdayanga
0ca2fe79mo ago

Upload Sinhala_NER.jsonl

kasunUdayanga
f52b44a9mo ago

Delete Sinhala_NER_Dataset.conll

kasunUdayanga
34015639mo ago

Upload sinhala_ner_dataset.csv

kasunUdayanga
0c1df179mo ago

Update README.md

kasunUdayanga
1c144829mo ago

Update README.md

kasunUdayanga
0a34be69mo ago

add description

kasunUdayanga
0fad9859mo ago

Upload Sinhala_NER_Dataset.conll

kasunUdayanga
abd110e9mo ago

initial commit

kasunUdayanga