CoolFace
Datasetpublic

chcaa/dansk-ner

Dataset Summary DANSK: Danish Annotations for NLP Specific TasKs is a dataset consisting of texts from multiple domains, sampled from the Danish GigaWord Corpus (DAGW). The dataset was created to fill in the gap of Danish NLP datasets from different domains, that are required for training models that generalize across domains. The Named-Entity annotations are moreover fine-grained and have a similar form to that of OntoNotes v5, which significantly broadens the use cases of the… See the full description on the dataset page: https://huggingface.co/datasets/chcaa/dansk-ner.

sourceHugging Faceupdated 2y agoView on Hugging Face
3likes93downloads
7 commits on main
8ee5bdd2y ago

Upload dataset

KennethEnevoldsen
7725dc33y ago

Update README.md

KennethEnevoldsen
43b509e3y ago

Update README.md

KennethEnevoldsen
6b24a233y ago

Add metadata tags (#2)

KennethEnevoldsen, saattrupdan
86656a33y ago

Update README.md

KennethEnevoldsen
1a7ce163y ago

Upload dataset

KennethEnevoldsen
a6880d63y ago

initial commit

KennethEnevoldsen