CoolFace
Datasetpublic

chcaa/dansk-ner

Dataset Summary DANSK: Danish Annotations for NLP Specific TasKs is a dataset consisting of texts from multiple domains, sampled from the Danish GigaWord Corpus (DAGW). The dataset was created to fill in the gap of Danish NLP datasets from different domains, that are required for training models that generalize across domains. The Named-Entity annotations are moreover fine-grained and have a similar form to that of OntoNotes v5, which significantly broadens the use cases of the… See the full description on the dataset page: https://huggingface.co/datasets/chcaa/dansk-ner.

sourceHugging Faceupdated 2y agoView on Hugging Face
3likes93downloads
settings

This repository belongs to chcaa on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

namedansk-ner
visibilitypublic
licencenot set
gatedno
ownerchcaa
Account settings
chcaa/dansk-ner · CoolFace