CoolFace
Datasetpublic

danish-foundation-models/danish-gigaword

Danish Gigaword Corpus Version: 1.0.0 License: See the respective dataset Dataset Summary The Danish Gigaword Corpus contains text spanning several domains and forms. This version does not include the sections containing tweets ("General Discussions" and "Parliament Elections"), "danavis", "Common Crawl" and "OpenSubtitles" due to potential privacy, quality and copyright concerns. Loading the dataset from datasets import load_dataset name =… See the full description on the dataset page: https://huggingface.co/datasets/danish-foundation-models/danish-gigaword.

sourceHugging Faceotherupdated 2y agoView on Hugging Face
9likes438downloads
settings

This repository belongs to danish-foundation-models on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

namedanish-gigaword
visibilitypublic
licenceother
gatedno
ownerdanish-foundation-models
Account settings
danish-foundation-models/danish-gigaword · CoolFace