CoolFace
Datasetpublic

ybracke/lexicon-dtak-transnormer-v1

Lexicon-DTAK-transnormer (v1.0) This dataset is derived from dtak-transnormer-full-v1, a parallel corpus of German texts from the period between 1600 to 1899, that aligns sentences in historical spelling with their normalizations. This dataset is a lexicon of ngram alignments between original and normalized ngrams observed in dtak-transnormer-full-v1 and their frequency. The ngram alignments in the lexicon are drawn from the sentence-level ngram alignments in… See the full description on the dataset page: https://huggingface.co/datasets/ybracke/lexicon-dtak-transnormer-v1.

sourceHugging Faceupdated 2y agoView on Hugging Face
0likes12downloads
settings

This repository belongs to ybracke on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

namelexicon-dtak-transnormer-v1
visibilitypublic
licencenot set
gatedno
ownerybracke
Account settings
ybracke/lexicon-dtak-transnormer-v1 · CoolFace