CoolFace
Datasetpublic

ybracke/lexicon-dtak-transnormer-v1

Lexicon-DTAK-transnormer (v1.0) This dataset is derived from dtak-transnormer-full-v1, a parallel corpus of German texts from the period between 1600 to 1899, that aligns sentences in historical spelling with their normalizations. This dataset is a lexicon of ngram alignments between original and normalized ngrams observed in dtak-transnormer-full-v1 and their frequency. The ngram alignments in the lexicon are drawn from the sentence-level ngram alignments in… See the full description on the dataset page: https://huggingface.co/datasets/ybracke/lexicon-dtak-transnormer-v1.

sourceHugging Faceupdated 2y agoView on Hugging Face
0likes12downloads
6 commits on main
1ac54b82y ago

Update README.md

ybracke
eda21f72y ago

Update README.md

ybracke
29b605b2y ago

Update README.md

ybracke
844b1be2y ago

Create README.md

ybracke
33cd98a2y ago

Upload folder using huggingface_hub

ybracke
319fbdf2y ago

initial commit

ybracke