ybracke/lexicon-dtak-transnormer-v1
Lexicon-DTAK-transnormer (v1.0) This dataset is derived from dtak-transnormer-full-v1, a parallel corpus of German texts from the period between 1600 to 1899, that aligns sentences in historical spelling with their normalizations. This dataset is a lexicon of ngram alignments between original and normalized ngrams observed in dtak-transnormer-full-v1 and their frequency. The ngram alignments in the lexicon are drawn from the sentence-level ngram alignments in… See the full description on the dataset page: https://huggingface.co/datasets/ybracke/lexicon-dtak-transnormer-v1.
012
Update README.md
Update README.md
Update README.md
Create README.md
Upload folder using huggingface_hub
initial commit
