CoolFace
Datasetpublic

ybracke/lexicon-dtak-transnormer-v1

Lexicon-DTAK-transnormer (v1.0) This dataset is derived from dtak-transnormer-full-v1, a parallel corpus of German texts from the period between 1600 to 1899, that aligns sentences in historical spelling with their normalizations. This dataset is a lexicon of ngram alignments between original and normalized ngrams observed in dtak-transnormer-full-v1 and their frequency. The ngram alignments in the lexicon are drawn from the sentence-level ngram alignments in… See the full description on the dataset page: https://huggingface.co/datasets/ybracke/lexicon-dtak-transnormer-v1.

sourceHugging Faceupdated 2y agoView on Hugging Face
0likes12downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
ybracke/lexicon-dtak-transnormer-v1 · CoolFace