CoolFace
Datasetpublic

fdemelo/spelling-correction-french-news

Spelling correction dataset (French) This dataset is generated by transforming/corrupting sentences of a French news corpus provided by the University of Leipzig. The following transformations are applied to words in the sentences: concatenation of pairs of words swapping of neighboring letters in words insertion deletion replacement (by neighboring characters in AZERTY keyboard) Generation ./scripts/get_data.py -t news -y 2023 -s 10K… See the full description on the dataset page: https://huggingface.co/datasets/fdemelo/spelling-correction-french-news.

sourceHugging Facemitupdated 1y agoView on Hugging Face
1likes69downloads
settings

This repository belongs to fdemelo on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

namespelling-correction-french-news
visibilitypublic
licencemit
gatedno
ownerfdemelo
Account settings
fdemelo/spelling-correction-french-news · CoolFace