CoolFace
Datasetpublic

TigreGotico/desacordo_ortografico

Portuguese Orthographies — Parallel Corpus A parallel corpus for detecting and converting between Portuguese orthographies. Each record is one Portuguese sentence written in five orthographic norms, so the same content can be aligned across the spelling reforms of the language. Norms (one column each) column norm etymological pre-1911 pseudo-etymological spelling pt_1973 pre-AO1990 European (Convenção 1945 + 1973 mini-reform) ao1990_pt Acordo… See the full description on the dataset page: https://huggingface.co/datasets/TigreGotico/desacordo_ortografico.

sourceHugging Faceapache-2.0updated 3mo agoView on Hugging Face
0likes28downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face