CoolFace
Datasetpublic

tartuNLP/pale-madlad-data

license: mit PaLe-MADLAD Data Data used for training the PaLe-MADLAD model to translate from Proper Karelian, Livvi, Ludian, and Veps to Russian and vice versa. Every dataset entry represents a single text and comes as a list of sentences supplemented (where possible) with a list of translations into Russian. Our sources include: VepKar: various articles, Biblical texts, folklore, and more in Proper Karelian, Livvi, Ludian, and Veps, mostly translated into… See the full description on the dataset page: https://huggingface.co/datasets/tartuNLP/pale-madlad-data.

sourceHugging Faceupdated 2y agoView on Hugging Face
0likes27downloads
10 commits on main
c7d766c2y ago

Update README.md

ozzwoy
16998b42y ago

Added dataset info

ozzwoy
6fe603e2y ago

Update README.md

LisaY
7beef302y ago

Update README.md

LisaY
8bd217e2y ago

Randomized data.

ozzwoy
89790582y ago

Fixed vepkar data.

ozzwoy
8610f4f2y ago

Swapped entries in README.

ozzwoy
63440a32y ago

Added dataset files.

ozzwoy
3c65fcc2y ago

Added config to readme.

ozzwoy
4047a3c2y ago

initial commit

ozzwoy