CoolFace
Datasetpublic

projecte-aina/ES-OC_Parallel_Corpus

Dataset Card for ES-OC Parallel Corpus Dataset Summary The ES-OC Parallel Corpus is a Spanish-Aranese dataset created to support the use of under-resourced languages from Spain, such as Aranese, in NLP tasks, specifically Machine Translation. Supported Tasks and Leaderboards The dataset can be used to train Bilingual Machine Translation models between Aranese and Spanish in any direction, as well as Multilingual Machine Translation models.… See the full description on the dataset page: https://huggingface.co/datasets/projecte-aina/ES-OC_Parallel_Corpus.

sourceHugging Facecc-by-sa-4.0updated 1y agoView on Hugging Face
2likes75downloads
8 commits on main
e7189561y ago

Update license to CC-BY-SA 4.0

fdelucaf
e113b0a1y ago

Fix typo

fdelucaf
f361d9e2y ago

Add license flag

fdelucaf
9ae86062y ago

add paper link

fdelucaf
1cda7362y ago

Update README.md

mmarimon
978b7fc2y ago

Create README.md

fdelucaf
352deed2y ago

Upload dataset files

fdelucaf
4d100662y ago

initial commit

fdelucaf