projecte-aina/ES-OC_Parallel_Corpus
Dataset Card for ES-OC Parallel Corpus Dataset Summary The ES-OC Parallel Corpus is a Spanish-Aranese dataset created to support the use of under-resourced languages from Spain, such as Aranese, in NLP tasks, specifically Machine Translation. Supported Tasks and Leaderboards The dataset can be used to train Bilingual Machine Translation models between Aranese and Spanish in any direction, as well as Multilingual Machine Translation models.… See the full description on the dataset page: https://huggingface.co/datasets/projecte-aina/ES-OC_Parallel_Corpus.
275
Update license to CC-BY-SA 4.0
Fix typo
Add license flag
add paper link
Update README.md
Create README.md
Upload dataset files
initial commit
