CoolFace
Datasetpublic

projecte-aina/ES-AST_Parallel_Corpus

Dataset Card for ES-AST Parallel Corpus Dataset Summary The ES-AST Parallel Corpus is a Spanish-Asturian dataset created to support the use of under-resourced languages from Spain, such as Asturian, in NLP tasks, specifically Machine Translation. Supported Tasks and Leaderboards The dataset can be used to train Bilingual Machine Translation models between Asturian and Spanish in any direction, as well as Multilingual Machine Translation models.… See the full description on the dataset page: https://huggingface.co/datasets/projecte-aina/ES-AST_Parallel_Corpus.

sourceHugging Facecc-by-sa-4.0updated 1y agoView on Hugging Face
3likes54downloads
filees-ast_corpus.parquet156.9 MBdownload

projecte-aina/ES-AST_Parallel_Corpus · main · files are served by the source, never re-hosted here