CoolFace
Datasetpublic

ds123DDD/opus-100

Dataset Card for OPUS-100 Dataset Summary OPUS-100 is an English-centric multilingual corpus covering 100 languages. OPUS-100 is English-centric, meaning that all training pairs include English on either the source or target side. The corpus covers 100 languages (including English). The languages were selected based on the volume of parallel data available in OPUS. Supported Tasks and Leaderboards Translation. Languages OPUS-100… See the full description on the dataset page: https://huggingface.co/datasets/ds123DDD/opus-100.

sourceHugging Faceunknownupdated 3mo agoView on Hugging Face
0likes164downloads
1 commits on main
ca815bd3mo ago

Duplicate from Helsinki-NLP/opus-100

ds123DDD, parquet-converter