CoolFace
Datasetpublic

Helsinki-NLP/opus-100

Dataset Card for OPUS-100 Dataset Summary OPUS-100 is an English-centric multilingual corpus covering 100 languages. OPUS-100 is English-centric, meaning that all training pairs include English on either the source or target side. The corpus covers 100 languages (including English). The languages were selected based on the volume of parallel data available in OPUS. Supported Tasks and Leaderboards Translation. Languages OPUS-100… See the full description on the dataset page: https://huggingface.co/datasets/Helsinki-NLP/opus-100.

sourceHugging Faceunknownupdated 3y agoView on Hugging Face
244likes23kdownloads

Helsinki-NLP/opus-100 · main · files are served by the source, never re-hosted here