Helsinki-NLP/opus-100
Dataset Card for OPUS-100 Dataset Summary OPUS-100 is an English-centric multilingual corpus covering 100 languages. OPUS-100 is English-centric, meaning that all training pairs include English on either the source or target side. The corpus covers 100 languages (including English). The languages were selected based on the volume of parallel data available in OPUS. Supported Tasks and Leaderboards Translation. Languages OPUS-100… See the full description on the dataset page: https://huggingface.co/datasets/Helsinki-NLP/opus-100.
Update dataset card (#6)
Convert dataset to Parquet (#5)
Delete legacy JSON metadata (#4)
rename configs to config_name
Remove tasks other than translation (#2)
Add translation task (#1)
add dataset_info in dataset metadata
remove dummmy data
Fix missing tags in dataset cards (#4891)
Fix titles in dataset cards (#4824)
Add `language_bcp47` tag (#4753)
Align more metadata with other repo types (models,spaces) (#4607)
Remove config names as yaml keys (#4367)
Update datasets task tags to align tags with models (#4067)
Update files from the datasets library (from 1.16.0)
Update files from the datasets library (from 1.7.0)
Update files from the datasets library (from 1.6.1)
Update files from the datasets library (from 1.6.0)
Update files from the datasets library (from 1.3.0)
Update files from the datasets library (from 1.2.0)
