Helsinki-NLP/europarl
Dataset Card for OPUS Europarl (European Parliament Proceedings Parallel Corpus) Dataset Summary A parallel corpus extracted from the European Parliament web site by Philipp Koehn (University of Edinburgh). The main intended use is to aid statistical machine translation research. More information can be found at http://www.statmt.org/europarl/ Supported Tasks and Leaderboards Tasks: Machine Translation, Cross Lingual Word Embeddings (CWLE) Alignment… See the full description on the dataset page: https://huggingface.co/datasets/Helsinki-NLP/europarl.
Update dataset card (#9)
Convert dataset to Parquet (#8)
Fix bug in language pairs (#7)
Add all language pairs (#6)
Add paperswithcode ID and pretty name (#5)
Fix citation issue in loading script (#4)
Update metadata (#3)
Update base url (#2)
Delete legacy JSON metadata (#1)
add dataset_info in dataset metadata
remove dummmy data
Align more metadata with other repo types (models,spaces) (#4607)
Update datasets task tags to align tags with models (#4067)
Update files from the datasets library (from 1.18.0)
Update files from the datasets library (from 1.7.0)
Update files from the datasets library (from 1.6.1)
Update files from the datasets library (from 1.6.0)
Update files from the datasets library (from 1.5.0)
