universityofbucharest/moroco
The MOROCO (Moldavian and Romanian Dialectal Corpus) dataset contains 33564 samples of text collected from the news domain. The samples belong to one of the following six topics: - culture - finance - politics - science - sports - tech
1166
Delete legacy JSON metadata (#3)
Replace YAML keys from int to str (#2)
Reorder split names (#1)
add dataset_info in dataset metadata
remove dummmy data
Add `language_bcp47` tag (#4753)
Align more metadata with other repo types (models,spaces) (#4607)
Remove config names as yaml keys (#4367)
Update files from the datasets library (from 1.16.0)
Update files from the datasets library (from 1.7.0)
Update files from the datasets library (from 1.6.0)
