CoolFace
Datasetpublic

universityofbucharest/moroco

The MOROCO (Moldavian and Romanian Dialectal Corpus) dataset contains 33564 samples of text collected from the news domain. The samples belong to one of the following six topics: - culture - finance - politics - science - sports - tech

sourceHugging Facecc-by-4.0updated 3y agoView on Hugging Face
1likes166downloads
11 commits on main
d64d9b83y ago

Delete legacy JSON metadata (#3)

albertvillanova
c183e4f4y ago

Replace YAML keys from int to str (#2)

albertvillanova
9faec9c4y ago

Reorder split names (#1)

albertvillanova
52df16c4y ago

add dataset_info in dataset metadata

lhoestq
71c8b8f4y ago

remove dummmy data

mariosasko
c58f8134y ago

Add `language_bcp47` tag (#4753)

lhoestq
85dc5414y ago

Align more metadata with other repo types (models,spaces) (#4607)

julien-c
c5e78a24y ago

Remove config names as yaml keys (#4367)

lhoestq
4e8d6875y ago

Update files from the datasets library (from 1.16.0)

system
51d7d935y ago

Update files from the datasets library (from 1.7.0)

system
f49ef355y ago

Update files from the datasets library (from 1.6.0)

system