CoolFace
Datasetpublic

waashk/yelp_2013

Dataset used in the paper: A thorough benchmark of automatic text classification From traditional approaches to large language models https://github.com/waashk/atcBench To guarantee the reproducibility of the obtained results, the dataset and its respective CV train-test partitions is available here. Each dataset contains the following files: data.parquet: pandas DataFrame with texts and associated encoded labels for each document. split_<k>.pkl: pandas DataFrame with k-cross validation… See the full description on the dataset page: https://huggingface.co/datasets/waashk/yelp_2013.

sourceHugging Facemitupdated 1y agoView on Hugging Face
0likes77downloads
3 commits on main
20cd7fb1y ago

Upload folder using huggingface_hub

waashk
bf21b2c1y ago

Update README.md

waashk
fbb762d1y ago

initial commit

waashk