CoolFace
Datasetpublic

BEE-spoke-data/gutenberg-en-v1-clean

gutenberg - clean dataset_info: - config_name: default features: - name: text dtype: string - name: label dtype: string - name: score dtype: float64 - name: sha256 dtype: string - name: word_count dtype: int64 splits: - name: train num_bytes: 3384868097 num_examples: 9978 - name: validation num_bytes: 195405579 num_examples: 574 - name: test num_bytes: 189439446 num_examples: 565 download_size: 2317462261… See the full description on the dataset page: https://huggingface.co/datasets/BEE-spoke-data/gutenberg-en-v1-clean.

sourceHugging Faceodc-byupdated 9mo agoView on Hugging Face
4likes223downloads

BEE-spoke-data/gutenberg-en-v1-clean · main · files are served by the source, never re-hosted here