datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
The_Gab_Hate_Corpus_ghc_test_originalmammut-corpus-venezuela-test-set
mammut-corpus-venezuela
HuggingFace Dataset for testing purposes. The train dataset is mammut/mammut-corpus-venezuela.
1. How to use
How to load this dataset directly with the datasets library:
>>> from datasets import load_dataset>>> dataset = load_dataset("mammut/mammut-corpus-venezuela")
2. Dataset Summary
mammut-corpus-venezuela is a dataset for Spanish language modeling. This dataset comprises a large number of Venezuelan and Latin-American Spanish texts… See the full description on the dataset page: https://huggingface.co/datasets/mammut/mammut-corpus-venezuela-test-set.georgian-corpus-testThe_Gab_Hate_Corpus_ghc_test_translate
