michsethowusu/pedi-sentiments-corpus
Pedi Sentiment Corpus Dataset Description This dataset contains sentiment-labeled text data in Pedi for binary sentiment classification (Positive/Negative). Sentiments are extracted and processed from the English meanings of the sentences using DistilBERT for sentiment classification. The dataset is part of a larger collection of African language sentiment analysis resources. Dataset Statistics Total samples: 422,975 Positive sentiment: 255703… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/pedi-sentiments-corpus.
Add dataset README for sentiment corpus
Upload README.md with huggingface_hub
Upload data/train-00000-of-00001-9dffadee04db7bfc.parquet with huggingface_hub
initial commit
