datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
BnSentMix
BnSentMix: A Diverse Bengali-English Code-Mixed Dataset for Sentiment Analysis
Dataset Overview
Column Title
Description
Data Sources
Facebook, YouTube, E-commerce Sites
#Samples
20000
Sentiment Labels
1:Positive, 2:Negative, 3:Neutral, 4:Mixed
Filtering Method
Automated using mBERT
#Annotators
64
Annotation/Sample
2 or 3 (if tie)
Dataset Statistics
Statistic
Value
Mean Character Length
62.77
Max Character Length
1985… See the full description on the dataset page: https://huggingface.co/datasets/rifat101/BnSentMix.Muria-Churn-Prediction
