datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
SIRENE_Clientdataset-analisis-sentimen-review-produk
📊 Dataset untuk Model indobert-analisis-sentimen-review-produk
Model ini dilatih untuk melakukan klasifikasi sentimen terhadap review produk dalam Bahasa Indonesia menggunakan IndoBERT sebagai base model.
📁 Jumlah Data
Dataset ini terdiri dari:
Kelas Sentimen
Jumlah Data
Positif
6.000
Negatif
4.589
🔧 Tahapan Preprocessing Data
Dataset telah melalui beberapa tahap preprocessing untuk memastikan kualitas data yang digunakan dalam pelatihan… See the full description on the dataset page: https://huggingface.co/datasets/siRendy/dataset-analisis-sentimen-review-produk.dataset-klasifikasi-sentimen-ulasan-produk1200_rows_dataset_siren_greetings_thanks_augmenteddataset-review-produk-untuk-pretrain-model-transformersSales-History-XGBdataset-sumarization-ulasan-produksirena_test1000_rows_dataset_siren_augmentedsocial-siren-synthetic
Social Siren Synthetic Multilingual Disaster-Tweet Dataset (v1)
5,000 synthetic, programmatically-generated tweets with multi-label annotations
for disaster detection, sarcasm, sentiment, and language, across English, Hindi,
Marathi, and code-mixed (Hinglish) text.
IMPORTANT — read before using in a paper. This is a synthetic dataset.
Every tweet was generated from templates, not collected from Twitter/X. It is
intended as a controlled testbed and demonstration fixture for… See the full description on the dataset page: https://huggingface.co/datasets/harrrrsh307/social-siren-synthetic.100_rows_dataset_siren_augmented
