datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
turkish-sentiment-analysis-dataset
Dataset
This dataset contains positive , negative and notr sentences from several data sources given in the references. In the most sentiment models , there are only two labels; positive and negative. However , user input can be totally notr sentence. For such cases there were no data I could find. Therefore I created this dataset with 3 class. Positive and negative sentences are listed below. Notr examples are extraced from turkish wiki dump. In addition, added some random text… See the full description on the dataset page: https://huggingface.co/datasets/winvoker/turkish-sentiment-analysis-dataset.Turkish_SentimentAnalysis_TRSAv1TRSAv1 (Turkish Sentiment Analysis Version 1) Dataset
This data set has been produced to contribute to Turkish NLP studies.
The dataset consists of a total of 150 thousand samples, 50 thousand negative, 50 thousand positive, and 50 thousand neutral.
It can be used in text classification and sentiment analysis studies by citing the related study.
Related Work
Aydoğan M, Kocaman V. TRSAv1: A new benchmark dataset for classifying user reviews on Turkish e-commerce websites. Journal of… See the full description on the dataset page: https://huggingface.co/datasets/maydogan/Turkish_SentimentAnalysis_TRSAv1.alihanuludag_turkish-universities-sentiment-analysis-dataset
Turkish Universities Sentiment Analysis Dataset
Sentiment Labels for Turkish University-related Tweets from X (Twitter)
Dataset Info
Source: Kaggle
Original Size: 4.78 MB
Kaggle Downloads: 17
Files: 5
Files
README.md.txt
real-synthetic_dataset.csv
real-synthetic_dataset.xlsx
real_dataset.csv
real_dataset.xlsx
Mirrored from Kaggle
keloglan-turkish-sentiment-analysis-dataset
🏰 Keloğlan Turkish Sentiment Analysis Dataset
This dataset represents one of the largest and most comprehensive sentiment analysis collections for the Turkish language, containing over 630,000 unique samples.
It was meticulously constructed by merging, cleaning, and deduplicating several major open-source Turkish sentiment datasets to train the Keloğlan model series.
📊 Dataset Statistics
Total Unique Samples: 631,166
Training Set: 599,607
Validation Set: 31,559
Label… See the full description on the dataset page: https://huggingface.co/datasets/engin1123/keloglan-turkish-sentiment-analysis-dataset.turkish-sentiment-analysis-dataset
Dataset
This dataset contains positive , negative and notr sentences from several data sources given in the references. In the most sentiment models , there are only two labels; positive and negative. However , user input can be totally notr sentence. For such cases there were no data I could find. Therefore I created this dataset with 3 class. Positive and negative sentences are listed below. Notr examples are extraced from turkish wiki dump. In addition, added some random text… See the full description on the dataset page: https://huggingface.co/datasets/BatuhanBM/turkish-sentiment-analysis-dataset.turkish-sentiment-analysis-miniturkish-sentiment-analysis-dataset_ENG
