CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01winvoker /turkish-sentiment-analysis-dataset Dataset This dataset contains positive , negative and notr sentences from several data sources given in the references. In the most sentiment models , there are only two labels; positive and negative. However , user input can be totally notr sentence. For such cases there were no data I could find. Therefore I created this dataset with 3 class. Positive and negative sentences are listed below. Notr examples are extraced from turkish wiki dump. In addition, added some random text… See the full description on the dataset page: https://huggingface.co/datasets/winvoker/turkish-sentiment-analysis-dataset.texttext-classification100K<n<1M49 likes699 downloads3y agoHugging Face02hugginglearners /amazon-reviews-sentiment-analysis Dataset Card for amazon reviews for sentiment analysis Dataset Summary One of the most important problems in e-commerce is the correct calculation of the points given to after-sales products. The solution to this problem is to provide greater customer satisfaction for the e-commerce site, product prominence for sellers, and a seamless shopping experience for buyers. Another problem is the correct ordering of the comments given to the products. The prominence of misleading… See the full description on the dataset page: https://huggingface.co/datasets/hugginglearners/amazon-reviews-sentiment-analysis.tabular1K<n<10K5 likes604 downloads4y agoHugging Face03ParsiAI /snappfood-sentiment-analysistexttext-classification10K<n<100K7 likes538 downloads2y agoHugging Face04ParsiAI /digikala-sentiment-analysistabulartext-classification1K<n<10K3 likes525 downloads2y agoHugging Face05Sp1786 /multiclass-sentiment-analysis-dataset Dataset Card for Dataset Name Dataset Summary This dataset card aims to be a base template for new datasets. It has been generated using this raw template. Supported Tasks and Leaderboards [More Information Needed] Languages [More Information Needed] Dataset Structure Data Instances [More Information Needed] Data Fields [More Information Needed] Data Splits [More Information Needed] Dataset Creation… See the full description on the dataset page: https://huggingface.co/datasets/Sp1786/multiclass-sentiment-analysis-dataset.tabulartext-classification10K<n<100K29 likes434 downloads3y agoHugging Face06prasadsawant7 /sentiment_analysis_preprocessed_datasetBrief idea about dataset: This dataset is designed for a Text Classification to be specific Multi Class Classification, inorder to train a model (Supervised Learning) for Sentiment Analysis. Also to be able retrain the model on the given feedback over a wrong predicted sentiment this dataset will help to manage those things using Other Features. Main Features text labels This feature variable has all sort of texts, sentences, tweets, etc. This target variable contains 3 types of… See the full description on the dataset page: https://huggingface.co/datasets/prasadsawant7/sentiment_analysis_preprocessed_dataset.tabulartext-classification100K<n<1M4 likes248 downloads3y agoHugging Face07yassiracharki /Amazon_Reviews_Binary_for_Sentiment_Analysis Dataset Card for Dataset Name The Amazon reviews polarity dataset is constructed by taking review score 1 and 2 as negative, and 4 and 5 as positive. Samples of score 3 is ignored. In the dataset, class 1 is the negative and class 2 is the positive. Each class has 1,800,000 training samples and 200,000 testing samples. Dataset Details Dataset Description The files train.csv and test.csv contain all the training samples as comma-sparated values. There are 3… See the full description on the dataset page: https://huggingface.co/datasets/yassiracharki/Amazon_Reviews_Binary_for_Sentiment_Analysis.texttext-classification1M<n<10M0 likes150 downloads2y agoHugging Face08maydogan /Turkish_SentimentAnalysis_TRSAv1TRSAv1 (Turkish Sentiment Analysis Version 1) Dataset This data set has been produced to contribute to Turkish NLP studies. The dataset consists of a total of 150 thousand samples, 50 thousand negative, 50 thousand positive, and 50 thousand neutral. It can be used in text classification and sentiment analysis studies by citing the related study. Related Work Aydoğan M, Kocaman V. TRSAv1: A new benchmark dataset for classifying user reviews on Turkish e-commerce websites. Journal of… See the full description on the dataset page: https://huggingface.co/datasets/maydogan/Turkish_SentimentAnalysis_TRSAv1.texttext-classification100K<n<1M15 likes144 downloads2y agoHugging Face09yassiracharki /Amazon_Reviews_for_Sentiment_Analysis_fine_grained_5_classes Dataset Card for Dataset Name The Amazon reviews full score dataset is constructed by randomly taking 600,000 training samples and 130,000 testing samples for each review score from 1 to 5. In total there are 3,000,000 trainig samples and 650,000 testing samples. Dataset Details Dataset Description The files train.csv and test.csv contain all the training samples as comma-sparated values. There are 3 columns in them, corresponding to class index (1 to 5)… See the full description on the dataset page: https://huggingface.co/datasets/yassiracharki/Amazon_Reviews_for_Sentiment_Analysis_fine_grained_5_classes.texttext-classification1M<n<10M4 likes144 downloads2y agoHugging Face10MaNaN-3 /twitter_sentiment_analysistext100K<n<1M1 likes107 downloads3y agoHugging Face11sinhala-nlp /sinhala-sentiment-analysistext1K<n<10K0 likes71 downloads2y agoHugging Face12xcz0 /Aspect-Based_Sentiment_Analysis_for_Catering 说明 数据集来源于AI Challenger 2018 sentiment_analysis_trainingset.csv 为训练集数据文件,共105000条评论数据 sentiment_analysis_validationset.csv 为验证集数据文件,共15000条评论数据 sentiment_analysis_testa.csv 为测试集A数据文件,共15000条评论数据 数据集分为训练、验证、测试A与测试B四部分。数据集中的评价对象按照粒度不同划分为两个层次,层次一为粗粒度的评价对象,例如评论文本中涉及的服务、位置等要素;层次二为细粒度的情感对象,例如“服务”属性中的“服务人员态度”、“排队等候时间”等细粒度要素。评价对象的具体划分如下表所示。 The dataset is divided into four parts: training, validation, test A and test B. This dataset builds a two-layer labeling system according to the… See the full description on the dataset page: https://huggingface.co/datasets/xcz0/Aspect-Based_Sentiment_Analysis_for_Catering.tabulartext-classification100K<n<1M0 likes64 downloads3y agoHugging Face13agentlans /twitter-sentiment-meta-analysis Twitter Sentiment Meta-Analysis Dataset Dataset Description This dataset contains sentiment analysis results for English tweets collected between September 2009 and January 2010. The tweets were processed and analyzed using 10 different sentiment classifiers, with the final sentiment score derived from principal component analysis (PCA). Source Data Original Data: Cheng-Caverlee-Lee Twitter Scrape (Sept 2009 - Jan 2010) Number of Tweets: 138 690 Language:… See the full description on the dataset page: https://huggingface.co/datasets/agentlans/twitter-sentiment-meta-analysis.tabulartext-classification10K<n<100K1 likes61 downloads2y agoHugging Face14md-nishat-008 /Code-Mixed-Sentiment-Analysis-Dataset Dataset Generation: Initially, we select the Amazon Review Dataset as our base data, referenced from Ni et al. (2019)[^1]. We randomly extract 100,000 instances from this dataset. The original labels in this dataset are ratings, scaled from 1 to 5. For our specific task, we categorize them into Positive (rating > 3), Neutral (rating = 3), and Negative (rating < 3), ensuring a balanced number of instances for each label. To generate the synthetic Code-mixed dataset, we apply two… See the full description on the dataset page: https://huggingface.co/datasets/md-nishat-008/Code-Mixed-Sentiment-Analysis-Dataset.text10K<n<100K0 likes59 downloads3y agoHugging Face15hafsteinn /faroese_sentiment_analysis Good or Bad News? Exploring GPT-4 for Sentiment Analysis on Faroese News Corpora This dataset is a part of the research from the paper "Good or Bad News? Exploring GPT-4 for Sentiment Analysis for Faroese on a Public News Corpora," that focuses on the application of GPT-4 for sentiment analysis on Faroese news texts. The study addresses the challenges of sentiment analysis in low-resource languages and evaluates the effectiveness of Large Language Models, specifically GPT-4, in… See the full description on the dataset page: https://huggingface.co/datasets/hafsteinn/faroese_sentiment_analysis.tabulartext-classificationn<1K1 likes58 downloads3y agoHugging Face16yassiracharki /Yelp_Reviews_for_Sentiment_Analysis_fine_grained_5_classes Dataset Card for Dataset Name The Yelp reviews full star dataset is constructed by randomly taking 130,000 training samples and 10,000 testing samples for each review star from 1 to 5. In total there are 650,000 trainig samples and 50,000 testing samples. Dataset Description The files train.csv and test.csv contain all the training samples as comma-sparated values. There are 2 columns in them, corresponding to class index (1 to 5) and review text. The review texts are… See the full description on the dataset page: https://huggingface.co/datasets/yassiracharki/Yelp_Reviews_for_Sentiment_Analysis_fine_grained_5_classes.texttext-classification100K<n<1M0 likes57 downloads2y agoHugging Face17moali-mkh-2000 /PersianTwitterDataset-SentimentAnalysis Dataset Card for Dataset Name This dataset card aims to be a base template for new datasets. It has been generated using this raw template. Dataset Details Dataset Description This dataset contains more than 3300 Persian tweets, crawled from X.com Each tweet is assigned a label, which is a number between 0 to 4. Label 0 indicates the sentiment of Happiness and Joy. Label 1 indicates the sentiment of Sadness. Label 2 indicates the sentiment of Anger and… See the full description on the dataset page: https://huggingface.co/datasets/moali-mkh-2000/PersianTwitterDataset-SentimentAnalysis.tabulartext-classification1K<n<10K0 likes52 downloads1y agoHugging Face18Danie1Arias /sentiment-analysis-catalan-reviews CSXSC: Classificador de Sentiments de Xarxes Socials en Català This repository contains the CSXSC (Classificador de Sentiments a Xarxes Socials en Català) dataset, a comprehensive corpus designed for sentiment analysis of Catalan-language content from social media. The dataset contains 23,788 text entries, each classified as positive, negative, or neutral. It was specifically constructed to address the significant class imbalance often found in user-generated content, resulting in a… See the full description on the dataset page: https://huggingface.co/datasets/Danie1Arias/sentiment-analysis-catalan-reviews.text10K<n<100K0 likes47 downloads1y agoHugging Face19jinunyachhyon /Sentiment_Analysis_Benchmarktext10K<n<100K0 likes45 downloads2y agoHugging Face20kalixlouiis /myanmar-sentiment-analysis Dataset Description This dataset consists of 1,665 annotated Myanmar language sentences designed for Sentiment Analysis tasks. The dataset is balanced across three sentiment categories, providing a robust foundation for training and fine-tuning machine learning models to understand Myanmar linguistic nuances, including sarcasm and compound sentence structures. Total size: 1,665 records Languages: Burmese (Myanmar) Task: Sentiment Classification (Positive, Negative, Neutral)… See the full description on the dataset page: https://huggingface.co/datasets/kalixlouiis/myanmar-sentiment-analysis.texttext-classification1K<n<10K4 likes41 downloads3mo agoHugging Face21LYTinn /sentiment-analysis-tweettext10K<n<100K1 likes40 downloads4y agoHugging Face22pryshlyak /sentiment_analysistext100K<n<1M0 likes40 downloads3y agoHugging Face23Senem /Nostalgic_Sentiment_Analysis_of_YouTube_Comments_Data Dataset Summary The dataset is a collection of Youtube Comments and it was captured using the YouTube Data API. The data set consists of 1500 nostalgic and non-nostalgic comments in English. Languages The language of the data is English. Citation If you find this dataset usefull for your study, please cite the paper as followed: @article{postalcioglu2020comparison, title={Comparison of Neural Network Models for Nostalgic Sentiment Analysis of YouTube… See the full description on the dataset page: https://huggingface.co/datasets/Senem/Nostalgic_Sentiment_Analysis_of_YouTube_Comments_Data.texttext-classification1K<n<10K5 likes38 downloads3y agoHugging Face24syedkhalid0 /Sentiment-Analysis Sentiment Analysis Dataset Overview This dataset is designed for sentiment analysis tasks, providing labeled examples across three sentiment categories: 0: Negative 1: Neutral 2: Positive It is suitable for training, validating, and testing text classification models in tasks such as social media sentiment analysis, customer feedback evaluation, and opinion mining. Dataset Details Key Features Type: CSV Language: English Labels: 0:… See the full description on the dataset page: https://huggingface.co/datasets/syedkhalid0/Sentiment-Analysis.texttext-classification100K<n<1M0 likes37 downloads2y agoHugging Face25anotherpolarbear /vietnamese-sentiment-analysistexttext-classification10K<n<100K1 likes36 downloads3y agoHugging Face26sudhanshusinghaiml /airlines-sentiment-analysistext10K<n<100K0 likes36 downloads2y agoHugging Face27ankitjha07 /Synthetic_dataset-MCAeConsultation_Sentiment_Analysistext10K<n<100K0 likes35 downloads21d agoHugging Face28chillies /course-review-multilabel-sentiment-analysistabular1K<n<10K0 likes32 downloads2y agoHugging Face29Aldoreni45 /sentiment_analysis_preprocessed_datasetBrief idea about dataset: This dataset is designed for a Text Classification to be specific Multi Class Classification, inorder to train a model (Supervised Learning) for Sentiment Analysis. Also to be able retrain the model on the given feedback over a wrong predicted sentiment this dataset will help to manage those things using Other Features. Main Features text labels This feature variable has all sort of texts, sentences, tweets, etc. This target variable contains 3 types of… See the full description on the dataset page: https://huggingface.co/datasets/Aldoreni45/sentiment_analysis_preprocessed_dataset.tabulartext-classification100K<n<1M0 likes32 downloads8mo agoHugging Face30mltrev23 /financial-sentiment-analysis Model Card for Sentiment Analysis on Financial News Overview This dataset contains sentiments for financial news headlines from the perspective of a retail investor. The data is derived from the research by Malo et al. (2014), which focuses on detecting semantic orientations in economic texts. Dataset Details Source: Malo, P., Sinha, A., Takala, P., Korhonen, P., and Wallenius, J. (2014). “Good debt or bad debt: Detecting semantic orientations in economic… See the full description on the dataset page: https://huggingface.co/datasets/mltrev23/financial-sentiment-analysis.text1K<n<10K4 likes31 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.