datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
amazon-reviews-sentiment-analysis
Dataset Card for amazon reviews for sentiment analysis
Dataset Summary
One of the most important problems in e-commerce is the correct calculation of the points given to after-sales products. The solution to this problem is to provide greater customer satisfaction for the e-commerce site, product prominence for sellers, and a seamless shopping experience for buyers. Another problem is the correct ordering of the comments given to the products. The prominence of misleading… See the full description on the dataset page: https://huggingface.co/datasets/hugginglearners/amazon-reviews-sentiment-analysis.multiclass-sentiment-analysis-dataset
Dataset Card for Dataset Name
Dataset Summary
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Supported Tasks and Leaderboards
[More Information Needed]
Languages
[More Information Needed]
Dataset Structure
Data Instances
[More Information Needed]
Data Fields
[More Information Needed]
Data Splits
[More Information Needed]
Dataset Creation… See the full description on the dataset page: https://huggingface.co/datasets/Sp1786/multiclass-sentiment-analysis-dataset.sentiment_analysis_preprocessed_datasetBrief idea about dataset:
This dataset is designed for a Text Classification to be specific Multi Class Classification, inorder to train a model (Supervised Learning) for Sentiment Analysis.
Also to be able retrain the model on the given feedback over a wrong predicted sentiment this dataset will help to manage those things using Other Features.
Main Features
text
labels
This feature variable has all sort of texts, sentences, tweets, etc.
This target variable contains 3 types of… See the full description on the dataset page: https://huggingface.co/datasets/prasadsawant7/sentiment_analysis_preprocessed_dataset.digikala-sentiment-analysisAspect-Based_Sentiment_Analysis_for_Catering
说明
数据集来源于AI Challenger 2018
sentiment_analysis_trainingset.csv 为训练集数据文件,共105000条评论数据
sentiment_analysis_validationset.csv 为验证集数据文件,共15000条评论数据
sentiment_analysis_testa.csv 为测试集A数据文件,共15000条评论数据
数据集分为训练、验证、测试A与测试B四部分。数据集中的评价对象按照粒度不同划分为两个层次,层次一为粗粒度的评价对象,例如评论文本中涉及的服务、位置等要素;层次二为细粒度的情感对象,例如“服务”属性中的“服务人员态度”、“排队等候时间”等细粒度要素。评价对象的具体划分如下表所示。
The dataset is divided into four parts: training, validation, test A and test B. This dataset builds a two-layer labeling system according to the… See the full description on the dataset page: https://huggingface.co/datasets/xcz0/Aspect-Based_Sentiment_Analysis_for_Catering.twitter-sentiment-meta-analysis
Twitter Sentiment Meta-Analysis Dataset
Dataset Description
This dataset contains sentiment analysis results for English tweets collected between September 2009 and January 2010. The tweets were processed and analyzed using 10 different sentiment classifiers, with the final sentiment score derived from principal component analysis (PCA).
Source Data
Original Data: Cheng-Caverlee-Lee Twitter Scrape (Sept 2009 - Jan 2010)
Number of Tweets: 138 690
Language:… See the full description on the dataset page: https://huggingface.co/datasets/agentlans/twitter-sentiment-meta-analysis.faroese_sentiment_analysis
Good or Bad News? Exploring GPT-4 for Sentiment Analysis on Faroese News Corpora
This dataset is a part of the research from the paper "Good or Bad News? Exploring GPT-4 for Sentiment Analysis for Faroese on a Public News Corpora," that focuses on the application of GPT-4 for sentiment analysis on Faroese news texts.
The study addresses the challenges of sentiment analysis in low-resource languages and evaluates the effectiveness of Large Language Models, specifically GPT-4, in… See the full description on the dataset page: https://huggingface.co/datasets/hafsteinn/faroese_sentiment_analysis.PersianTwitterDataset-SentimentAnalysis
Dataset Card for Dataset Name
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Dataset Details
Dataset Description
This dataset contains more than 3300 Persian tweets, crawled from X.com
Each tweet is assigned a label, which is a number between 0 to 4.
Label 0 indicates the sentiment of Happiness and Joy.
Label 1 indicates the sentiment of Sadness.
Label 2 indicates the sentiment of Anger and… See the full description on the dataset page: https://huggingface.co/datasets/moali-mkh-2000/PersianTwitterDataset-SentimentAnalysis.sentiment_analysis_preprocessed_datasetBrief idea about dataset:
This dataset is designed for a Text Classification to be specific Multi Class Classification, inorder to train a model (Supervised Learning) for Sentiment Analysis.
Also to be able retrain the model on the given feedback over a wrong predicted sentiment this dataset will help to manage those things using Other Features.
Main Features
text
labels
This feature variable has all sort of texts, sentences, tweets, etc.
This target variable contains 3 types of… See the full description on the dataset page: https://huggingface.co/datasets/Aldoreni45/sentiment_analysis_preprocessed_dataset.course-review-multilabel-sentiment-analysisE-commerce-Product-Review-Sentiment-Analysisamazon-reviews-sentiment-analysis
Dataset Card for amazon reviews for sentiment analysis
Dataset Summary
One of the most important problems in e-commerce is the correct calculation of the points given to after-sales products. The solution to this problem is to provide greater customer satisfaction for the e-commerce site, product prominence for sellers, and a seamless shopping experience for buyers. Another problem is the correct ordering of the comments given to the products. The prominence of… See the full description on the dataset page: https://huggingface.co/datasets/Zhengyif/amazon-reviews-sentiment-analysis.BilTweetNews-sentiment-analysis
Turkish Sentiment Analysis Tweet Dataset: BilTweetNews
The dataset contains tweets related to six major events from Turkish news sources between May 4, 2015
and Jan 8, 2017.
The dataset covers 6 major events:
May 25, 2015 One of the popular football clubs in Turkey, Galatasaray, wins the 2015
Turkish Super League.
Sep 6, 2015 A terrorist group, called PKK, attacked to soldiers in Dağlıca, a village in
southeastern Turkey.
Oct 7, 2015 A Turkish scientist, Aziz Sancar, won the 2015… See the full description on the dataset page: https://huggingface.co/datasets/ctoraman/BilTweetNews-sentiment-analysis.multiclass-sentiment-analysis-dataset
Dataset Card for Dataset Name
Dataset Summary
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Supported Tasks and Leaderboards
[More Information Needed]
Languages
[More Information Needed]
Dataset Structure
Data Instances
[More Information Needed]
Data Fields
[More Information Needed]
Data Splits
[More Information Needed]
Dataset Creation… See the full description on the dataset page: https://huggingface.co/datasets/Ajain1391/multiclass-sentiment-analysis-dataset.abdelmalekeladjelet_sentiment-analysis-dataset
Sentiment Analysis Dataset
Dataset for text classification
Dataset Info
Source: Kaggle
Original Size: 8.68 MB
Kaggle Downloads: 3,928
Files: 1
Files
sentiment_data.csv
Mirrored from Kaggle
Eye-tracking-and-Sentiment-Analysis-Dataset-IIEye-tracking and Sentiment Analysis Dataset-II (without fixation data)
(1) "text_and_annotations.csv" - Contains Sentences taken for our experiment and Annotation results.
Columns:
Text_ID - Id of the text
Text - Sentence(s).
Default_Polarity - Gold polarity label [-1 for negative sentiment and 1 for positive sentiment]
Aspect - Entity with respect to which sentiment is expressed.
Source - From where the text has been obtained.
Sarcasm - Whether the text contains irony/sarcasm or not.
[P1… See the full description on the dataset page: https://huggingface.co/datasets/shiv213/Eye-tracking-and-Sentiment-Analysis-Dataset-II.multiclass-sentiment-analysis-dataset
Dataset Card for Dataset Name
Dataset Summary
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Supported Tasks and Leaderboards
[More Information Needed]
Languages
[More Information Needed]
Dataset Structure
Data Instances
[More Information Needed]
Data Fields
[More Information Needed]
Data Splits
[More Information Needed]
Dataset Creation… See the full description on the dataset page: https://huggingface.co/datasets/Shivams123/multiclass-sentiment-analysis-dataset.multiclass-sentiment-analysis-dataset
Dataset Card for Dataset Name
Dataset Summary
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Supported Tasks and Leaderboards
[More Information Needed]
Languages
[More Information Needed]
Dataset Structure
Data Instances
[More Information Needed]
Data Fields
[More Information Needed]
Data Splits
[More Information Needed]
Dataset Creation… See the full description on the dataset page: https://huggingface.co/datasets/sayuyuyyy/multiclass-sentiment-analysis-dataset.tweets_sentiment_analysissentiment_analysis_dataset-9846258f-16b8-422d-a291-965686769dcasentiment_analysis_dataset-92f2ad6c-0376-4ce9-be36-3c6f26f144d7sentiment_analysis_dataset-920376de-dcd1-48e0-965d-597cb1bf5ac8amazon-reviews-sentiment-analysis
Dataset Card for amazon reviews for sentiment analysis
Dataset Summary
One of the most important problems in e-commerce is the correct calculation of the points given to after-sales products. The solution to this problem is to provide greater customer satisfaction for the e-commerce site, product prominence for sellers, and a seamless shopping experience for buyers. Another problem is the correct ordering of the comments given to the products. The prominence of misleading… See the full description on the dataset page: https://huggingface.co/datasets/Ha1200/amazon-reviews-sentiment-analysis.multiclass-sentiment-analysis-dataset
Dataset Card for Dataset Name
Dataset Summary
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Supported Tasks and Leaderboards
[More Information Needed]
Languages
[More Information Needed]
Dataset Structure
Data Instances
[More Information Needed]
Data Fields
[More Information Needed]
Data Splits
[More Information Needed]
Dataset Creation… See the full description on the dataset page: https://huggingface.co/datasets/Raihanultomal/multiclass-sentiment-analysis-dataset._1342_political_sentiment_analysissentiment_analysis_dataset-9fb4d6b9-87e1-4d47-a0bb-b265fc66d1e1multiclass-sentiment-analysis-dataset
Dataset Card for Dataset Name
Dataset Summary
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Supported Tasks and Leaderboards
[More Information Needed]
Languages
[More Information Needed]
Dataset Structure
Data Instances
[More Information Needed]
Data Fields
[More Information Needed]
Data Splits
[More Information Needed]
Dataset Creation… See the full description on the dataset page: https://huggingface.co/datasets/Renture666/multiclass-sentiment-analysis-dataset.amazon-reviews-sentiment-analysis
Dataset Card for amazon reviews for sentiment analysis
Dataset Summary
One of the most important problems in e-commerce is the correct calculation of the points given to after-sales products. The solution to this problem is to provide greater customer satisfaction for the e-commerce site, product prominence for sellers, and a seamless shopping experience for buyers. Another problem is the correct ordering of the comments given to the products. The prominence of misleading… See the full description on the dataset page: https://huggingface.co/datasets/tvyaishavi/amazon-reviews-sentiment-analysis.4000-Stories-with-sentiment-analysisPatent_sentiment_analysis
Dataset for annotation
Files and their description
combined_all_stats_csv
This is the CSV File containing statistics for the patent dataset
can be downloaded from the Git Repo.
What are the statistics it shows?
This file contains columns such as filename, total publications, total positive samples, total negative samples, and total neutral samples per every week of the year from 2010-2020.
combined_neut_data_csv
This is the CSV File, that contains details… See the full description on the dataset page: https://huggingface.co/datasets/Renukswamy/Patent_sentiment_analysis.
