datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Nostalgic_Sentiment_Analysis_of_YouTube_Comments_Data
Dataset Summary
The dataset is a collection of Youtube Comments and it was captured using the YouTube Data API.
The data set consists of 1500 nostalgic and non-nostalgic comments in English.
Languages
The language of the data is English.
Citation
If you find this dataset usefull for your study, please cite the paper as followed:
@article{postalcioglu2020comparison,
title={Comparison of Neural Network Models for Nostalgic Sentiment Analysis of YouTube… See the full description on the dataset page: https://huggingface.co/datasets/Senem/Nostalgic_Sentiment_Analysis_of_YouTube_Comments_Data.sentiment_analysis_financial_news_dataData_Analysis_Workflow_20240509_134737Data_Analysis_Workflow_20240509_121811olympic-data-analysis
The analysis behind the glory: 120 Years of Data
Project Overview
In this project, I performed a comprehensive Exploratory Data Analysis (EDA) on a dataset covering 120 years of Olympic history. My main goal was to transform a messy, historical dataset into a clean, analyzed resource to uncover the key physical, demographic, and geopolitical factors that determine an athlete's success.
Dataset Description
The dataset provides a wide view of Olympic athletes… See the full description on the dataset page: https://huggingface.co/datasets/grasimus/olympic-data-analysis.Data_Analysis_Workflow_20240504_054612sentiment-analysis-for-mental-health-Combined-DataData_Analysis_Workflow_20240504_054711data-table-analysisThis dataset can be used for benchmarking LLM Structured Outputs via the code here:
https://github.com/cleanlab/structured-output-benchmark/
Encoding_Mismatch_Analysis_Data
Encoding Mismatch Analysis Data
This repository publishes the prepared numerical analysis artifacts associated
with From Per-Image Low-Rank to Encoding Mismatch: Rethinking Feature
Distillation in Vision Transformers. It is analysis data, not an image or
model-training dataset, and it does not redistribute ImageNet.
Links
Paper: https://arxiv.org/abs/2511.15572
Hugging Face paper page: https://huggingface.co/papers/2511.15572
Code and analysis scripts:… See the full description on the dataset page: https://huggingface.co/datasets/Huiyuancs/Encoding_Mismatch_Analysis_Data.Data_Analysis_Workflow_20240503_223029Data_Analysis_Workflow_20240610_100454Data_Analysis_Workflow_20240504_052456Data_Analysis_Workflow_20240504_061548pandas_data_analysis_questionsData_Analysis_Workflow_20240612_123344Data_Analysis_Workflow_20240503_224446Data_Analysis_Workflow_20240503_195700Data_Analysis_Workflow_20240503_195731Data_Analysis_Workflow_20240509_164435Data_Analysis_Workflow_20240422_074629Data_Analysis_Workflow_20240429_065730NCSS_2023_Data_AnalysisData_Analysis_Workflow_20240509_140347Data_Analysis_Workflow_20240424_173553Data_Analysis_Workflow_20240503_195322Data_Analysis_Workflow_20240503_200819Data_Analysis_Workflow_20240608_134433Data_Analysis_Workflow_20240610_094328Data_Analysis_Workflow_20240610_102312
