CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01edmundmiller /rocketleague-analysis Rocket League Analysis Local Rocket League replay analysis using Ballchasing API exports and plain DuckDB. The report is meant to answer one practical question: what should I work on next from my saved replay sample? Quick Start uv sync --locked UV_CACHE_DIR=/tmp/rocketleague-uv-cache \ uv run --locked pytest -v uv run --locked python scripts/analyze_scenarios.py \ --replay-dir /path/to/Rocket\ League/TAGame/Demos \ --limit 10 Start with CONTRIBUTING.md… See the full description on the dataset page: https://huggingface.co/datasets/edmundmiller/rocketleague-analysis.imagen<1K0 likes1.3k downloads11d agoHugging Face02ramankamran /retina-age-analysis Retina Age Analysis Dataset Dataset Description This dataset contains 9,857 retinal fundus images from 5,393 patients for age prediction tasks. Dataset Summary Task: Age prediction from retinal fundus images Images: 9,857 high-quality retinal images Patients: 5,393 unique patients Age Range: 5-97 years Image Format: JPEG Average Image Size: ~1 MB Supported Tasks Regression: Predict continuous age (5-97 years) Classification: Predict age group (5… See the full description on the dataset page: https://huggingface.co/datasets/ramankamran/retina-age-analysis.imageimage-classification1K<n<10K0 likes596 downloads11mo agoHugging Face03hugginglearners /amazon-reviews-sentiment-analysis Dataset Card for amazon reviews for sentiment analysis Dataset Summary One of the most important problems in e-commerce is the correct calculation of the points given to after-sales products. The solution to this problem is to provide greater customer satisfaction for the e-commerce site, product prominence for sellers, and a seamless shopping experience for buyers. Another problem is the correct ordering of the comments given to the products. The prominence of misleading… See the full description on the dataset page: https://huggingface.co/datasets/hugginglearners/amazon-reviews-sentiment-analysis.tabular1K<n<10K5 likes588 downloads4y agoHugging Face04ParsiAI /digikala-sentiment-analysistabulartext-classification1K<n<10K3 likes517 downloads2y agoHugging Face05NuBerea /source-analysisgated NuBerea Source Analysis Source-critical analysis of the Hebrew Bible, Septuagint, New Testament, Vulgate, and Second Temple literature. The dataset carries machine-generated source and tradition annotations at the verse level — the classical concerns of source criticism (documentary strata in the Old Testament, corpus structure in the New Testament, the pathway of Old Testament traditions into New Testament citation) expressed as structured data — together with semantic-domain… See the full description on the dataset page: https://huggingface.co/datasets/NuBerea/source-analysis.tabularfeature-extraction100K<n<1M0 likes517 downloads2d agoHugging Face06timchen0618 /browsecomp-plus-selected-tools-analysis-v1 BrowseComp-Plus: Selected Tools Analysis Side-by-side view of selected tool calls from a reference trajectory alongside the new agent trajectory conditioned on those steps. Retrieval model: Qwen3-Embedding-8BAgent model: gpt-oss-120bRun: traj_summary_ext_selected_tools_gpt-oss-120b_seed0 Columns Column Description query_id Query identifier rationale GPT rationale for why these k steps were selected from the reference trajectory selected_indices Step indices… See the full description on the dataset page: https://huggingface.co/datasets/timchen0618/browsecomp-plus-selected-tools-analysis-v1.tabularn<1K0 likes511 downloads6mo agoHugging Face07NuBerea /translation-analysisgated NuBerea Translation Verse Texts Verse-level texts of historical Bible translations (Clementine Vulgate, Luther Bible 1545, Matthew's Bible 1537). Part of the NuBerea curated corpus estate of biblical and historical texts. Attribution Upstream Data Sources Source License Clementine Vulgate, NOCR Public Domain Luther Bible 1545, NOCR Public Domain Matthew's Bible 1537, Textus Receptus Bibles Public Domain NuBerea project. Licensed… See the full description on the dataset page: https://huggingface.co/datasets/NuBerea/translation-analysis.tabularfeature-extraction100K<n<1M0 likes458 downloads9d agoHugging Face08NuBerea /pseudepigrapha-analysisgated NuBerea Pseudepigrapha Analysis Derived linguistic datasets over pseudepigraphal literature, part of the NuBerea curated corpus estate. Covers the Greek and Latin witnesses of these texts along with a multilingual view across the available witness languages. License CC BY 4.0. Attribution Source Link License NuBerea project https://huggingface.co/NuBerea CC BY 4.0 tabularfeature-extraction10K<n<100K0 likes437 downloads2d agoHugging Face09Sp1786 /multiclass-sentiment-analysis-dataset Dataset Card for Dataset Name Dataset Summary This dataset card aims to be a base template for new datasets. It has been generated using this raw template. Supported Tasks and Leaderboards [More Information Needed] Languages [More Information Needed] Dataset Structure Data Instances [More Information Needed] Data Fields [More Information Needed] Data Splits [More Information Needed] Dataset Creation… See the full description on the dataset page: https://huggingface.co/datasets/Sp1786/multiclass-sentiment-analysis-dataset.tabulartext-classification10K<n<100K29 likes432 downloads3y agoHugging Face10NuBerea /lxx-analysisgated NuBerea Research: LXX Translation-Technique Noise Model Quantitative study of Septuagint translation technique: verse-by-verse measurements of where the ancient Greek translation (LXX) diverges from the Hebrew Masoretic Text, with book-level statistical summaries. The material lets researchers distinguish a translator's habitual working style — free versus literal rendering — from genuine textual anomalies worth close scholarly attention, putting on a measurable footing what LXX… See the full description on the dataset page: https://huggingface.co/datasets/NuBerea/lxx-analysis.tabularfeature-extraction10K<n<100K0 likes431 downloads2d agoHugging Face11NuBerea /septuagint-analysisgated NuBerea Septuagint Textual Analysis Curated datasets for study of the Septuagint (the ancient Greek translation of the Hebrew Bible), part of the NuBerea corpus estate of biblical and patristic texts. It gathers Septuagint verse texts, apparatus notes, and edition-comparison material into a set of ready-to-load configurations. Attribution This dataset derives from the following upstream sources, which require attribution: Source License Rahlfs… See the full description on the dataset page: https://huggingface.co/datasets/NuBerea/septuagint-analysis.tabularfeature-extraction10K<n<100K0 likes429 downloads2mo agoHugging Face12GildasLeDrogoff /spotify-huge-track-analysis-dataset Spotify Track Analysis Dataset General Description This dataset provides a large-scale, research-oriented analytical representation of Spotify music data. It is centered on tracks as musical recordings (track_id), while preserving explicit artist attribution as defined by Spotify’s native credit model. Each row corresponds to a track–artist association, identified by: a Spotify track identifier (track_id) a credited artist name (artist_name) A single track may appear on… See the full description on the dataset page: https://huggingface.co/datasets/GildasLeDrogoff/spotify-huge-track-analysis-dataset.tabulartabular-classification10M<n<100M5 likes426 downloads7mo agoHugging Face13jtviegas /ticker_analysis_articlestabular10K<n<100K0 likes421 downloads8h agoHugging Face14lia-prop13 /startup-Investments-analysis 📊 StartUp Investments EDA 1. Background & Objectives This project explores a comprehensive dataset of startup investments (sourced from Crunchbase) to uncover the primary factors that predict a startup's survival and trajectory in a competitive market. Through this Exploratory Data Analysis (EDA), we analyze historical funding data, investment rounds, and market categories to determine which variables drive specific company outcomes - namely, whether a business… See the full description on the dataset page: https://huggingface.co/datasets/lia-prop13/startup-Investments-analysis.imagetabular-classification1K<n<10K1 likes349 downloads21d agoHugging Face15attik /Instacart-Market-Basket-Analysistabular1M<n<10M0 likes331 downloads8mo agoHugging Face16OALL /details_deep-analysis-research__D2IL-Arabic-Qwen2.5-72B-Instruct-v0.2_v2 Dataset Card for Evaluation run of deep-analysis-research/D2IL-Arabic-Qwen2.5-72B-Instruct-v0.2 Dataset automatically created during the evaluation run of model deep-analysis-research/D2IL-Arabic-Qwen2.5-72B-Instruct-v0.2. The dataset is composed of 116 configuration, each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_deep-analysis-research__D2IL-Arabic-Qwen2.5-72B-Instruct-v0.2_v2.tabular100K<n<1M0 likes321 downloads1y agoHugging Face17deep-analysis-research /details_D2IL-Arabic-Qwen2.5-72B-Instruct-v0.1tabular10K<n<100K0 likes299 downloads1y agoHugging Face18jtviegas /ticker_analysis_pricestabular10K<n<100K0 likes298 downloads8h agoHugging Face19eliel2003 /student-burnout-analysis2026 🔥 Predicting Academic Burnout: A Multivariate Analysis of Student Stressors Exploring how financial pressure, family expectations, and social support shape burnout in university students. Project Overview & Data Walkthrough 📋 Abstract Academic burnout is an increasingly recognized phenomenon with far-reaching consequences for student wellbeing and performance. This study investigates the relationship between external environmental stressors —… See the full description on the dataset page: https://huggingface.co/datasets/eliel2003/student-burnout-analysis2026.tabulartabular-regression1K<n<10K0 likes291 downloads6mo agoHugging Face20princeton-nlp /QuRatedPajama-1B_tokens_for_analysis QuRatedPajama Paper: QuRating: Selecting High-Quality Data for Training Language Models This dataset is a 1B token subset derived from princeton-nlp/QuRatedPajama-260B, which is a subset of cerebras/SlimPajama-627B annotated by princeton-nlp/QuRater-1.3B with sequence-level quality ratings across 4 criteria: Educational Value - e.g. the text includes clear explanations, step-by-step reasoning, or questions and answers Facts & Trivia - how much factual and trivia knowledge the text… See the full description on the dataset page: https://huggingface.co/datasets/princeton-nlp/QuRatedPajama-1B_tokens_for_analysis.tabular1M<n<10M6 likes260 downloads2y agoHugging Face21gretelai /gretel-financial-risk-analysis-v1 gretelai/gretel-financial-risk-analysis-v1 This dataset contains synthetic financial risk analysis text generated by fine-tuning Phi-3-mini-128k-instruct on 14,306 SEC filings (10-K, 10-Q, and 8-K) from 2023-2024, utilizing differential privacy. It is designed for training models to extract key risk factors and generate structured summaries from financial documents while demonstrating the application of differential privacy to safeguard sensitive information. This dataset showcases… See the full description on the dataset page: https://huggingface.co/datasets/gretelai/gretel-financial-risk-analysis-v1.tabulartext-classification1K<n<10K12 likes254 downloads2y agoHugging Face22prasadsawant7 /sentiment_analysis_preprocessed_datasetBrief idea about dataset: This dataset is designed for a Text Classification to be specific Multi Class Classification, inorder to train a model (Supervised Learning) for Sentiment Analysis. Also to be able retrain the model on the given feedback over a wrong predicted sentiment this dataset will help to manage those things using Other Features. Main Features text labels This feature variable has all sort of texts, sentences, tweets, etc. This target variable contains 3 types of… See the full description on the dataset page: https://huggingface.co/datasets/prasadsawant7/sentiment_analysis_preprocessed_dataset.tabulartext-classification100K<n<1M4 likes249 downloads3y agoHugging Face23guyshilo12 /diabetes_eda_analysis Diabetes Dataset — Exploratory Data Analysis (EDA) This repository contains a diabetes-related tabular dataset and a complete Exploratory Data Analysis (EDA).The main objective of this project was to learn how to conduct a structured EDA, apply best practices, and extract meaningful insights from real-world health data. The analysis includes correlations, distributions, group comparisons, class balance exploration, and statistical interpretations that illustrate how different… See the full description on the dataset page: https://huggingface.co/datasets/guyshilo12/diabetes_eda_analysis.imagetabular-classificationn<1K1 likes237 downloads10mo agoHugging Face24Hanno-Labs /harvey-labs-llm-artifact-analysis Harvey Labs LLM artifact analysis This dataset contains artifacts from a non-LLM analysis of the Harvey Labs DOCX corpus. The analysis used filename similarity, document extraction heuristics, a small manually labeled seed set, CatBoost native text features, and a native CatBoost embedding feature built from a mean Word2Vec representation. It was designed to find documents where an LLM refused the requested task and returned a safe alternative instead. Source… See the full description on the dataset page: https://huggingface.co/datasets/Hanno-Labs/harvey-labs-llm-artifact-analysis.tabular10K<n<100K0 likes191 downloads2mo agoHugging Face25SaguaroCapital /sentiment-analysis-in-commodity-market-gold Dataset Card for Sentiment Analysis of Commodity News (Gold) This is a news dataset for the commodity market which has been manually annotated for 10,000+ news headlines across multiple dimensions into various classes. The dataset has been sampled from a period of 20+ years (2000-2021). The dataset was curated by Ankur Sinha and Tanmay Khandait and is detailed in their paper "Impact of News on the Commodity Market: Dataset and Results." It is currently published by the authors on… See the full description on the dataset page: https://huggingface.co/datasets/SaguaroCapital/sentiment-analysis-in-commodity-market-gold.tabulartext-classification10K<n<100K6 likes185 downloads2y agoHugging Face26mqraitem /Deforest-Analysis NRT Forest-Loss Test Set for Student Analysis This package contains the fixed held-out test split used for a study of near-real-time forest-loss detection from four HLS observations. It is an analysis release: it includes inputs, labels, model outputs, and visual renders, but no checkpoints or GPU-dependent code. The intended analyses are prediction-shape comparison, per-connected-component performance, and seasonal performance. Do not use this test set to select model… See the full description on the dataset page: https://huggingface.co/datasets/mqraitem/Deforest-Analysis.imageimage-segmentationn<1K0 likes181 downloads2mo agoHugging Face27slliac /isom5240-td-traffic-analysistabularn<1K0 likes162 downloads1y agoHugging Face28drukeroni /airline-satisfaction-analysis Airline Passenger Satisfaction – EDA Report This project analyzes the Airline Passenger Satisfaction Dataset, containing 103,904 rows and 25 columns describing passenger demographics, flight information, and service ratings.The goal is to understand which factors influence satisfaction, identify important service features,and compare satisfaction between different traveler types and flight classes. Dataset Overview The dataset includes: Passenger demographics (age… See the full description on the dataset page: https://huggingface.co/datasets/drukeroni/airline-satisfaction-analysis.tabular100K<n<1M0 likes162 downloads10mo agoHugging Face29wangd12 /XBRL_analysis XBRL Extraction Dataset The is the official dataset introduced in the paper FinLoRA: Benchmarking LoRA Methods for Fine-Tuning LLMs on Financial Datasets tabular10K<n<100K2 likes151 downloads1y agoHugging Face30ag00dman /student-depression-analysis Assignment #1: EDA & Dataset Predicting and Preventing Student Depression Student: Amit GoodmanProgram: Economics & Entrepreneurship, Reichman University (RUNI)Date: March 2026 Project Overview In this project, I explore the "Student Depression Dataset" to build a narrative around student well-being. By analyzing academic pressure, financial stress, and lifestyle habits, I aim to identify predictable risk factors and uncover actionable protective measures.… See the full description on the dataset page: https://huggingface.co/datasets/ag00dman/student-depression-analysis.tabulartabular-classification10K<n<100K2 likes150 downloads5mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.