CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Weiyun1025 /InternVL-Performanceimagen<1K0 likes8.2k downloads1y agoHugging Face02yeigen /fannie-mae-loan-performancetabular1B<n<10B0 likes2.6k downloads7mo agoHugging Face03whitphx /transformersjs-performance-leaderboard-results0 likes1.3k downloads11mo agoHugging Face04docling-project /performance-dataset-bo10kdocument10K<n<100K0 likes719 downloads2mo agoHugging Face05docling-project /performance-dataset-bo767documentn<1K1 likes712 downloads5mo agoHugging Face06jason1966 /alinaboulsi_digital-marketing-performance-dataset Digital Marketing Performance Dataset A Synthetic, Benchmark-Based Dataset for Multi-Platform Marketing Analytics & BI Dataset Info Source: Kaggle Original Size: 1.83 MB Kaggle Downloads: 145 Files: 3 Files README_DATASET.md data_dictionary.csv digital_marketing_dataset_30k.csv Mirrored from Kaggle 0 likes686 downloads6mo agoHugging Face07cnmat /human_performance_dataset Human performance dataset Piano MIDI dataset used for training and evaluation. Layout and processing: Source: MAESTRO (or similar) with a prompt (or user_prompt) per piece. Augmentation: dataset_preprocess/augment_dataset.py adds pitch/time/velocity variants; output CSV lists originals and augmented files. Grouping: dataset_preprocess/group_dataset.py copies MIDIs into grouped/ as composer/genre/filename.mid (and augmented/composer/genre/ for augmented). Genre is taken from the… See the full description on the dataset page: https://huggingface.co/datasets/cnmat/human_performance_dataset.1 likes309 downloads6mo agoHugging Face08wanshenl /pgsql-performance-raw0 likes265 downloads2y agoHugging Face09alirezaaminzadeh /solverport-solver-performance SolverPort Solver Performance Dataset Solver benchmark results across 8 solvers and 10 optimization problem families. Metrics per Run Runtime (seconds) Optimality gap (%) Feasibility status Time to first feasible solution Best bound and gap improvement rate Search speed PAR10 penalty score Oracle regret vs best solver Solvers Profiled CP-SAT, HiGHS, CBC, SCIP, GLPK, Gurobi, MiniZinc, ALNS Files {instance_id}_performance.json — full… See the full description on the dataset page: https://huggingface.co/datasets/alirezaaminzadeh/solverport-solver-performance.textn<1K0 likes258 downloads1mo agoHugging Face10Alej0909 /fannie-mae-loan-performance-rawtabular1B<n<10B0 likes251 downloads7mo agoHugging Face11FastVideo /performance-tracking0 likes172 downloads17d agoHugging Face12mstz /student_performance Student performance The Student performance dataset from Kaggle. Configuration Task Description encoding Encoding dictionary showing original values of encoded features. math Binary classification Has the student passed the math exam? writing Binary classification Has the student passed the writing exam? reading Binary classification Has the student passed the reading exam? Usage from datasets importload_dataset dataset =… See the full description on the dataset page: https://huggingface.co/datasets/mstz/student_performance.tabulartabular-classification1K<n<10K3 likes161 downloads1y agoHugging Face13slymachenko /image-deblurring-performance-analysis0 likes137 downloads1y agoHugging Face14alirezaaminzadeh /frontierco-solver-performance frontierco-solver-performance Solver performance benchmark dataset produced by the FrontierCO Solver Arena. Source Built on profiles aligned with CO-Bench/FrontierCO. Contents Field Description instance_id Unique instance identifier problem_type One of 8 CO problems size small / medium / large difficulty easy / hard time_budget_sec 10 / 30 / 60 / 300 solver_id One of 13 solvers optimality_gap_pct Gap to known optimum… See the full description on the dataset page: https://huggingface.co/datasets/alirezaaminzadeh/frontierco-solver-performance.tabularothern<1K0 likes135 downloads1mo agoHugging Face15Sri-Vigneshwar-DJ /Performance-Marketing-Data Performance Marketing Expert Dataset Dataset Description This dataset contains comprehensive performance marketing knowledge and logical reasoning patterns for Meta (Facebook/Instagram), Google Ads, and TikTok advertising platforms. It's designed for fine-tuning language models to understand brand verticals, performance marketing strategies, and develop reasoning capacity for creating winning ad campaigns. Dataset Structure Each example follows an… See the full description on the dataset page: https://huggingface.co/datasets/Sri-Vigneshwar-DJ/Performance-Marketing-Data.texttext-generationn<1K4 likes113 downloads1y agoHugging Face16LLM-OS-Models /korean-embedding-performance-v1-performance-1m Korean Embedding Performance v1 — 1M Qwen3-Embedding-8B의 한국어 retrieval data-scale 실험을 위한 정확히 1,000,000-row 연구·비상업 contrastive dataset이다. release_eligible: false, 통합 라이선스 other이며 upstream source 조건을 재허가하지 않는다. 구성 계열 Rows 비율 역할 nlpai-lab/ko-triplet-v1.0 600,254 60.03% 넓은 한국어 QA/retrieval core F2 Korean QA/instruction 287,000 28.70% webfaq, mqa, koalpaca, realQA, komagpie F2 retrieval task train-family 4,146 0.41% MIRACL, MrTidy, MLDR F2 PAWS-X… See the full description on the dataset page: https://huggingface.co/datasets/LLM-OS-Models/korean-embedding-performance-v1-performance-1m.textsentence-similarity1M<n<10M0 likes104 downloads2mo agoHugging Face17AAU-NLP /effective-performance-measurement Effective Performance Measurement: KPI Extraction Datasets This dataset repository accompanies the ACL 2026 (Industry Track) paper: "Effective Performance Measurement: Challenges and Opportunities in KPI Extraction from Earnings Calls". It contains three novel benchmarks and a prediction set designed to evaluate the extraction of Key Performance Indicators (KPIs) from unstructured financial texts, specifically comparing highly regulated SEC filings to conversational earnings calls.… See the full description on the dataset page: https://huggingface.co/datasets/AAU-NLP/effective-performance-measurement.text-classification3 likes100 downloads5mo agoHugging Face18LLM-OS-Models /korean-embedding-performance-v1-ablation-200k Korean Embedding Performance v1 — Ablation 200K Qwen3-Embedding-8B의 한국어 retrieval continued fine-tuning에서 LoRA/DoRA/부분 및 full fine-tuning, loss, hard-negative 전략을 비교하기 위한 200,000-row 연구·비상업 성능 데이터다. release_eligible: false이며 통합 라이선스는 other다. upstream source별 조건을 재허가하지 않는다. 구성 계열 Rows 역할 nlpai-lab/ko-triplet-v1.0@1f5d72d 100,254 넓은 한국어 QA/retrieval core F2 Korean QA/instruction 68,000 webfaq, mqa, koalpaca, realQA, komagpie F2 retrieval task… See the full description on the dataset page: https://huggingface.co/datasets/LLM-OS-Models/korean-embedding-performance-v1-ablation-200k.textsentence-similarity100K<n<1M0 likes98 downloads2mo agoHugging Face19bethgelab /frequency_determines_performanceFrequency estimation results and tagged samples: counts_and_indices.zip contains all the result jsons (for the estimated frequencies for image-only, text-only and image-text searches) and the sample indices that are tagged to each concept for the LAION400m/LAION-Aesthetics datasets. Constructed dictionaries and other pretraining and downstream data artefacts: Due to the large size of all our data artefacts, we release our dictionaries and other feature artefacts as split files of a 110GB… See the full description on the dataset page: https://huggingface.co/datasets/bethgelab/frequency_determines_performance.zero-shot-classificationn<1K4 likes96 downloads2y agoHugging Face20LLM-OS-Models /korean-embedding-performance-v1-pilot-50k Korean Embedding Performance v1 — Pilot 50K 주의: 이 revision은 공개 benchmark 성능 후보 학습에 사용하면 안 된다. 사후 15-task exact text-hash 감사에서 평가 query 고유 hash 4개가 확인됐다. 파이프라인·최적화 진단과 contamination ablation에만 남기며, 교체본은 ablation-200k이다. Qwen3-Embedding 계열의 한국어 retrieval 성능 실험을 위한 50,000-row 연구용 contrastive dataset이다. 각 row는 instruction-aware query, positive passage 1개, hard/easy negative passage 1–7개를 ms-swift embedding message schema로 저장한다. 사용 조건과 공개 범위 이 저장소의 통합 라이선스는 other다.… See the full description on the dataset page: https://huggingface.co/datasets/LLM-OS-Models/korean-embedding-performance-v1-pilot-50k.texttext-retrieval10K<n<100K0 likes95 downloads2mo agoHugging Face21neuralsorcerer /student-performance Student Performance Synthetic Dataset v2 A reproducible, explicitly structured synthetic high-school student dataset for machine-learning benchmarking, data-engineering tests, educational analytics method development, and controlled fairness experiments. Scope and non-claims The data are entirely synthetic. Generator parameters are designed for internal coherence and are not calibrated to a specific real school system, demographic population, causal effect, or… See the full description on the dataset page: https://huggingface.co/datasets/neuralsorcerer/student-performance.tabulartabular-classification10M<n<100M2 likes87 downloads4d agoHugging Face22alaa1234ah /fbref_football_player_performance_2024-2025 FBref Football Player Performance Dataset (2024-2025 Season) Dataset Description This dataset contains comprehensive performance statistics for 2273 professional football players during the 2024-2025 season. Sourced from FBref, it includes both traditional metrics (goals, assists) and advanced analytics (xG, xAG, progressive actions) across top European leagues. Curated by: FBref License: Publicly available football statistics (check FBref terms for redistribution)… See the full description on the dataset page: https://huggingface.co/datasets/alaa1234ah/fbref_football_player_performance_2024-2025.tabular1K<n<10K0 likes84 downloads9mo agoHugging Face23michaelozon /student-performance-factors-analysis-michael-ozon🎓 Student Performance Factors — EDA & Insights Michael Ozon — Assignment #1 (EDA & Dataset) Reichman University – Data Science Course 🎥 Presentation Video https://drive.google.com/drive/folders/1cAXLzcZflMgv12EDlVTeQoKxzVumOjbd?usp=drive_link 📌 Project Overview This project explores the Student Performance Factors dataset, containing 6,607 student records and 20 academic, behavioral, lifestyle, and demographic features. The goal of this Exploratory Data Analysis (EDA) is to understand which… See the full description on the dataset page: https://huggingface.co/datasets/michaelozon/student-performance-factors-analysis-michael-ozon.imagen<1K1 likes74 downloads10mo agoHugging Face24whitphx /transformersjs-performance-leaderboard-results-dev20 likes73 downloads11mo agoHugging Face25datametrik /b2b-digital-marketing-performance-benchmarks B2B & Ecommerce Performance Marketing Benchmarks Maintained and published by Datametrik — Performance Marketing and Growth Agency. textn<1K0 likes70 downloads1mo agoHugging Face26federicomoreno /marketing-campaign-performance-200k Marketing Campaign Performance Dataset (200k) Mirror de Kaggle: Marketing Campaign Performance Dataset (manishabhatt22), 200.000 filas de campañas de marketing (verificado: rango de fechas 2021-01-01 a 2021-12-31, no dos años como dice la card de Kaggle). Subido aquí para poder cargarlo con datasets.load_dataset() sin credenciales de Kaggle. Contenido data/marketing_campaign_dataset.csv — 200.000 filas, ~27 MB, 16 columnas. data/data_dictionary.md — descripción… See the full description on the dataset page: https://huggingface.co/datasets/federicomoreno/marketing-campaign-performance-200k.tabular100K<n<1M0 likes67 downloads6d agoHugging Face27ibm-research /LLM_Fine-Tuning_Performance LLM Fine-Tuning Performance Benchmark Dataset Dataset Summary This dataset contains performance benchmarks for Large Language Model (LLM) fine-tuning across various hardware and software configurations. It includes throughput measurements (tokens per second) for 959 valid configurations, collected over 1000 GPU hours on a Kubernetes cluster. The dataset is designed for research on predictive performance modeling, specifically for evaluating methods that handle Categorical… See the full description on the dataset page: https://huggingface.co/datasets/ibm-research/LLM_Fine-Tuning_Performance.tabular-regressionn<1K2 likes66 downloads4mo agoHugging Face283zden /fbref_football_player_performance_2024-2025 FBref Football Player Performance Dataset (2024-2025 Season) Dataset Description This dataset contains comprehensive performance statistics for 2273 professional football players during the 2024-2025 season. Sourced from FBref, it includes both traditional metrics (goals, assists) and advanced analytics (xG, xAG, progressive actions) across top European leagues. Curated by: FBref License: Publicly available football statistics (check FBref terms for redistribution)… See the full description on the dataset page: https://huggingface.co/datasets/3zden/fbref_football_player_performance_2024-2025.tabular1K<n<10K9 likes63 downloads1y agoHugging Face29matthewcox /paragru-performance-audit-embeddings0 likes61 downloads2mo agoHugging Face30LLM-OS-Models /korean-embedding-performance-v1-sionic-retrieval-train-family-4146 Korean Sionic Retrieval Train-Family 4,146 F2LLM-v2가 공개한 Korean MIRACL, MrTidy, MLDR train-family row만 1M decontaminated curriculum에서 lossless 추출한 target-adaptation dataset이다. 공개 evaluation query는 포함하지 않으며 current-student HN7 mining 전의 source artifact다. 구성과 목적 source rows 역할 f2_miracl_ko_train 700 MIRACL Korean retrieval train-family f2_mrtidy_korean_train 1,200 MrTidy Korean train f2_mldr_ko_train 2,246 MLDR Korean long-document train-family 합계 4… See the full description on the dataset page: https://huggingface.co/datasets/LLM-OS-Models/korean-embedding-performance-v1-sionic-retrieval-train-family-4146.texttext-retrieval1K<n<10K0 likes56 downloads2mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.