datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
mpac-predictions
MPAC variant effect predictions
MPAC (Malinois with Parallel Aggregated Cross-validation) predicts cis-regulatory
activity of 200 bp human sequences in K562, HepG2 and SK-N-SH, and the allelic skew
caused by non-coding variants. This dataset holds the published predictions.
Identifying non-coding variant effects at scale via machine learning models of
cis-regulatory reporter assays.
The paper is the source of truth for how these tables were made and what they mean.
This is a… See the full description on the dataset page: https://huggingface.co/datasets/saarantras1/mpac-predictions.clinical-trial-outcomes-predictions
Clinical Trial Outcomes Prediction Dataset
A dataset of 1,366 binary forecasting questions about clinical trial outcomes, automatically generated and labeled using Lightning Rod Labs' Future-as-Label methodology.
Dataset Description
This dataset contains questions about pharmaceutical clinical trials from 2023-2024, paired with verified outcomes (success/failure). Each question asks whether a specific trial will meet its endpoints, receive FDA approval, or complete by a… See the full description on the dataset page: https://huggingface.co/datasets/3rdSon/clinical-trial-outcomes-predictions.negative-pi-predictions
negative-pi-predictions
what if π had digits before 3?
this dataset contains 100,000 digits predicted by a neural network at negative positions of π.
yes, this is exactly as stupid as it sounds.
what is this?
normally, we index the fractional digits of π like this:
position: 1 2 3 4 5 6 7 8 9 ...
digit: 1 4 1 5 9 2 6 5 3 ...
so:
π = 3.141592653589793...
↑
position 1
i trained a neural network to predict the digit at a given position using only… See the full description on the dataset page: https://huggingface.co/datasets/akaruineko/negative-pi-predictions.Prediction-Smartphone-Addiction-Submission-in-Kaggle
📱 Prediction Smartphone Addiction - Competition Submission in Kaggle
This dataset contains the data and/or prediction results used for a Kaggle competition related to smartphone addiction prediction.
The project focuses on analyzing smartphone usage and related behavioral or demographic features to build machine learning models capable of predicting smartphone addiction levels.
🎯 Project Overview
Smartphone usage has become an important part of everyday life.… See the full description on the dataset page: https://huggingface.co/datasets/Qamro/Prediction-Smartphone-Addiction-Submission-in-Kaggle.esnlir-llm-predictions
ESNLIR-LLM — per-pair predictions
Per-pair predictions for every model evaluated in An Analysis of the Performance of Large Language
Models in Spanish NLI Datasets with Causal Relationships (IBERAMIA 2026, to appear). Code in
Pacolas/NLI-via-LLM; part of the
ESNLIR-LLM
collection.
These are the raw outputs behind the paper's tables, so results can be re-scored, sliced by genre or
domain, or compared pair by pair without re-running any model.
Files… See the full description on the dataset page: https://huggingface.co/datasets/Flaglab/esnlir-llm-predictions.worldcup-2026-predictions
World Cup 2026 — Multi-Model Match Predictions
Pre-match win/draw/loss predictions for the FIFA World Cup 2026 fixtures,
produced by four independent prediction models and served live by a public
demo app. The opening two fixtures predate three model pipelines; those missing
pre-match forecasts remain null rather than being reconstructed after the fact.
This is a faithful export of data generated by a Salesforce-backed prediction
pipeline — no synthetic rows.
The… See the full description on the dataset page: https://huggingface.co/datasets/sergiopesch/worldcup-2026-predictions.aftermath_predictions
Aftermath of DrawEduMath
This contains predictions.csv, for recreating the results of the paper titled "The Aftermath of DrawEduMath: Vision Language Models Underperform with Struggling Students and Misdiagnose Errors".
This file contains model predictions for DrawEduMath QA from eleven vision-language models. These models include:
GPT-4.1
GPT-4.5 Preview
o4-mini
GPT-5
Claude Sonnet 3.7
Claude Sonnet 4
Claude Sonnet 4.5
Gemini 2.0 Flash
Gemini 2.5 Pro
Gemini 2.5 Pro Preview
Llama… See the full description on the dataset page: https://huggingface.co/datasets/lucy3/aftermath_predictions.Crash_Predictionsv2
Dataset Card for Dataset Name
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Dataset Details
Dataset Description
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Dataset Sources [optional]
Repository: [More… See the full description on the dataset page: https://huggingface.co/datasets/PratikGanesh/Crash_Predictionsv2.LegalSeg_RhetoricLLaMA_Predictionsabdulmaliklodhra_predictions-2026-csv
Life Expectancy from Birth to Death Prediction2026
Life Expectancy from Birth to Death of Human beings and Prediction about 2026
Dataset Info
Source: Kaggle
Original Size: 0.01 MB
Kaggle Downloads: 2
Files: 1
Files
predictions_2026.csv
Mirrored from Kaggle
LegalSeg_MTL_Predictionsrsuper-predictionsThis dataset contains the results of reproducing the paper "Learning Segmentation from Radiology Reports" using only public data, following R-Super Demo with Public Data.
Download using:
pip install -U "huggingface_hub[cli]"
# download
hf download edomerli/rsuper-predictions --repo-type dataset --local-dir ./rsuper-predictions
# extract
cd rsuper-predictions
tar -xzf preds_merlin.tar.gz
tar -xzf preds_pants.tar.gz
w2v_predictions_and_metricsplain_T5_small_predictionsprecog_miner_predictions_april_may15jokes-final-predictionswc2026-match-predictionsNuExtract3.4_27B-SFT_predictionssemantic_prime_T5_predictionssemantic_prime_T5_predictions_and_metricstourism-package-predictionspredictionssnappfood_predictions
SnappFood Sentiment Predictions
این فایل شامل پیشبینی احساسات نظرات مشتریان است.
