datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
alfa-scoring-trxhttps://ods.ai/competitions/dl-fintech-card-transactions
https://boosters.pro/championship/alfabattle2
scoring_data
scoring_data
At a state: several action chunks proposed from it, and how each one actually ended.
A branch the search dropped was cut off mid-episode, so it is resumed from its own snapshot
and carried to a finish — the action nobody executed still gets an answer to would this
have worked.
1037 searches · 13 tasks · 39,016 nodes, each with its own
state and image.
This repo hosts the data. What it means, how it was produced and how to use it live in
the code that wrote it:… See the full description on the dataset page: https://huggingface.co/datasets/mahgoobi/scoring_data.ToM-auto-scoring-basecaptini-scoring-referencessora-video-generation-style-likert-scoring
Rapidata Video Generation Preference Dataset
If you get value from this dataset and would like to see more in the future, please consider liking it.
This dataset was collected in ~1 hour using the Rapidata Python API, accessible to anyone and ideal for large scale data annotation.
Overview
In this dataset, ~6000 human evaluators were asked to rate AI-generated videos based on their visual appeal, without seeing the prompts used to generate them. The specific… See the full description on the dataset page: https://huggingface.co/datasets/Rapidata/sora-video-generation-style-likert-scoring.sora-video-generation-physics-likert-scoring
Rapidata Video Generation Physics Dataset
If you get value from this dataset and would like to see more in the future, please consider liking it.
This dataset was collected in ~1 hour using the Rapidata Python API, accessible to anyone and ideal for large scale data annotation.
Overview
In this dataset, ~6000 human evaluators were asked to rate AI-generated videos based on if gravity and colisions make sense, without seeing the prompts used to generate them.… See the full description on the dataset page: https://huggingface.co/datasets/Rapidata/sora-video-generation-physics-likert-scoring.AES2-essay-scoringhttps://www.kaggle.com/competitions/learning-agency-lab-automated-essay-scoring-2/data
ScoringLeaderboardsora-video-generation-alignment-likert-scoring
Rapidata Video Generation Prompt Alignment Dataset
If you get value from this dataset and would like to see more in the future, please consider liking it.
This dataset was collected in ~1 hour using the Rapidata Python API, accessible to anyone and ideal for large scale data annotation.
Overview
In this dataset, ~6000 human evaluators were asked to evaluate AI-generated videos based on how well the generated video matches the prompt. The specific question… See the full description on the dataset page: https://huggingface.co/datasets/Rapidata/sora-video-generation-alignment-likert-scoring.credit-scoring-training-datasetThe training dataset includes all addresses that had undertaken at least one borrow transaction on Aave v2 Ethereum or Compound v2 Ethereum any time between 7 May 2019 and 31 August 2023, inclusive (called the observation window).
Data Structure & Shape
There are almost 0.5 million observations with each representing a single borrow event. Therefore, all feature values are calculated as at the timestamp of a borrow event and represent the cumulative positions just before the borrow event's… See the full description on the dataset page: https://huggingface.co/datasets/spectrallabs/credit-scoring-training-dataset.Scoring-Verifiers
Scoring Verifiers
Scoring Verifiers is a set of 4 benchmarks that evaluate the scoring and ranking capabilities of synthetic verifiers such as test case generation and reward modelling. You can find our paper Scoring Verifiers: Evaluating Synthetic Verification for Code and Reasoning which explains in more detail our methodology, benchmark details and findings.
Datasets
In this repository, we include 4 benchmarks that are code scoring and ranking versions of… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/Scoring-Verifiers.german-credit-risk_credit-scoring_mlp
🏦 German Credit Risk - Dataset para MLP
Este dataset es parte del curso de Deep Learning impartido en el canal de YouTube de inGeniia. Se utiliza para demostrar la implementación de un Perceptrón Multicapa (MLP) para tareas de clasificación binaria (riesgo crediticio).
Descripción del Proyecto
El objetivo de este dataset es predecir si un cliente representa un buen o mal riesgo crediticio basándose en una serie de atributos financieros y personales.
Problema:… See the full description on the dataset page: https://huggingface.co/datasets/inGeniia/german-credit-risk_credit-scoring_mlp.editlens-scoring-data
Editlens Scoring Data
A collection of datasets of human writing for evaluating AI-generated text detectors.
Citations:
Crossley, S. (2025). A large-scale corpus for assessing source-based writing quality: ASAP 2.0. Zenodo. https://doi.org/10.5281/zenodo.14781349
J. Schler, M. Koppel, S. Argamon and J. Pennebaker (2006). Effects of Age and Gender on Blogging in Proceedings of 2006 AAAI Spring Symposium on Computational Approaches for Analyzing Weblogs. URL:… See the full description on the dataset page: https://huggingface.co/datasets/awdllt03/editlens-scoring-data.leadforge-lead-scoring-v1
LeadForge: Synthetic B2B Lead Scoring Dataset (leadforge-lead-scoring-v1)
A relational, reproducible, three-tier synthetic CRM dataset family for
teaching lead scoring at scale. Created by
Shay Palachy Affek and generated by
leadforge, an
open-source Python framework for synthetic CRM/funnel data. The
framework version is decoupled from the dataset version: the package
stays at 1.x; the dataset is published under the explicit …-v1
tag.
Why lead scoring matters in… See the full description on the dataset page: https://huggingface.co/datasets/shaypal5/leadforge-lead-scoring-v1.lead-scoring-x
Lead Scoring Dataset
Overview
This dataset contains lead scoring data for X Education, a company that provides online courses. The dataset is designed for binary classification to predict whether a lead will convert to a customer using an LLM.
Source: Kaggle - Lead Scoring Dataset
Target Variable: Converted (0 = Not Converted, 1 = Converted)
Features
The processed dataset includes the following 7 key features:
Prospect ID - Unique identifier for each… See the full description on the dataset page: https://huggingface.co/datasets/shawhin/lead-scoring-x.geolip-sdxl-fid-scoringspx-sustainalytics-esg-scoresdeckanalyst-scoring-methodology
DeckAnalyst Pitch Deck Scoring Methodology
Dataset Description
The complete, transparent methodology behind DeckAnalyst, an AI-powered pitch deck evaluation engine developed by Unbiased Ventures (Foxsmart Systems GmbH, Zurich). This dataset documents the scoring framework used to evaluate startup pitch decks across 8 dimensions with stage-aware weighting, hard gating rules, missing-data handling, and peer benchmarking against 6,586 companies.
Purpose… See the full description on the dataset page: https://huggingface.co/datasets/peterweisz/deckanalyst-scoring-methodology.Automated-Essay-Scoring-2.0alfa-scoring-bkihttps://ods.ai/competitions/dl-fintech-bki
hf-video-scoring
NFL Play Scoring Inference Demo
This repository demonstrates a lightweight video classification inference pipeline using PyTorchVideo's X3D-M model to score 2-second NFL play clips as part of a "start-of-play" or "end-of-play" detection task.
Overall Architecture
This inference pipeline is part of a larger AWS-based NFL play analysis system. The X3D model component (this repository) fits into the following architecture:
+---------------+ +---------------------+
|… See the full description on the dataset page: https://huggingface.co/datasets/rocket-wave/hf-video-scoring.Dataset_Automatic_Essay_Scoring_Essay-EssayScore_and_24_textual_featuresautonomous-driving-rss-traffic-flow-coherence-state-scoring-v0.1What this dataset tests
Whether a system can score traffic-flow coherence
before and after an ego action.
This is not collision detection.
It measures systemic stability.
Required outputs
pre_action_coherence_score
post_action_coherence_score
coherence_delta
shockwave_generation_flag
braking_propagation_depth
systemic_risk_score
Scoring conventions
coherence scores range 0 to 1
coherence_delta may be negative or positive
shockwave flag is 0 or 1
braking propagation depth… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/autonomous-driving-rss-traffic-flow-coherence-state-scoring-v0.1.essay-scoringunlv-olympic-scoring-datasetautomatic_essay_scoringafrica-synth-telecom-telecom-credit-scoring-data-nigeria
Africa Synth Telecom Telecom Credit Scoring Data Nigeria | Africa (Electric Sheep Africa metadata inventory)
Size category: 100K<n<1M - Formats: parquet - Sector: economics_finance - Engineered by Electric Sheep Africa
TL;DR
This dataset is part of the Electric Sheep Africa catalog on Hugging Face. It is indexed for African data discovery with standardized metadata, loading guidance, provenance notes, and analyst-oriented context.
What This Dataset… See the full description on the dataset page: https://huggingface.co/datasets/electricsheepafrica/africa-synth-telecom-telecom-credit-scoring-data-nigeria.assets-credit-scoring-mlopsarabic-cv-scoring-dataset
Arabic CV Scoring Dataset
Dataset Summary
This dataset contains ~7,220 synthetically generated Arabic CVs, each paired
with a job category, an ATS (Applicant Tracking System) compatibility score,
and a suitability score/class label. It was built to train and evaluate the
Arabic CV Analyzer —
an NLP pipeline that scores, classifies, and generates improvement suggestions
for Arabic CVs targeting the Arab job market, where no equivalent
ATS-optimization tooling… See the full description on the dataset page: https://huggingface.co/datasets/omaraboelmaaty/arabic-cv-scoring-dataset.essay-scoringThis is the link to the github repo used to create an Essay-scoring LLM.
https://github.com/RSDP101/Essay-scoring-LLM
task_categories:
- text-classification
- text-retrieval
language:
- en
tags:
- text-classification
- essays
- score
- scores
- essay
- nlp
size_categories:
- n<1K
