datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
geometry-dash-levels
Geometry Dash Level Dataset
Subsets
2024_300k
Dump of ~300k levels from the Geometry Dash servers, sorted by the number of likes. 66 JSONL shards (~660MB each, ~43GB total).
Files: 2024_300k/levels-v1-00000.jsonl through 2024_300k/levels-v1-00065.jsonl
2026_50k_rated
~50k rated/featured levels scraped from the Geometry Dash servers in February 2026. 39 JSONL shards (~500MB each, ~19GB total).
Files: 2026_50k_rated/levels-v2-00000.jsonl through… See the full description on the dataset page: https://huggingface.co/datasets/yusp48/geometry-dash-levels.llamascope2-dashboardD_ASAP-AES
D_ASAP-AES
This is the train, test, and validation split of the ASAP Automated Essay Scoring dataset,
prepared for use with the S-GRADES benchmark.
Ground truth labels have been removed to prevent leakage during evaluation.
For the original dataset with labels, see below.
Original Dataset
🔗 ASAP-AES on Kaggle
Citation
If you use this dataset, please cite the original:
@misc{asap_aes,
title={ASAP Automated Essay Scoring}… See the full description on the dataset page: https://huggingface.co/datasets/nlpatunt/D_ASAP-AES.lm-eval-results-abideen-AlphaMonarch-daser-private
Dataset Card for Evaluation run of abideen/AlphaMonarch-daser
Dataset automatically created during the evaluation run of model abideen/AlphaMonarch-daser
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-abideen-AlphaMonarch-daser-private.geometry-dash-retro-levelsFork of https://huggingface.co/datasets/yusp48/geometry-dash-levels.
Contains only retro levels with id < 11000000.
Use my gdparse library: pip install gdparse
quants
QuAnTS: Question Answering on Time Series
QuAnTS is a challenging dataset designed to bridge the gap in question-answering research on time series data.
The dataset features a wide variety of questions and answers concerning human movements, presented as tracked skeleton trajectories.
QuAnTS also includes human reference performance to benchmark the practical usability of models trained on this dataset.
At present, there is no official leaderboard for this dataset.
Dataset… See the full description on the dataset page: https://huggingface.co/datasets/dasyd/quants.DAS-0910-1This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"fps": 30,
"features": {
"observation.imu": {
"dtype": "float32",
"shape": [
6
],
"names": [
"angular_velocity.x",
"angular_velocity.y",
"angular_velocity.z",
"linear_acceleration.x",
"linear_acceleration.y"… See the full description on the dataset page: https://huggingface.co/datasets/fza1796262052/DAS-0910-1.lelab-test_20260906_162149This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"fps": 30,
"features": {
"action": {
"dtype": "float32",
"shape": [
6
],
"names": [
"shoulder_pan.pos",
"shoulder_lift.pos",
"elbow_flex.pos",
"wrist_flex.pos",
"wrist_roll.pos",
"gripper.pos"… See the full description on the dataset page: https://huggingface.co/datasets/dasfas1132/lelab-test_20260906_162149.pd-discovery-benchmark-dashboard
Parkinson's Disease Discovery Benchmark Dashboard
Reusable benchmark, knowledge graph, manuscript resource, and Streamlit dashboard for Parkinson's disease target-to-intervention discovery.
This repository integrates evidence-synthesis priority scores, target tractability, omics/pathway recurrence, ChEMBL compound activity, RDKit physicochemical heuristics, Human Protein Atlas cell-type context, iPSC/stem-cell validation mappings, and publication-ready figures.… See the full description on the dataset page: https://huggingface.co/datasets/hssling/pd-discovery-benchmark-dashboard.DasanCallDial
DasanCallDial
DasanCallDial is the first large-scale Korean benchmark built specifically for
dialogue-level ASR error correction. It contains 1,974 real civil-complaint call
dialogues (115,460 utterances) placed to the 120 Dasan Call Foundation, Seoul's
municipal civic-information hotline. Each utterance pairs the transcription produced by a
production speech-recognition system with a human-verified ground truth.
Unlike datasets built by injecting synthetic noise, DasanCallDial… See the full description on the dataset page: https://huggingface.co/datasets/zgold5670/DasanCallDial.DAS-Bench
DAS-Bench
DAS-Bench is a 30-topic multi-domain benchmark for automatically generated academic surveys. It is paired with DAS-Eval, a 16-criterion evaluation suite for publication-oriented academic surveys covering scholarly citation, taxonomic synthesis, hierarchical discourse, and manuscript reliability.
Paper: Deep Academic Survey
Project page: DAS
Source and evaluation toolkit: ZhikaiXu24/DAS
Literature metadata lake: ZhikaiXu24/DAS-2M
This repository is the Hugging Face… See the full description on the dataset page: https://huggingface.co/datasets/ZhikaiXu24/DAS-Bench.DAS-0910This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"fps": 30,
"features": {
"observation.imu": {
"dtype": "float32",
"shape": [
6
],
"names": [
"angular_velocity.x",
"angular_velocity.y",
"angular_velocity.z",
"linear_acceleration.x",
"linear_acceleration.y"… See the full description on the dataset page: https://huggingface.co/datasets/fza1796262052/DAS-0910.trace-dashed-lineThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"fps": 30,
"features": {
"action": {
"dtype": "float32",
"names": [
"shoulder_pan.pos",
"shoulder_lift.pos",
"elbow_flex.pos",
"wrist_flex.pos",
"wrist_roll.pos",
"gripper.pos"
],
"shape": [
6… See the full description on the dataset page: https://huggingface.co/datasets/Yunis147/trace-dashed-line.DAS-0910-2This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"fps": 30,
"features": {
"observation.imu": {
"dtype": "float32",
"shape": [
6
],
"names": [
"angular_velocity.x",
"angular_velocity.y",
"angular_velocity.z",
"linear_acceleration.x",
"linear_acceleration.y"… See the full description on the dataset page: https://huggingface.co/datasets/fza1796262052/DAS-0910-2.KoMultiText
KoMultiText: Korean Multi-task Dataset for Classifying Biased Speech
Dataset Summary
KoMultiText is a comprehensive Korean multi-task text dataset designed for classifying biased and harmful speech in online platforms. The dataset focuses on tasks such as Preference Detection, Profanity Identification, and Bias Classification across multiple domains, enabling state-of-the-art language models to perform multi-task learning for socially responsible AI applications.… See the full description on the dataset page: https://huggingface.co/datasets/Dasool/KoMultiText.D_ASAP-SAS
D_ASAP-SAS
This is the train, test, and validation split of the ASAP Short Answer Scoring dataset, prepared for use with the S-GRADES benchmark. Ground truth labels have been removed to prevent leakage during evaluation.
For the original dataset with labels, see below.
Original Dataset
🔗 ASAP-SAS on Kaggle
Citation
If you use this dataset, please cite the original:
@misc{asapsas2012,
author={Barbara and Hamner, Ben and Morgan, Jaison and lynnvandev and… See the full description on the dataset page: https://huggingface.co/datasets/nlpatunt/D_ASAP-SAS.mqtt_robot_demo_with_camThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "mqtt_robot",
"total_episodes": 1,
"total_frames": 1282,
"total_tasks": 1,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:1"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/dashuai2025/mqtt_robot_demo_with_cam.trace-dashed-line_20260820_141731This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"fps": 30,
"features": {
"action": {
"dtype": "float32",
"names": [
"shoulder_pan.pos",
"shoulder_lift.pos",
"elbow_flex.pos",
"wrist_flex.pos",
"wrist_roll.pos",
"gripper.pos"
],
"shape": [
6… See the full description on the dataset page: https://huggingface.co/datasets/Yunis147/trace-dashed-line_20260820_141731.DashboardQA
DashboardQA: Benchmarking Multimodal Agents for Question Answering on Interactive Dashboards
🤗Dataset | 🖥️Code | 📄Paper
The abstract of the paper states that:
Dashboards are powerful visualization tools for data-driven decision-making, integrating multiple interactive views that allow users to explore, filter, and navigate data. Unlike static charts, dashboards support rich interactivity, which is essential for uncovering insights in real-world analytical workflows. However… See the full description on the dataset page: https://huggingface.co/datasets/ahmed-masry/DashboardQA.running-dashboard-dataD_ASAP_plus_plus
D_ASAP_plus_plus
This is the train, test, and validation split of the ASAP++ dataset, prepared for use with the S-GRADES benchmark. Ground truth labels have been removed to prevent leakage during evaluation.
Original Dataset
ASAP++ enriches the original ASAP dataset with attribute-specific essay scores (content, organization, style, etc.).
🔗 ASAP++ Official Page
Citation
If you use this dataset, please cite the original:
@inproceedings{mathias2018asap++… See the full description on the dataset page: https://huggingface.co/datasets/nlpatunt/D_ASAP_plus_plus.gemma-dashboard-session-redacted
Redacted Codex Session Export
Source session:
019ead55-e227-7f81-bd3b-3f3dad3ba5e4
Source file:
/root/.codex/sessions/2026/06/09/rollout-2026-06-09T17-02-27-019ead55-e227-7f81-bd3b-3f3dad3ba5e4.jsonl
Redactions applied:
HF access tokens
HF username / identity strings
SSH key path
raw IP addresses
The session content and timestamps are otherwise preserved.
cwe-eval-dashboard-data
cwe-eval-dashboard-data
Backing data for the CWE eval dashboard Space: per-(dataset, prompt-style) rollouts + verdicts. Not a training dataset.
hf-coding-tools-dashboard-v2
HuggingFace AI Coding Tools Dashboard (Enhanced)
Enhanced benchmark data from the HuggingFace AI Dashboard — includes query metadata (query_set, intent), run metadata (run_name, run_date), and freshness flags for stale references.
This is the v2 enhanced dataset. The original dataset is at davidkling/hf-coding-tools-dashboard.
Dataset Structure
Split
Description
Rows
results
Enhanced results with query/run metadata and freshness flags
9146
queries… See the full description on the dataset page: https://huggingface.co/datasets/davidkling/hf-coding-tools-dashboard-v2.hf-coding-tools-dashboard-all
HuggingFace AI Coding Tools Dashboard
Benchmark data from the HuggingFace AI Dashboard — tracking how AI coding tools (Claude Code, Codex, Copilot, Cursor) recommend HuggingFace products across 32 developer categories.
Dataset Structure
Split
Description
Rows
results
Full benchmark results with LLM responses, cost, tokens, latency, and product detection
9603
queries
Benchmark query definitions across 32 categories
404
runs
Run metadata and tool/model… See the full description on the dataset page: https://huggingface.co/datasets/davidkling/hf-coding-tools-dashboard-all.hf-coding-tools-dashboard
HuggingFace AI Coding Tools Dashboard
Benchmark data from the HuggingFace AI Dashboard — tracking how AI coding tools (Claude Code, Codex, Copilot, Cursor) recommend HuggingFace products across 32 developer categories.
Dataset Structure
Split
Description
Rows
results
Full benchmark results with LLM responses, cost, tokens, latency, and product detection
9146
queries
Benchmark query definitions across 32 categories
404
runs
Run metadata and tool/model… See the full description on the dataset page: https://huggingface.co/datasets/davidkling/hf-coding-tools-dashboard.africa-zimbabwe-unicef-southern-africa-dashboard-situation-and-response-1-59a7f97a
UNICEF Southern Africa Dashboard Situation and Response 1 January - 30 June 2017 | Africa (Zimbabwe official open data)
471 rows - 1 Africa country - 2016-2017 - Repackaged by Electric Sheep Africa
TL;DR
This dataset packages one official XLSX resource from Zimbabwe as
ML-ready Parquet. The source file is the provenance boundary; all usable
indicators or tabular columns from the resource stay together in this repo.
About the source
Source:… See the full description on the dataset page: https://huggingface.co/datasets/electricsheepafrica/africa-zimbabwe-unicef-southern-africa-dashboard-situation-and-response-1-59a7f97a.hf-coding-tools-dashboard-run-april12
HuggingFace AI Coding Tools Dashboard
Benchmark data from the HuggingFace AI Dashboard — tracking how AI coding tools (Claude Code, Codex, Copilot, Cursor) recommend HuggingFace products across 32 developer categories.
Dataset Structure
Split
Description
Rows
results
Full benchmark results with LLM responses, cost, tokens, latency, and product detection
8875
queries
Benchmark query definitions across 32 categories
263
runs
Run metadata and tool/model… See the full description on the dataset page: https://huggingface.co/datasets/davidkling/hf-coding-tools-dashboard-run-april12.mqtt_robot_with_01camThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "mqtt_robot",
"total_episodes": 2,
"total_frames": 2565,
"total_tasks": 1,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:2"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/dashuai2025/mqtt_robot_with_01cam.trackio-dashboard-dataset
