datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
KarantaOCR-Bench
KarantaOCR - Bench
KarantaOCR-Bench is a unit-test–style evaluation dataset, similar to olmOCR-bench. It consists of 70 PDF documents and 300 test cases spanning multiple document types.
All the tests were manually verified by us.
KarantaOCR-Bench is designed specifically to evaluate document text extraction for Documents with diacritics and special characters, covering a diverse range of document formats and languages commonly under-represented in existing OCR benchmarks.… See the full description on the dataset page: https://huggingface.co/datasets/taresco/KarantaOCR-Bench.tarih-ders-kitaplari-1931
Tarih I ve Tarih II (1931) — makine-okunabilir tam metin
Türk Tarihi Tetkik Cemiyeti tarafından hazırlanıp Maarif Vekâleti emriyle basılan ve 1931-1941 arasında Türkiye'de liselerde resmî tarih ders kitabı olarak okutulan Tarih I (Tarihtenevelki Zamanlar ve Eski Zamanlar) ile Tarih II (Ortazamanlar) ciltlerinin tam metni.
Her basılı sayfa ayrı bir kayıttır: kalıcı kimliği, hazır künyesi ve kaynak
taramadaki tam konumu vardır. Bir iddia basılı sayfaya kadar izlenebilir.… See the full description on the dataset page: https://huggingface.co/datasets/asayimusa19/tarih-ders-kitaplari-1931.tarotoo-tarot-card-meanings
Tarotoo Tarot Card Meanings
A complete, structured dataset of all 78 tarot cards (22 Major Arcana + 56 Minor Arcana) in the Rider–Waite–Smith tradition. Published by Tarotoo. These are the card meanings that ground the AI-generated readings on Tarotoo.com.
Dataset details
Curated by: Tarotoo (tarotoo.com)
Language: English
License: MIT
Rows: 78 (one per card) · Fields: 22
DOI (Zenodo, cite this): 10.5281/zenodo.21514483
Concept DOI (Zenodo, always resolves to the… See the full description on the dataset page: https://huggingface.co/datasets/Tarotoo/tarotoo-tarot-card-meanings.tarkov-customs-trajectory
CUSTOMS — Fable 5.1 agent trajectory
Full Claude Code session log (raw JSONL trajectory) of Fable 5.1 (claude-fable-5) building
tarkov-customs — a single-file Tarkov-style
extraction shooter — end to end: engine, AI, ballistics, UI, and tests.
Play: https://inonono66.github.io/tarkov-customs/
Code: https://github.com/INONONO66/tarkov-customs
File: trajectory/5881b249-26a1-4c50-a720-1cee6e722991.jsonl (509 events)
swev-trm-trajectories-25models
SWE-Bench Verified TRM Trajectories (25 Models, Verified Labels)
Trajectories from 25 LLMs attempting SWE-Bench Verified tasks, formatted for
training a Trajectory Reward Model (TRM). Each record is one model's full
multi-turn attempt at one task, labeled with the real SWE-bench harness
verdict (scores.resolved).
Splits
Split
Records
Tasks
Pos
Neg
train
10,107
405
6,108
3,999
val
2,366
95
1,440
926
Train/val are task-disjoint (stable hash on task_id… See the full description on the dataset page: https://huggingface.co/datasets/tarsur385/swev-trm-trajectories-25models.tarot-rws-historical-meanings
StarTarot English RWS Historical Meanings and Golden Dawn Correspondences
Version 1.1.0 · English · 78 cards · DOI: https://doi.org/10.5281/zenodo.22917681 ·
all versions: https://doi.org/10.5281/zenodo.21381779
Project website ·
Dataset documentation ·
Tarot card guide ·
Tarot spreads
Summary
This dataset gives one structured record for each of the 78 cards of the
Rider–Waite–Smith (RWS) tarot deck. Each record combines:
A. E. Waite's divinatory meanings from… See the full description on the dataset page: https://huggingface.co/datasets/StarTarotOnline/tarot-rws-historical-meanings.uk-pods
uk-pods - speech datasets of Ukrainian podcasts.
Preparation
Clone the dataset repository and extract the content of clips.tar.gz archive.
git clone https://huggingface.co/datasets/taras-sereda/uk-pods
cd uk-pods && tar -zxvf clips.tar.gz
To use these manifests for training/inference with NeMo [1] modify audio_filepath to absolute locations of audio files extracted in previous step.
# data_root=<clonned_repo_dir> # /home/ubuntu/uk-pods
data_root=$(realpath .)
sed -i… See the full description on the dataset page: https://huggingface.co/datasets/taras-sereda/uk-pods.tartanaviation-adsb-19k-clean
TartanAviation ADS-B Dataset (19.7K Clean Samples)
Dataset Description
19,714 high-quality ADS-B trajectory datapoints from aircraft operations, rigorously cleaned and validated. Perfect for machine learning research in aviation, reinforcement learning, and trajectory prediction.
Key Features
19,714 clean samples (no missing data, no duplicates)
17 comprehensive features including aircraft ID, timestamp components, altitude, speed, heading, geolocation, and… See the full description on the dataset page: https://huggingface.co/datasets/SANIKKI/tartanaviation-adsb-19k-clean.tartanaviation-adsb-19k-clean
TartanAviation ADS-B Dataset (19.7K Clean Samples)
Dataset Description
19,714 high-quality ADS-B trajectory datapoints from aircraft operations, rigorously cleaned and validated. Perfect for machine learning research in aviation, reinforcement learning, and trajectory prediction.
Key Features
19,714 clean samples (no missing data, no duplicates)
17 comprehensive features including aircraft ID, timestamp components, altitude, speed, heading, geolocation, and… See the full description on the dataset page: https://huggingface.co/datasets/Pathange/tartanaviation-adsb-19k-clean.swe-verified-gemini3-flash-trajectories
SWE-bench Verified — Gemini-3-flash agent trajectories (graded, 3 samples/instance)
Agent trajectories from gemini-3-flash-preview (high reasoning, temperature 0.8) run with the
OpenHands agent on SWE-bench Verified, in Modal sandboxes. For each of 100 instances
we sampled multiple trajectories and graded them with the SWE-bench harness; this dataset holds the
3 graded samples per instance = 296 trajectories, 198 resolved (67%).
pass@1 ≈ 66/100, pass@3 (oracle) = 73/100.… See the full description on the dataset page: https://huggingface.co/datasets/tarsur385/swe-verified-gemini3-flash-trajectories.EstCOPA
Estonian Choice of Plausible Alternatives (EstCOPA)
Dataset Summary
EstCOPA is an extended version of XCOPA that was created with a goal to further investigate Estonian language understanding of large language models. EstCOPA provides two new versions of train, eval and test datasets in Estonian: firstly, a machine translated (En->Et) version of original English COPA (Roemmele et al., 2011) and secondly, a manually post-edited version of the same machine translated data.… See the full description on the dataset page: https://huggingface.co/datasets/tartuNLP/EstCOPA.zh-bo-instructtartanaviation-adsb-19k-clean
TartanAviation ADS-B Dataset (19.7K Clean Samples)
Dataset Description
19,714 high-quality ADS-B trajectory datapoints from aircraft operations, rigorously cleaned and validated. Perfect for machine learning research in aviation, reinforcement learning, and trajectory prediction.
Key Features
19,714 clean samples (no missing data, no duplicates)
17 comprehensive features including aircraft ID, timestamp components, altitude, speed, heading… See the full description on the dataset page: https://huggingface.co/datasets/RunningCFOP/tartanaviation-adsb-19k-clean.tartanaviation-adsb-19k-clean
TartanAviation ADS-B Dataset (19.7K Clean Samples)
Dataset Description
19,714 high-quality ADS-B trajectory datapoints from aircraft operations, rigorously cleaned and validated. Perfect for machine learning research in aviation, reinforcement learning, and trajectory prediction.
Key Features
19,714 clean samples (no missing data, no duplicates)
17 comprehensive features including aircraft ID, timestamp components, altitude, speed, heading, geolocation, and… See the full description on the dataset page: https://huggingface.co/datasets/Gogul001/tartanaviation-adsb-19k-clean.teleop_7hr_tardrugbank_drug_target_label_mapping_amino_acid_pairtarotoo-tarot-card-meanings
Tarotoo Tarot Card Meanings
A complete, structured dataset of all 78 tarot cards (22 Major Arcana + 56 Minor Arcana) in the Rider–Waite–Smith tradition. Published by Tarotoo. These are the card meanings that ground the AI-generated readings on Tarotoo.com.
Dataset details
Curated by: Tarotoo (tarotoo.com)
Language: English
License: MIT
Rows: 78 (one per card) · Fields: 22
DOI (Zenodo, cite this): 10.5281/zenodo.21514483
Concept DOI (Zenodo, always resolves to the… See the full description on the dataset page: https://huggingface.co/datasets/Clouds4days/tarotoo-tarot-card-meanings.adaption-mission-target-pairs
This dataset is a remastered version prepared using Adaption's Adaptive Data platform.
adaption-mission_target_pairs
This dataset consists of prompt-completion pairs mapping unique mission identifiers to specific celestial bodies or astronomical targets. The prompts follow a 'Mission-XXX' format, while the completions include planets, moons, dwarf planets, and stars such as Mars, Europa, and Betelgeuse. It is structured for training or evaluating models on space mission target… See the full description on the dataset page: https://huggingface.co/datasets/Charley890/adaption-mission-target-pairs.gamma-g1-334-vast-g1-333-targeted9-semantic-gate-20260625Tarek07__Thalassic-Alpha-LLaMa-70B-details
Dataset Card for Evaluation run of Tarek07/Thalassic-Alpha-LLaMa-70B
Dataset automatically created during the evaluation run of model Tarek07/Thalassic-Alpha-LLaMa-70B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Tarek07__Thalassic-Alpha-LLaMa-70B-details.Tarek07__Progenitor-V1.1-LLaMa-70B-details
Dataset Card for Evaluation run of Tarek07/Progenitor-V1.1-LLaMa-70B
Dataset automatically created during the evaluation run of model Tarek07/Progenitor-V1.1-LLaMa-70B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Tarek07__Progenitor-V1.1-LLaMa-70B-details.Sakalti__tara-3.8B-details
Dataset Card for Evaluation run of Sakalti/tara-3.8B
Dataset automatically created during the evaluation run of model Sakalti/tara-3.8B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Sakalti__tara-3.8B-details.BBQ_target_locgamma-g1-331-vast-g1-326-targeted9-direct-port-first-gate-20260625Sakalti__Tara-3.8B-v1.1-details
Dataset Card for Evaluation run of Sakalti/Tara-3.8B-v1.1
Dataset automatically created during the evaluation run of model Sakalti/Tara-3.8B-v1.1
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Sakalti__Tara-3.8B-v1.1-details.crisis-negotiation-dataTartan-Explore
CMU Landmarks Dataset
Dataset Description
This dataset contains a curated collection of 100+ Carnegie Mellon University landmarks, including their names, categories, geographic coordinates, ratings, dwell times, and indoor/outdoor classifications. It serves as the primary data source for the CMU Explorer ML application, enabling features like content-based recommendations, rating prediction, and route optimization.
Dataset Details
Source
The… See the full description on the dataset page: https://huggingface.co/datasets/ysakhale/Tartan-Explore.
