datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
hack-ignition-benchmark
hack-ignition benchmark — data, v0.1.6
Training trajectories of reinforcement-learning runs on exploitable graders, for studying and predicting when RL
comes to produce exploits. Each family is a set of GRPO runs over configurations of (start model, prompt,
training set, grader / reward structure, recipe), with one or more seeds per configuration. Every family stores
what its training logs contain — per-step exploit, task and reward rates, the item × step exploit record… See the full description on the dataset page: https://huggingface.co/datasets/EleutherAI/hack-ignition-benchmark.deep-ignorance-pretraining-mix
Deep Ignorance Model Suite
We explore an intuitive yet understudied question: Can we prevent LLMs from learning unsafe technical capabilities (such as CBRN) by filtering out enough of the relevant pretraining data before we begin training a model? Research into this question resulted in the Deep Ignorance Suite. In our experimental setup, we find that filtering pretraining data prevents undesirable knowledge, doesn't sacrifice general performance, and results in models that are… See the full description on the dataset page: https://huggingface.co/datasets/EleutherAI/deep-ignorance-pretraining-mix.genesis-1k-ignorekanari-wildfire-ignitions
kanari — worldwide wildfire ignitions archive
Continuously updated archive of significant wildfire events worldwide (136 countries), produced by
kanari, a free near-real-time map of wildfire ignitions.
Each row is one fire event: satellite hotspots from NASA FIRMS (VIIRS 375 m), NOAA GOES and
EUMETSAT Meteosat MTG are clustered into events (≈4 km cells); the first detection is the proxy for
ignition time. Public witness reports (Bluesky, press via GDELT, Telegram) are geoparsed… See the full description on the dataset page: https://huggingface.co/datasets/expansia/kanari-wildfire-ignitions.seal_ur5_ignore-lolThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "ur5",
"total_episodes": 30,
"total_frames": 7620,
"total_tasks": 2,
"total_videos": 60,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 20,
"splits": {
"train": "0:30"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/iantc104/seal_ur5_ignore-lol.ign-earthquakes
IGN Spanish Earthquake Catalogue
This dataset is a research-ready snapshot of the official earthquake catalogue maintained by the Spanish Instituto Geografico Nacional (IGN). It contains 207,222 seismic events from 3 March 1373 through 12 August 2026 within the catalogue search area used by IGN for Spain, the Canary Islands, and nearby regions.
The data is observational and has no target label or predefined classes. The complete table is provided as one train split because… See the full description on the dataset page: https://huggingface.co/datasets/hsilvosa/ign-earthquakes.ignorance-classifier-testing-datalm-eval-EleutherAI_tampered-deep-ignorance-random-init-fp-adversarial-20251104_051748
Dataset Card for Evaluation run of EleutherAI/tampered-deep-ignorance-random-init-fp-adversarial-20251104_051748
Dataset automatically created during the evaluation run of model EleutherAI/tampered-deep-ignorance-random-init-fp-adversarial-20251104_051748
The dataset is composed of 2 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 11 run(s). Each run can be found as a specific split in each configuration, the split being… See the full description on the dataset page: https://huggingface.co/datasets/EleutherAI/lm-eval-EleutherAI_tampered-deep-ignorance-random-init-fp-adversarial-20251104_051748.Yuma42__Llama3.1-IgneousIguana-8B-details
Dataset Card for Evaluation run of Yuma42/Llama3.1-IgneousIguana-8B
Dataset automatically created during the evaluation run of model Yuma42/Llama3.1-IgneousIguana-8B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Yuma42__Llama3.1-IgneousIguana-8B-details.so101_cup_eraser_target_v1_10ep
so101_cup_eraser_target_v1_10ep
LeRobot dataset snapshot.
Source dataset repo: ignite-dohyun/so101_cup_eraser_target_v1
Snapshot episodes: 10
Snapshot frames: 12068
Robot type: bi_so_follower
FPS: 30
This repo is intended as an immutable training snapshot. Continue recording into
the source dataset repo, and create new snapshot repos for later episode counts.
so101_cube_50mm_mac_pilotThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "so_follower",
"total_episodes": 1,
"total_frames": 600,
"total_tasks": 1,
"chunks_size": 1000,
"data_files_size_in_mb": 100,
"video_files_size_in_mb": 200,
"fps": 30,
"splits": {
"train": "0:1"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/ignite-dohyun/so101_cube_50mm_mac_pilot.lm-eval-EleutherAI_deep-ignorance-unfiltered
Dataset Card for Evaluation run of EleutherAI/deep-ignorance-unfiltered
Dataset automatically created during the evaluation run of model EleutherAI/deep-ignorance-unfiltered
The dataset is composed of 1 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/EleutherAI/lm-eval-EleutherAI_deep-ignorance-unfiltered.ttz-redblock-yellowbox-old-newgripper-merged-ignorecodec-04This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "ned2",
"total_episodes": 138,
"total_frames": 39340,
"total_tasks": 3,
"chunks_size": 1000,
"data_files_size_in_mb": 100,
"video_files_size_in_mb": 200,
"fps": 21,
"splits": {
"train": "0:138"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/AlexanderRoempke/ttz-redblock-yellowbox-old-newgripper-merged-ignorecodec-04.so101_red_cube_5cm_blue_square_v1This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "so_follower",
"total_episodes": 70,
"total_frames": 64104,
"total_tasks": 1,
"chunks_size": 1000,
"data_files_size_in_mb": 100,
"video_files_size_in_mb": 200,
"fps": 30,
"splits": {
"train": "0:70"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/ignite-dohyun/so101_red_cube_5cm_blue_square_v1.DreadPoor__TEST03-ignore-details
Dataset Card for Evaluation run of DreadPoor/TEST03-ignore
Dataset automatically created during the evaluation run of model DreadPoor/TEST03-ignore
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DreadPoor__TEST03-ignore-details.DreadPoor__TEST06-ignore-details
Dataset Card for Evaluation run of DreadPoor/TEST06-ignore
Dataset automatically created during the evaluation run of model DreadPoor/TEST06-ignore
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DreadPoor__TEST06-ignore-details.DreadPoor__TEST02-Ignore-details
Dataset Card for Evaluation run of DreadPoor/TEST02-Ignore
Dataset automatically created during the evaluation run of model DreadPoor/TEST02-Ignore
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DreadPoor__TEST02-Ignore-details.DreadPoor__TEST07-ignore-details
Dataset Card for Evaluation run of DreadPoor/TEST07-ignore
Dataset automatically created during the evaluation run of model DreadPoor/TEST07-ignore
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DreadPoor__TEST07-ignore-details.DreadPoor__TEST08-ignore-details
Dataset Card for Evaluation run of DreadPoor/TEST08-ignore
Dataset automatically created during the evaluation run of model DreadPoor/TEST08-ignore
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DreadPoor__TEST08-ignore-details.lm-eval-EleutherAI_deep-ignorance-random-init
Dataset Card for Evaluation run of EleutherAI/deep-ignorance-random-init
Dataset automatically created during the evaluation run of model EleutherAI/deep-ignorance-random-init
The dataset is composed of 0 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/EleutherAI/lm-eval-EleutherAI_deep-ignorance-random-init.spydr-pockiris-claseasl
