datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
nucleotide_transformer_downstream_tasks_revised
Dataset Card for Dataset Name
The nucleotide_transformer_downstream_tasks dataset features the 18 downstream tasks presented in the Nucleotide Transformer paper. They consist of both binary and multi-class classification tasks that aim at providing a consistent genomics benchmark.
We note that this is an updated version of this benchmark after the paper has been through peer-review. We highly encourage to move to this version in detriment of the older version.Keypoints about the… See the full description on the dataset page: https://huggingface.co/datasets/InstaDeepAI/nucleotide_transformer_downstream_tasks_revised.2026-09-14-dataset-refresh-revised-pilot-audit
Failed pilots for moral low-stakes and nonmoral craft advice refresh; audit evidence only
field
value
experiment
Failed pilots for moral low-stakes and nonmoral craft advice refresh; audit evidence only
date_generated
20260914_230322
constitution
constitutions/claude_distilled_09_principles/constitution.md; low-stakes principle generation, nonmoral compatibility review only
source_repo
https://github.com/Matthew-Bozoukov/Lessons_from_constituitional_AFT @… See the full description on the dataset page: https://huggingface.co/datasets/dougalldeepmind/2026-09-14-dataset-refresh-revised-pilot-audit.bcp-plan-revisions-v1
bcp-plan-revisions-v1
Plan text revision history across 8 planning conditions on BrowseComp-Plus (first 50 queries, Qwen3-Embedding-8B retriever). Each row is one revision entry from plan_text_history in the trajectory metadata.
Dataset Info
Rows: 2063
Columns: 8
Columns
Column
Type
Description
condition
Value('string')
Run condition name (planning version + optional gemini start extension)
run_id
Value('string')
Trajectory filename (without… See the full description on the dataset page: https://huggingface.co/datasets/timchen0618/bcp-plan-revisions-v1.sweep-single-revised-v3This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "xarm",
"total_episodes": 491,
"total_frames": 436429,
"total_tasks": 1,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 50,
"splits": {
"train": "0:491"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path": null,
"features": {… See the full description on the dataset page: https://huggingface.co/datasets/gozdebaydogmus/sweep-single-revised-v3.UTK-Face-Revised
Dataset Card for "UTK-Face-Revised"
More Information needed
REVISOR-25k
REVISOR-25k
A multi-task video understanding dataset for training video LLMs with reinforcement learning (GRPO). The dataset contains ~25k samples spanning Video QA and Temporal Grounding tasks.
Dataset Structure
The dataset is organized into 4 subsets:
Subset
Task
Samples
Description
video_r1
Video QA
20,855
Multiple-choice video question answering
time_r1
Temporal Grounding
2,500
Locate time intervals in videos
cg_bench
Temporal Grounding
1,167… See the full description on the dataset page: https://huggingface.co/datasets/williamljz/REVISOR-25k.code-review-instruct-critique-revision
Dataset Card for "code-review-instruct-critique-revision"
More Information needed
helical_dna_smaller_randomization_revisePAD3-Dataset-Revisi
Dataset Card for Dataset Name
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Dataset Details
Dataset Description
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Dataset Sources [optional]… See the full description on the dataset page: https://huggingface.co/datasets/capstone-pad3/PAD3-Dataset-Revisi.revisitopBioASQ-Task-B-Revisede3c-sentences-EN-revisedThe files in this repo fall under two extensions:
.tsv files, which contain the BIO files for clinical entities
.txt files, which contain the PubTator version for relation extraction
code-review-instruct-critique-revision-pythonResearchArcade-openreview-revisionsNavQA_Revised
NavQA Revised
NavQA Revised is a re-annotated version of the NaVQA dataset released with NVIDIA ReMEmbR. The original annotations are distributed in remembr/data/navqa/data.csv.
This repository provides the revised annotations as JSONL files:
navqa.jsonl
navqa_over_sequence.jsonl
Each line is one question-answer example. navqa.jsonl is the primary revised annotation file. navqa_over_sequence.jsonl uses the same schema and examples, but uses the beginning of the full sequence as… See the full description on the dataset page: https://huggingface.co/datasets/DoneHans/NavQA_Revised.NavQA_Revised
NavQA Revised
NavQA Revised is a re-annotated version of the NaVQA dataset released with NVIDIA ReMEmbR. The original annotations are distributed in remembr/data/navqa/data.csv.
This repository provides the revised annotations as JSONL files:
navqa.jsonl
navqa_over_sequence.jsonl
Each line is one question-answer example. navqa.jsonl is the primary revised annotation file. navqa_over_sequence.jsonl uses the same schema and examples, but uses the beginning of the full sequence as the… See the full description on the dataset page: https://huggingface.co/datasets/Qiuchen-Wang/NavQA_Revised.repro-revisiting-zeroth-order-hessian-approximation-policy-lens-traces
Agent traces
Agent sessions published from a Trackio Logbook.
repro-revisiting-zeroth-order-hessian-approximation-traces
Agent traces
Agent sessions published from a Trackio Logbook.
sdsd-revisions
Self Directed Synthetic Dialogues (SDSD) Revisions v0
This dataset is an experiment in procedurally generating synthetic dialogues and revisions between two language models, along the lines of Constitutional AI.
For each dialogue, one model, acting as a "user" generates a plan based on a topic, subtopic, and goal for a conversation.
Next, this model attempts to act on this plan and generating synthetic data.
Along with the plan is a principle which the model, in some successful… See the full description on the dataset page: https://huggingface.co/datasets/allenai/sdsd-revisions.32B-predict-revise-v2.rule-r-1.0-k-8.L-512.statml-arxi2026-09-10-nonmoral-grounded-revision-pilot-audit
Grounded nonmoral full-response revision: two bounded pilots; candidate stopped before production
field
value
experiment
Grounded nonmoral full-response revision: two bounded pilots; candidate stopped before production
date_generated
2026-09-10
constitution
none applied to model prompts; nonmoral preferences file required as a SynthDoc container only
source_repo
https://github.com/Matthew-Bozoukov/teaching_claude_why_replication.git @… See the full description on the dataset page: https://huggingface.co/datasets/dougalldeepmind/2026-09-10-nonmoral-grounded-revision-pilot-audit.visual_genome_revisede3c-sentences-PL-revisedGraphWiz-Revisedflight-mh370-revisited-data
Flight MH370 Revisited — oversized source files
Companion data for the research repository
gmkf7vfyfb-web/flight-mh370-revisited.
That repository holds the complete project tree — model code, source data,
reports, figures, posterior outputs, the 148-file source-library audit and the
handoff dossier — from the consolidated snapshot of 14 August 2026. Seven files
were too large to keep in Git (one is 336 MB, above GitHub's hard 100 MB
per-file limit), so they live here instead.… See the full description on the dataset page: https://huggingface.co/datasets/peteabiome/flight-mh370-revisited-data.ResearchArcade-openreview-papers-revisionse3c-sentences-IT-revisedThe files in this repo fall under two extensions:
.tsv files, which contain the BIO files for clinical entities
.txt files, which contain the PubTator version for relation extraction
e3c-sentences-GR-revisede3c-sentences-SK-revisedrevision_data_split_0_translated
Dataset Card for "revision_data_split_0_translated"
More Information needed
