datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
pdf_science_questions_verifiable_r1_traces__2_24_25
Dataset card for pdf_science_questions_verifiable_r1_traces__2_24_25
This dataset was made with Curator.
Dataset details
A sample from the dataset:
{
"url": "https://www.ttcho.com/_files/ugd/988b76_01ceeff230b24cbbb0125b2bfa3f3475.pdf",
"filename": "988b76_01ceeff230b24cbbb0125b2bfa3f3475.pdf",
"success": true,
"page_count": 37,
"page_number": 1,
"question_choices_solutions": "QUESTION: What is the identity of X in the reaction 14N + 1n \u2192… See the full description on the dataset page: https://huggingface.co/datasets/mlfoundations-dev/pdf_science_questions_verifiable_r1_traces__2_24_25.magpie-reasoning-v1-20k-math-verifiable-step-by-step-rationalenuminamath_verifiable_cleaned_wo_geo_mc_difficultymagpie-reasoning-v1-100k-math-verifiableraw_parsed_pdfs_for_verifiable_qavllm-verifiable-control-arena
vLLM Verifiable Main Tasks Dataset
Programmatically verifiable coding tasks generated from vLLM git commits for the ControlArena vLLM setting.
Dataset Description
This dataset contains 30 coding tasks automatically generated from git commits in the vLLM repository. Each task represents a real-world coding challenge derived from actual development work.
23 tasks (76%) are programmatically verifiable with pytest tests that can automatically validate task completion.… See the full description on the dataset page: https://huggingface.co/datasets/RoganInglis/vllm-verifiable-control-arena.verifiable-constraints-v2magpie-reasoning-v1-20k-thought-summary-math-verifiabletextbook-qa-nepali-verifiableverifiable-ai-provenance-bench
Verifiable AI Provenance Bench (TTTPS)
25 real timestamp-provenance receipts generated on 2026-08-04 by calling the
live KPP (Kenosian Protocol Platform) provenance API
(POST /v1/anchor, POST /v1/verify), which implements the TTTPS (Time-Token
Tamper-evident Provenance Seal) scheme. Each row is one real API round trip:
a content_hash was submitted to /v1/anchor, the returned receipt_id was then
submitted to /v1/verify, and both raw responses are recorded.
This dataset was built… See the full description on the dataset page: https://huggingface.co/datasets/Pittro/verifiable-ai-provenance-bench.magpie-reasoning-v1-20k-math-verifiable-verificationcleaned_NuminaMath-RL-Verifiable_with_proofteam-akiyama-short_cleaned_NuminaMath-RL-Verifiable_with_proofspai-ss6-corpus-medical-o1-verifiable
SPAI SS6 Medical O1 Verifiable Thai Index
Index repo for the imported Thai medical verifiable-problem dataset config.
This is a lightweight index dataset repo. It does not duplicate the full corpus.
The full Parquet data lives in the canonical repository config below.
Canonical Data
Canonical repo: SPAISS6F1/spai-ss6-llm-1b-thai-corpus
Canonical config: medical_o1_verifiable_problem_thai
Rows in canonical config: 40,906
Parquet size in canonical config: 0.00 GB… See the full description on the dataset page: https://huggingface.co/datasets/SPAISS6F1/spai-ss6-corpus-medical-o1-verifiable.short_cleaned_NuminaMath-RL-Verifiable_with_proofmagpie-reasoning-v1-20k-math-verifiable-verification-min-400rm-dataset-v0.0.1
Verifiable Labs RM dataset v0.0.1
Reward-model training data for the Verifiable Labs SDK,
produced by the Phase 29 reward-distillation pipeline.
Stats
Rows: 840
With frontier judgment: 194
Frontier judge: anthropic/claude-sonnet-4 (when judged)
Source mix:
env: 646
judgment: 194
Schema
Each row is a JSON object with the following fields:
field
type
meaning
row_id
str
unique id
env_id
str
env that produced the row
prompt
str
task… See the full description on the dataset page: https://huggingface.co/datasets/verifiablelabs/rm-dataset-v0.0.1.verifiable-constraints-v6pdf_r1_annotation_verifiable_questionscleaned_NuminaMath-RL-Verifiablegrpo-3-harder-verifiabletest-problems-verifiable
Dataset Card for test-problems-verifiable
This dataset has been created with distilabel.
Dataset Summary
This dataset contains a pipeline.yaml which can be used to reproduce the pipeline that generated it in distilabel using the distilabel CLI:
distilabel pipeline run --config "https://huggingface.co/datasets/rntc/test-problems-verifiable/raw/main/pipeline.yaml"
or explore the configuration:
distilabel pipeline info --config… See the full description on the dataset page: https://huggingface.co/datasets/rntc/test-problems-verifiable.verifiable-constraints-1119verifiable-constraints-1126WebInstruct-Kmean-V0-Verifiable
