datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
gspc-hub-cards
GSPC hub cards — mill cards, not board axes
SWIFT census (live): https://councilof.ai/api/swift
XRPL reader (live): https://councilof.ai/api/xrpl
One row per signed measurement card: one model, one axis, one date, Ed25519 over the body. A row is MEASURED only when a signed card verifies. Absent (model, axis) pairs are absent — not zero.
Measurement, not certification. Cards are evidence of bytes on a frozen bank at a time — never approval, rating, or safety guarantee.… See the full description on the dataset page: https://huggingface.co/datasets/csoai/gspc-hub-cards.hub-queue
Hub queue — the census, never a measured count
SWIFT census (live): https://councilof.ai/api/swift
XRPL reader (live): https://councilof.ai/api/xrpl
One flat table (queue.parquet / queue.jsonl): rank, id, downloads, pipeline_tag, status, card_id, as_of, measured_axes. SUMMARY.json is the only place counts live (n, n_measured, n_measured_axes) — read them there, never type them here.
Most rows: UNMEASURED, empty card_id.
A few rows may carry a verifying card sha256 for models… See the full description on the dataset page: https://huggingface.co/datasets/csoai/hub-queue.vocab-bloom-hub-en
Vocab Bloom Hub — English
A structured English lexical dataset with translations into Russian, Spanish, French, German, Portuguese, Chinese and Arabic, maintained by the Vocab Bloom Hub project — documentation, the API reference and a playground at vocab-bloom-hub.com.
Every entry carries IPA transcription, a CEFR level, one or more sense-level definitions with usage examples, synonym and antonym links per sense, translations per sense in seven languages, and inflected forms —… See the full description on the dataset page: https://huggingface.co/datasets/Fristail27/vocab-bloom-hub-en.PolyGuardPolyGuard: Massive Multi-Domain Safety Policy-Grounded Guardrail Dataset
hf-hub-session-pi-traces
dacorvo/hf-hub-session-pi-traces
pi coding-agent session traces produced by
agentcap runs. Each run
contributes one folder under data/<run_id>/; inside, one file per
session in pi's native export format.
The on-the-wire HTTP captures for these same runs live in
dacorvo/hf-hub-session-captures.
Both belong to the
hf-hub-session Collection
— join on run_id to align captures with traces.
slack-bench
Slack Bench
Slack Bench contains 37 tasks for evaluating AI agents on Slack API operations.
Task Categories
Basic messaging: Send messages, DMs, group conversations
Channel operations: Create channels, invite users, archive
Rich text formatting: Block Kit (bold, italic, code blocks, lists, tables)
Thread replies: Reply to specific messages in threads
Cross-channel operations: Copy/summarize content between channels
User management: Add/remove users from channels… See the full description on the dataset page: https://huggingface.co/datasets/hubertmarek/slack-bench.jsonl-mls-hubert_large_ll60k-layer_22hubei_Enrollment_Information
Dataset Card for Dataset Name
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Dataset Details
Dataset Description
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Language(s) (NLP): [简体中文]
License: [More Information Needed]
Dataset Sources [optional]
Repository: [More Information Needed]
Paper… See the full description on the dataset page: https://huggingface.co/datasets/xzitao/hubei_Enrollment_Information.pentabrid-reproducibility
Pentabrid 27B: reproducibility package
Everything required to recompute the results of a controlled evaluation of fine-tuning
configurations for medical question answering. Openly available with no access
restrictions.
Contents
Path
Description
per_item/medxpertqa_*.jsonl
Per-item predictions for all six checkpoints on 2,450 MedXpertQA-Text items. Fields: id, gold, extracted_answer, correct, explicit_marker_present, n_markers, response_chars… See the full description on the dataset page: https://huggingface.co/datasets/Clinical-Reasoning-Hub/pentabrid-reproducibility.Gene-Embedding-Hub-contributions-stagingcuda-nsys-training
Qwythos Nsight Systems Profiling Agent Dataset
Multi-turn GPU profiling agent trajectories for fine-tuning Qwythos-9B (and similar tool-calling models) on NVIDIA Nsight Systems (nsys) + CUDA-L1 / KernelBench workloads.
Generated autonomously on an RTX 5090 by the model itself driving real profiling tools for ~33 hours.
Code: ai-hpc/prof-dataset-gen
Stats
Split
Rows
Notes
train
5,884
Accepted episodes (quality ≥ 0.55)
eval
309
5% holdout from accepted… See the full description on the dataset page: https://huggingface.co/datasets/gittensor-model-hub/cuda-nsys-training.research-program-hub
shimo4228 Research Program Hub — Knowledge Graph
JSON-LD federation index linking the five sibling research lines of the shimo4228 research program — three agent-design lines (Agent Knowledge Cycle, Contemplative Agent, Agent Attribution Practice) and two cross-cutting lines (Authorship Strategy, Attention Not Self) — together with their ecosystem repositories. The program's through-line is value-layer harness engineering: extending the engineered agent harness past tools… See the full description on the dataset page: https://huggingface.co/datasets/shimo4228/research-program-hub.huberman-lab-transcripts
Huberman Lab Transcript Dataset
Cleaned English transcripts from 438 videos published on the Huberman Lab YouTube channel.
Dataset
438 videos
9,833 transcript chunks
~114 million characters
JSONL format
Each record contains:
ext
ideo_id
itle
Processing
The transcripts were collected from YouTube captions and processed by normalizing whitespace, removing common caption artifacts, removing repeated words, splitting into coherent chunks, and… See the full description on the dataset page: https://huggingface.co/datasets/Noothi/huberman-lab-transcripts.giskard-hub-demo-retailnmr-retraction-hub
NMR Retraction Hub
Synthetic forward-DAG belief-revision episodes for evaluating policy-bounded
hub retraction. Each row is one complete policy variant of a shared base case.
The train.jsonl and test.jsonl files contain the evaluation splits;
variants.jsonl is the complete prototype collection and
policy_comparison.json is a compact review table.
Rules have one premise. Internal hub nodes have two incoming and two outgoing
rules, while internal simple nodes have one incoming and… See the full description on the dataset page: https://huggingface.co/datasets/leo-bjpark/nmr-retraction-hub.duckdb-nsql-scoresgiskard-hub-demo-healthcarereboot-hub-drone-industry-evidence-briefs
Reboot Hub Drone Industry Evidence Briefs
This repository preserves concise, source-bounded evidence briefs derived from Reboot Hub Industry Hotspot analysis. Each brief separates what public evidence establishes from what remains unknown, then turns that boundary into a practical decision framework.
This is not a manufacturer release feed, a copy of the source articles, or a product endorsement. Reboot Hub is an independent drone lifecycle and technical-reference company and is… See the full description on the dataset page: https://huggingface.co/datasets/Thomas0229/reboot-hub-drone-industry-evidence-briefs.persona-hub-toolhttps://github.com/tencent-ailab/persona-hub
sql-console-prompt
SQL Console Text 2 SQL Prompt
GitHub Gist
Feedback is welcome 🤗. This prompt was based on performance from Qwen on the DuckDB NSQL Benchmark, common dataset types and tasks typical for exploring HF Datasets.
This is the prompt used for the Text2SQL inside the SQL Console on Datasets.
Example Table Context
For the {table_context} we use the SQL DDL CREATE TABLE statement.
CREATE TABLE datasets (
"_id" VARCHAR,
"id" VARCHAR,
"author" VARCHAR… See the full description on the dataset page: https://huggingface.co/datasets/duckdb-nsql-hub/sql-console-prompt.docker-hub-scraper
Docker Hub Scraper · Images, Tags, Pulls & Publishers
Scrape Docker Hub container images, tags, pull counts, star counts, publishers, official vs community status, and categories.
Rows in this dataset
2,843
Fields
15
Collector runs behind it
50
Most recent observation
2026-08-04
Browsable presentation
https://reapx.dev/data/docker-hub-scraper/ — 2,805 entity pages
Run the collector yourself
https://apify.com/reapx/docker-hub-scraper
What… See the full description on the dataset page: https://huggingface.co/datasets/reapxdev/docker-hub-scraper.hub-stats-papers-last-7-days
hub-stats papers last 7 days
Snapshot extracted from cfahlgren1/hub-stats (papers subset) for the rolling last 7 days.
Source dataset: cfahlgren1/hub-stats
Window: NOW() - INTERVAL 7 DAY at extraction time
Rows: 109
Files:
papers_last_7_days.csv
papers_last_7_days.json
query.sql
Generated at: 2026-03-06 22:59:59 UTC
svlm-preprocessed-datasets-v2market-brief-data-hub
Market Brief Dataset Hub (ship with V3 model)
Companion data for Market Brief Adapter V3 (Mixtral-8x7B LoRA, win rate 0.7806)Adaption AutoScientist Challenge 2026 · Market Analysis & News
Story
Artifact
Role
Model V3
Ship — grounded briefs + richer TrueNorth-enriched train path + Mixtral
This hub
Full data program: pilot → synthetic → TrueNorth → merged → Adaptive Data exports
V3 training used Adaptive Data on the merged seed (8e47668b… /… See the full description on the dataset page: https://huggingface.co/datasets/kongclaves/market-brief-data-hub.persona-hub-mathhttps://github.com/tencent-ailab/persona-hub
duckdb-nsql-predictionstwinkle_hub_finetune_dataset
twinkle_hub_finetune_dataset
MCP tool-calling SFT 資料集,由 Agent Tools Fine-Tuning Platform 以「反向生成 + teacher solver 驗證」流程產生。
語言:繁體中文
工具(來自 MCP server):search_datasets, get_dataset, query_rows, materialize_dataset, search_patents, get_patent_body, search_exam, search_exam_questions, get_exam_paper, search_teacher_exam, search_teacher_exam_questions, get_teacher_exam_paper, search_teacher_recruit, search_teacher_recruit_questions, get_teacher_recruit_paper, search_taiwan_md… See the full description on the dataset page: https://huggingface.co/datasets/Simon-Liu/twinkle_hub_finetune_dataset.persona-hub-npc https://github.com/tencent-ailab/persona-hub
linear-bench-mini
Agent-Diff: Linear Bench Mini
This dataset is part of the Agent-Diff benchmark, presented in the paper Agent-Diff: Benchmarking LLM Agents on Enterprise API Tasks via Code Execution with State-Diff-Based Evaluation.
Website | GitHub | Paper
Context
The Linear Bench suite runs inside the Agent Diff isolation engine, with its own Postgres schema replaying the Linear GraphQL API. Agents interact via Linear's public surface area to satisfy CRUD-style tasks (create issues… See the full description on the dataset page: https://huggingface.co/datasets/hubertmarek/linear-bench-mini.hub24-financial-conversation-sample1
Dataset Card for Dataset Name
Dataset Summary
Financial conversation with the provided customer profile
Supported Tasks and Leaderboards
[More Information Needed]
Languages
[More Information Needed]
Dataset Structure
Data Instances
[More Information Needed]
Data Fields
[More Information Needed]
Data Splits
[More Information Needed]
Dataset Creation
Curation Rationale
[More Information… See the full description on the dataset page: https://huggingface.co/datasets/jsonfin17/hub24-financial-conversation-sample1.
