datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
sample_clevrceo-quotes-verified-sample
🎙️ CEO Transcripts — Verified Executive Interviews
The World's Largest Database of Verified C-Suite Transcripts
20,000+ Executives · 100,000+ Transcripts · 400,000+ Quotes · S&P 500 + NASDAQ + Global Leaders
🔥 What's In This Sample?
This is a free evaluation sample from CEOInterviews.ai featuring 9 of the most market-moving voices in finance, tech, and policy.
Executive
Role
Why They Matter
Jensen Huang
CEO, NVIDIA
Every AI… See the full description on the dataset page: https://huggingface.co/datasets/codelucas/ceo-quotes-verified-sample.openspaces-depth-aware-32-samples
OpenSpaces Depth-Aware Visual QA Dataset
This is a 32-sample visual question answering (VQA) dataset that includes:
RGB images from the OpenSpaces dataset
Predicted depth maps generated using Depth Anything
3 depth-aware QA pairs per image:
Yes/No question (e.g., “Is there a person near the door?”)
Short answer question (e.g., “What color is the man’s coat?”)
Spatial sorting question (e.g., “Sort the objects from closest to farthest”)
Intended Use
This dataset is… See the full description on the dataset page: https://huggingface.co/datasets/srimoyee12/openspaces-depth-aware-32-samples.autolaparo-qna-sample
AutoLaparo QnA Sample (spec-aligned subtypes)
A small, spec-aligned QnA dataset for surgical-video VQA.
Each record follows the same schema as the production file output/qna_dataset.json,
and each question_subtype is one of the subtypes defined in the project's spec/*.md.
Built from AutoLaparo Task 1 (workflow phase recognition, video 14) and Task 3
(instrument and anatomy segmentation).
11 records / 11 subtypes — exactly one record per supported subtype to keep the
sample… See the full description on the dataset page: https://huggingface.co/datasets/phamdt2003/autolaparo-qna-sample.NULLSPACE-sample
NULLSPACE
267 JEE questions frontier models get wrong — every answer key independently verified, every row tagged for contamination. This is a 10-row sample. Access is auto-approved — fill in your details to view the data.
Full dataset (267 questions · 77 held out): Nalandadata/NULLSPACE — requires access request.
NULLSPACE is a failure-only benchmark. A question enters the set only when multiple frontier models independently get it wrong and the answer key has been… See the full description on the dataset page: https://huggingface.co/datasets/Nalandadata/NULLSPACE-sample.
