datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
bigbench_zero_shotZeroshot-Audio-Classification-Instructions
Zeroshot-Audio-Classification-Instructions
Convert audio classification dataset into zero-shot format speech instructions, support both single label and multi-label,
VGGSound
FSD50k
Nonspeech7k
urbansound8K
VocalSound
Emotion
Gender
ESD Emotion
Age
Language
TAU Urban Acoustic Scenes 2022
CochlScene
BirdCLEF_2021
EmoBox
AudioSet
We also converted huge WAV files into MP3 16k sample rate to reduce storage size.To prevent leakage, please do not include test set in training session.… See the full description on the dataset page: https://huggingface.co/datasets/mesolitica/Zeroshot-Audio-Classification-Instructions.PlantCAD2_zero_shot_tasks
🌱 PlantCAD2 Zero-Shot Tasks
Zero-shot evaluation tasks for plant genomics using PlantCAD2.This dataset contains tasks designed to evaluate model performance without task-specific training.
📂 Available Tasks
🔬 Cross-species Evolutionary Conservation
Task Name
Description
Samples
Metric
conservation_within_andropogoneae
Predict conserved vs non-conserved sites using alignments within 35 Andropogoneae genomes
19,030 vs 19,030
AUROC… See the full description on the dataset page: https://huggingface.co/datasets/plantcad/PlantCAD2_zero_shot_tasks.PlantCAD2_zero_shot_tasks
🌱 PlantCAD2 Zero-Shot Tasks
Zero-shot evaluation tasks for plant genomics using PlantCAD2.This dataset contains tasks designed to evaluate model performance without task-specific training.
📂 Available Tasks
🔬 Cross-species Evolutionary Conservation
Task Name
Description
Samples
Metric
conservation_within_andropogoneae
Predict conserved vs non-conserved sites using alignments within 35 Andropogoneae genomes
19,030 vs 19,030
AUROC… See the full description on the dataset page: https://huggingface.co/datasets/Yangximiao/PlantCAD2_zero_shot_tasks.synthetic_zeroshot_mixtral_v0.1zero-shot-label-nlitasksource classification tasks recasted as natural language inference.
This dataset is intended to improve label understanding in zero-shot classification HF pipelines.
Inputs that are text pairs are separated by a newline (\n).
from transformers import pipeline
classifier = pipeline(model="sileod/deberta-v3-base-tasksource-nli")
classifier(
"I have a problem with my iphone that needs to be resolved asap!!",
candidate_labels=["urgent", "not urgent", "phone", "tablet", "computer"],
)… See the full description on the dataset page: https://huggingface.co/datasets/tasksource/zero-shot-label-nli.medical-5day-zeroshotsf_fold
SF Fold
Real-world data for general robotics.
🤖 What is SF Fold?
This dataset contains real-world residential t-shirt folding demonstrations, collected by trained data collectors using hand-held grippers in diverse home environments.
📖 Table of Contents
Features
Terminology
Specifications
Dataset Composition
Environment Composition
Garment Composition
Trajectory Specifications
Hardware Specifications
Physical… See the full description on the dataset page: https://huggingface.co/datasets/zeroshotdata/sf_fold.zeroshot_test_downsampledautoeval-eval-autoevaluate__zero-shot-classification-sample-autoevalu-912bbb-1484454284
Dataset Card for AutoTrain Evaluator
This repository contains model predictions generated by AutoTrain for the following task and dataset:
Task: Zero-Shot Text Classification
Model: mathemakitten/opt-125m
Dataset: autoevaluate/zero-shot-classification-sample
Config: autoevaluate--zero-shot-classification-sample
Split: test
To run new evaluation jobs, visit Hugging Face's automatic model evaluator.
Contributions
Thanks to @mathemakitten for evaluating this model.
molmoact2-so101-zero-shot-eval
MolmoAct2 SO-101 Zero-Shot Evaluation Traces
This repository contains evaluation traces from running allenai/MolmoAct2-SO100_101 zero-shot on an SO-101 robot arm using the official LeRobot MolmoAct2 integration plus a remote async inference setup.
This is an evaluation artifact, not a training dataset or model checkpoint. The model under test is AllenAI's released MolmoAct2 SO-100/SO-101 checkpoint.
Summary
MolmoAct2 remote inference was successfully brought up on… See the full description on the dataset page: https://huggingface.co/datasets/abdul004/molmoact2-so101-zero-shot-eval.math-classification-zero-shot-resultsso100_test3This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "so100_bimanual",
"total_episodes": 6,
"total_frames": 9213,
"total_tasks": 1,
"total_videos": 12,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:6"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/ZeroShotLaz/so100_test3.Stack-Overflow-Zero-Shot-Classification
Dataset Card for "Stack-Overflow-Zero-Shot-Classification"
Automatic Stack Overflow Question Classifier
Important
All credit goes to huggingface user MoritzLaurer as his model is the basis for this project.
Introduction
The Automatic Stack Overflow Question Classifier harnesses the latest advancements in artificial intelligence to systematically categorize questions on Stack Overflow. Its primary goal is to streamline the process of sorting queries… See the full description on the dataset page: https://huggingface.co/datasets/amaye15/Stack-Overflow-Zero-Shot-Classification.so100_bimanualThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "so100_bimanual",
"total_episodes": 5,
"total_frames": 4025,
"total_tasks": 1,
"total_videos": 5,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:5"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/ZeroShotLaz/so100_bimanual.cartesia-sonic-preview-ztts1-zero-shot-sample
Cartesia Sonic on ZTTS1 zero-shot — sample with reference audio
100 utterances per language (700 rows) from the
zero-shot subsets of ZTTS1-Eval, synthesized with Cartesia Sonic (preview) in
voice-cloning mode.
Unlike the full set, every row carries the reference recording as well as the
synthesized clip, so a take can be compared against the voice it was cloning
without checking out the benchmark.
Columns
column
meaning
audio
the clip the model produced… See the full description on the dataset page: https://huggingface.co/datasets/jaeyong2/cartesia-sonic-preview-ztts1-zero-shot-sample.medical-3day-zeroshot-freshexps-test-no-contextmedical-4day-zeroshot-freshexps-test-no-contextzero_shotmedical-2day-zeroshot-freshexps-testso100_bimanual_coffee2This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "so100_bimanual",
"total_episodes": 6,
"total_frames": 15411,
"total_tasks": 1,
"total_videos": 6,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:6"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/ZeroShotLaz/so100_bimanual_coffee2.medical-1day-zeroshot-freshexps-test-no-contextzero_shot_open_llm_leaderboardmultilingual-zero-shot-label-nlimtasksource classification tasks recasted as natural language inference.
This dataset is intended to improve label understanding in zero-shot classification HF pipelines.
Inputs that are text pairs are separated by a newline (\n).
from transformers import pipeline
classifier = pipeline(model="sileod/mdeberta-v3-base-tasksource-nli")
classifier(
"I have a problem with my iphone that needs to be resolved asap!!",
candidate_labels=["urgent", "not urgent", "phone", "tablet", "computer"]… See the full description on the dataset page: https://huggingface.co/datasets/tasksource/multilingual-zero-shot-label-nli.rollout-pi05-zeroshot-stack-two-blocks-2026-02-19This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "franka",
"total_episodes": 5,
"total_frames": 2505,
"total_tasks": 1,
"chunks_size": 1000,
"data_files_size_in_mb": 100,
"video_files_size_in_mb": 200,
"fps": 15,
"splits": {
"train": "0:5"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/ankile/rollout-pi05-zeroshot-stack-two-blocks-2026-02-19.medical-1day-zeroshot-freshexps-test-no-context-llama3370binstructagain3PIPer-envbench-zeroshot-rlPaper | Code
zero-shot-emotions-8-4-1.25-85-65-75From the base score of 100, subtract 8 first for Rank 2, then subtract an additional 4 for Rank 3+. For Rank 4+, multiply the subtraction by 1.25, except for the emotions field. Next, divide by 100. Then, for each Rank, multiply this base score by Rank Number If the top original score is less than 65, then penalize all the scores by 75% of the original. If the top score is greater than 85, penalize those that are less than 65.
so100_bimanual_coffeeThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "so100_bimanual",
"total_episodes": 4,
"total_frames": 8908,
"total_tasks": 1,
"total_videos": 4,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:4"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/ZeroShotLaz/so100_bimanual_coffee.ActionRoutes_Phi2_ZeroShot
