datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
transformers_circleci_workflow_runsdataset_with_scriptThis is a test dataset.zenlibrispeech_asr_dummyaudiofolder_two_configs_in_metadataaudiofolder_single_config_in_metadatavideos-testProgramBench-Tests
ProgramBench Generated Tests
This dataset contains the AI-generated behavioral test suites used to evaluate model solutions in ProgramBench.
ProgramBench is a benchmark that evaluates whether language models can rebuild programs from scratch. Given only a compiled binary and its documentation, AI agents must architect and implement a complete codebase that reproduces the original program's behavior. These test suites are used to assess whether a candidate solution is behaviorally… See the full description on the dataset page: https://huggingface.co/datasets/programbench/ProgramBench-Tests.swe-bench-dummy-test-datasetdatasets-tests-compressionmulti_dir_datasetimagefolder_with_metadatadataset_with_data_filesaudiofolder_no_configs_in_metadatadiffusers-imagestiny-testtests-raw-jsonldocumentation-mediaDatasetWithCapitalLettersalpaca_2k_testtestraw_jsonlfixtures_image_utils\\nfixtures_ade20kIFBench_test
License
This dataset is licensed under ODC-BY-1.0. It is intended for research and educational use in accordance with Ai2's Responsible Use Guidelines. This dataset includes output data generated from third party models that are subject to separate terms governing their use.
Citation
Please cite:
@misc{pyatkin2025generalizing,
title={Generalizing Verifiable Instruction Following},
author={Valentina Pyatkin and Saumya Malik and Victoria Graf and Hamish Ivison and… See the full description on the dataset page: https://huggingface.co/datasets/allenai/IFBench_test.testing_alpaca_small
Dataset Card for "testing_alpaca_small"
More Information needed
lora_testtokenizers-benchwds_objectnet_test
ObjectNet (Test set only)
Original paper: ObjectNet: A large-scale bias-controlled dataset for pushing the limits of object recognition models
Homepage: https://objectnet.dev/
Bibtex:
@inproceedings{NEURIPS2019_97af07a1,
author = {Barbu, Andrei and Mayo, David and Alverio, Julian and Luo, William and Wang, Christopher and Gutfreund, Dan and Tenenbaum, Josh and Katz, Boris},
booktitle = {Advances in Neural Information Processing Systems},
editor = {H. Wallach and H. Larochelle and… See the full description on the dataset page: https://huggingface.co/datasets/djghosh/wds_objectnet_test.tokenizers-test-data
tokenizers-test-data
Test and benchmark fixtures for huggingface/tokenizers,
pulled on demand by the repo Makefiles (make test / make bench / make fixtures
via hf download).
Layout
fixtures/ — multilingual + modality corpora for cross-language encode
benchmarks. Organized, documented, and reproducible: see
fixtures/FIXTURES.md for provenance and
fixtures/fixtures_manifest.json for
exact sources, pinned revisions, and sizes. Rebuild any file with… See the full description on the dataset page: https://huggingface.co/datasets/hf-internal-testing/tokenizers-test-data.
