datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
ImageNet_TA_IA
Library: https://github.com/lucasdegeorge/T2I-ImageNet
How far can we go with ImageNet for Text-to-Image generation?
Lucas Degeorge, Arijit Ghosh, Nicolas Dufour, David Picard, Vicky Kalogeiton
This dataset has the captions used during the training of the models from the paper "How far can we go with ImageNet for Text-to-Image generation?"
The core idea is that text-to-image generation models typically rely on vast datasets, prioritizing quantity over quality. The usual… See the full description on the dataset page: https://huggingface.co/datasets/Lucasdegeorge/ImageNet_TA_IA.imagenet-12k-metadata
ImageNet-12k Split Metadata
Metadata files defining the splits for ImageNet-12k subset of fall11_whole.tar (2011 ImageNet full release) used in some timm models (see dataset building code in https://github.com/rwightman/imagenet-12k).
ImageNetVC
Dataset Card for ImageNetVC
Dataset Summary
This is the dataset for our paper "ImageNetVC: Zero-Shot Visual Commonsense Evaluation on 1000 ImageNet Categories".
Supported Tasks
Zero- and Few-shot visual commonsense evaluation.
Languages
English
Dataset Structure
Data Instances
4076
Data Fields
Color, shape, material, component, and others.
Data Splits
Dev set for few-shot demonstrations and the main set for… See the full description on the dataset page: https://huggingface.co/datasets/hemingkx/ImageNetVC.ImageNet-CJ
JPEG Re-encoding Confound Control Dataset
A controlled-experiment dataset that isolates one acknowledged-but-unmeasured confound in
ImageNet-C. Hendrycks & Dietterich (Benchmarking Neural Network Robustness to Common
Corruptions and Perturbations, ICLR 2019, arXiv:1903.12261)
save every corrupted image as a lightly compressed JPEG. The benchmark therefore never measures a
corruption c applied to an image x in isolation — it measures JPEG(c(x)). This dataset lets
you quantify how… See the full description on the dataset page: https://huggingface.co/datasets/atharvadagaonkar/ImageNet-CJ.imagenet-closeimagenet1k_classes
ImageNet-1k Individual Class Datasets Hub
This dataset serves as a central hub and index for the 1,000 individual class datasets derived from the original imagenet-1k dataset. Each class has been separated into its own repository for easy access and analysis.
This repository does not contain any images itself. Instead, it provides a class_mapping.csv file that maps each class ID and name to its corresponding dataset repository on the Hugging Face Hub.
How to Use
The… See the full description on the dataset page: https://huggingface.co/datasets/mlnomad/imagenet1k_classes.imagenet_safety_annotatedThis is a safety annotation set for ImageNet. It uses the LlavaGuard-13B model for annotating.
The annotations entail a safety category (image-category), an explanation (assessment), and a safety rating (decision). Furthermore, it contains the unique ImageNet id class_sampleId, i.e. n04542943_1754.
These annotations allow you to train your model on only safety-aligned data. Plus, you can define yourself what safety-aligned means, i.e. discard all images where decision=="Review Needed" or… See the full description on the dataset page: https://huggingface.co/datasets/AIML-TUDA/imagenet_safety_annotated.imagenet1k_captions_minigpt4
ImageNet1k Captions Generated with MiniGPT-4
MiniGPT-4 captions generated for ImageNet1k images. Can be used for training/finetuning diffusion models for image generation.
ImageNet1k: link
MiniGPT-4: link
imagenet-partially-captionedimagenet_hard_review_data_r3
