datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
autotrain-data-logo-identifier-v3-medium
AutoTrain Dataset for project: logo-identifier-v3-medium
Dataset Description
This dataset has been automatically processed by AutoTrain for project logo-identifier-v3-medium.
Languages
The BCP-47 code for the dataset's language is unk.
Dataset Structure
Data Instances
A sample from this dataset looks as follows:
[
{
"image": "<100x72 RGB PIL image>",
"target": 47
},
{
"image": "<100x63 RGB PIL image>",
"target": 82… See the full description on the dataset page: https://huggingface.co/datasets/fsuarez/autotrain-data-logo-identifier-v3-medium.SAVANT-CODALM-medium
SAVANT CODALM Medium Dataset
This dataset is part of the SAVANT framework described in the SAVANT paper, currently under peer review.
This repository is provided for peer-review purposes only. After the review process, the dataset will be made publicly available through the authors' main account.
Dataset Description
CODALM medium was created by combining automated framework evaluation with human validation. Starting with the full CODA dataset (9,640 images), we used… See the full description on the dataset page: https://huggingface.co/datasets/u94fmn391j/SAVANT-CODALM-medium.Nexora-vision-dataset-v2-medium
Nexora Vision Dataset v2 Medium
The Nexora Vision Dataset v2 Medium is a scalable, mixed-resolution image dataset designed for generative AI experimentation, diffusion model workflows, and computer vision research.
Developed and curated by ArkDevLabs / ArkAiLab (ADL).
Official Website: https://arkdevlabs.com
Dataset Summary
Nexora Vision Dataset v2 Medium contains 9,236 curated images packaged in both:
Raw image format
Optimized Parquet format
This release prioritizes:… See the full description on the dataset page: https://huggingface.co/datasets/ArkAiLab-Adl/Nexora-vision-dataset-v2-medium.sd3-medium-scm-corpus
Shamima/sd3-medium-scm-corpus
Synthetic image corpus generated with Stable Diffusion 3 medium for studying the
Stereotype Content Model (SCM) structure of text-to-image latent space.
Images: 6,600
Categories: 66 occupation/identity groups
Prompt template: "A portrait of a [group], high quality."
Generator: Stable Diffsuion 3 medium, DPM++ 2M Karras, 30 steps, CFG 7.0
Resolution: 512 x 512
Fields
field
description
image
RGB JPEG
category
Group/occupation… See the full description on the dataset page: https://huggingface.co/datasets/Shamima/sd3-medium-scm-corpus.
