datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
StreetView-Image-Dataset-10K
Urban Streetscape Dataset for Vision Language Models
A curated subset of 10,000 street view images with 25 essential features optimized for training vision language models on urban environment analysis tasks.
Dataset Description
This dataset contains street view imagery paired with comprehensive annotations covering infrastructure characteristics, visual perception metrics, environmental context, and semantic segmentation data.
This comprehensive dataset represents a… See the full description on the dataset page: https://huggingface.co/datasets/Sadhana-24/StreetView-Image-Dataset-10K.LogoBrief-10K
LogoBrief-10K
10,000 brand logos, each with its original SVG and a clean raster render, plus design annotations: open-vocabulary style tags, a one-sentence motif description, a full generated design brief, and text-region bounding boxes. Domain sampling is stratified by web-popularity rank rather than selected for recognizable brands. Every included domain was checked for AI-training opt-out signals immediately before publication (see Opt-out audit evidence).
Video… See the full description on the dataset page: https://huggingface.co/datasets/Logolabs/LogoBrief-10K.sdxl-generated-10k
SDXL Generated Images Dataset (10,000 images)
This dataset contains 10,000 AI-generated images created with Stable Diffusion XL for training an AI image detector.
Dataset Details
Model: Stable Diffusion XL Base 1.0
Total Images: 10,000
Resolution: 1024×1024 pixels
Format: JPEG (quality 95)
Inference Steps: 10
Guidance Scale: 7.0
Random Seeds: Unique per image for maximum diversity
Generation Date: 2025-12-30
Prompt Diversity
Images generated with diverse… See the full description on the dataset page: https://huggingface.co/datasets/ash12321/sdxl-generated-10k.
