datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
getting-started-validation-clip-pred
Dataset Card for labeled_validation_predicted_clip
This is a FiftyOne dataset with 143 samples.
Installation
If you haven't already, install FiftyOne:
pip install -U fiftyone
Usage
import fiftyone as fo
from fiftyone.utils.huggingface import load_from_hub
# Load the dataset
# Note: other available arguments include 'max_samples', etc
dataset = load_from_hub("TheSteve0/getting-started-validation-clip-pred")
# Launch the App
session =… See the full description on the dataset page: https://huggingface.co/datasets/Voxel51/getting-started-validation-clip-pred.military-labeled-clip
Military-Labeled CLIP Crops (DVIDS sourced)
Per-object crops extracted from DVIDS military imagery, each accompanied by a
Gemini-VLM caption suitable for CLIP fine-tuning or zero-shot evaluation.
Files
crops/dvids_image_{id}_{class}_{idx}.jpg — 4,844 cropped objects
captions.jsonl — per-crop metadata: {crop_path, image_id, class_name, class_id, bbox, caption, image_caption, branch, source, source_url}
Classes (12) — Distribution
Class
Crops… See the full description on the dataset page: https://huggingface.co/datasets/AMANMP0007/military-labeled-clip.clide_synthetic_datasets
CLIDE Synthetic Image Datasets
📄 Paper • 💻 Code • 🌐 Webpage • 🎥 Video
A collection of synthetic images generated by modern text-to-image models, organized by domain and generator.
The dataset is designed to support analysis and evaluation of generated-image detection methods under domain and generator shifts.
🗂️ Dataset Structure
The dataset contains two visual domains:
💥🚗 Damaged Cars
Synthetic images of damaged cars generated by multiple… See the full description on the dataset page: https://huggingface.co/datasets/Fujitsu-FRE/clide_synthetic_datasets.military-labeled-clip
Military-Labeled CLIP Crops (DVIDS sourced)
Per-object crops extracted from DVIDS military imagery, each accompanied by a
Gemini-VLM caption suitable for CLIP fine-tuning or zero-shot evaluation.
Files
crops/dvids_image_{id}_{class}_{idx}.jpg — 4,844 cropped objects
captions.jsonl — per-crop metadata: {crop_path, image_id, class_name, class_id, bbox, caption, image_caption, branch, source, source_url}
Classes (12) — Distribution
Class
Crops… See the full description on the dataset page: https://huggingface.co/datasets/p14ton/military-labeled-clip.military-labeled-clip
Military-Labeled CLIP Crops (DVIDS sourced)
Per-object crops extracted from DVIDS military imagery, each accompanied by a
Gemini-VLM caption suitable for CLIP fine-tuning or zero-shot evaluation.
Files
crops/dvids_image_{id}_{class}_{idx}.jpg — 4,844 cropped objects
captions.jsonl — per-crop metadata: {crop_path, image_id, class_name, class_id, bbox, caption, image_caption, branch, source, source_url}
Classes (12) — Distribution
Class
Crops… See the full description on the dataset page: https://huggingface.co/datasets/llama-farm/military-labeled-clip.movies_CLIP_ViT-L14
🎬 Movie Frame & Caption Dataset
📖 Introduction
This dataset was created from multiple movies across 10 genres, with approximately 3 movies per genre.From each movie, frames were extracted periodically, and AI-generated captions (BLIP) were assigned to each frame.A total of 93,813 frames were extracted.
This dataset can be used for tasks such as:
Video understanding
Multimodal learning (image + text)
Image captioning
Vision-language retrieval
📂 Data… See the full description on the dataset page: https://huggingface.co/datasets/thaotien/movies_CLIP_ViT-L14.urban-climate-green-infrastructure
Urban Climate & Green Infrastructure Visual Dataset
Rows: 18,107
Dataset Description
Urban Climate & Green Infrastructure Visual Dataset is a global wildlife image dataset and geospatial computer vision dataset focused on street-level imagery of city features that support lower-carbon and more resilient planning. The labels cover Street Tree, Bike Lane, Solar Panel, EV Charger, Rain Garden, and Green Roof, produced through Outerview's query-driven embedding… See the full description on the dataset page: https://huggingface.co/datasets/Outerview/urban-climate-green-infrastructure.clip-insect-sex-data
Gryllus bimaculatus Insect Sex Dataset
Image dataset for binary classification: male and female.
Species
Gryllus bimaculatus
Structure
augmented_data/
male/
female/
Labels
male
female
Source and curation
Images were captured by the dataset owner.
Augmented variants were generated from owner-captured source images for training.
Access and permission terms
This dataset is shared for viewing/research reference.
Reuse… See the full description on the dataset page: https://huggingface.co/datasets/yashm/clip-insect-sex-data.cliptrace-baseline-data
CLIPTrace 2026 Baseline Data
Participant data for the reproducible CLIPTrace 2026 baseline. The repository
mirrors the full-size Imagenette train and validation images used by the
baseline and preserves the original class-directory layout.
This repository is intended to be public with access gating. Before
downloading, users must accept the repository terms and the upstream image-use
conditions configured by the organizers on the Hugging Face settings page.
Dataset… See the full description on the dataset page: https://huggingface.co/datasets/cliptrace-2026/cliptrace-baseline-data.
