datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
dual-stream-image-prompts
Dual-Stream Image Prompts
Multi-dialect image-prompt SFT dataset for training an LLM to route prompts to the
right diffusion model at inference time. Given a concept and a target_model, the
model learns to emit the correct prompt dialect (FLUX T5-XXL prose, SDXL dual-clip
tokens, a compact caption, or steering modifiers).
Routing lives in the instruction prefix, not in a nested output object — keeping the
LoRA's task simple and maximizing structural diversity for generalization.… See the full description on the dataset page: https://huggingface.co/datasets/Limbicnation/dual-stream-image-prompts.DualStream-Foundational-Manifests
Dual-Stream DeepFake Foundational Baseline Connectors
This repository provides standardized data connectors, download manifests, and partition splits for the 8 foundational baseline datasets used in the Dual-Stream Deepfake Detection Framework.
📊 Dual-Stream Model Allocation
🖼️ Model 1: General Vision & Signal Model (>578,000 samples)
NTIRE-RobustAIGenDetection (~120,000 samples): Multi-generator synthetic artifacts (MSU 2024).
CIFAKE (120,000… See the full description on the dataset page: https://huggingface.co/datasets/ThangCao/DualStream-Foundational-Manifests.
