datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
pixel-art-character
Pixel Art Character Dataset
⚠️ CONTENT WARNING: This dataset contains partially NSFW content. Some images may include suggestive themes, violence, or mature content. Viewer discretion advised.
A dataset of 500 pixel art character sprites for training LoRA models.
License
Derived License: Apache 2.0
This dataset is provided under the Apache 2.0 License, inherited from the base models used for generation.
Copyright 2026 Limbicnation
Licensed under the Apache License… See the full description on the dataset page: https://huggingface.co/datasets/Limbicnation/pixel-art-character.limbic-eval-tool-use-mcp
Dataset Summary
The MCP Tool Call Evaluation Test Dataset is a synthetic dataset designed for evaluating and benchmarking language models' ability to correctly execute function calls in the context of Model Context Protocol (MCP) tools. This dataset contains 9,813 test examples that assess a model's proficiency in:
Tool Selection: Choosing the correct function from available tools
Parameter Structure: Providing all required parameters with correct names
Parameter Values: Supplying… See the full description on the dataset page: https://huggingface.co/datasets/quotientai/limbic-eval-tool-use-mcp.dual-stream-image-prompts
Dual-Stream Image Prompts
Multi-dialect image-prompt SFT dataset for training an LLM to route prompts to the
right diffusion model at inference time. Given a concept and a target_model, the
model learns to emit the correct prompt dialect (FLUX T5-XXL prose, SDXL dual-clip
tokens, a compact caption, or steering modifiers).
Routing lives in the instruction prefix, not in a nested output object — keeping the
LoRA's task simple and maximizing structural diversity for generalization.… See the full description on the dataset page: https://huggingface.co/datasets/Limbicnation/dual-stream-image-prompts.deforum-prompt-lora-dataset
De Forum Cinematic Prompt Dataset
A specialized dataset for fine-tuning language models to generate cinematic video diffusion prompts in the style of "The Deforum Art Film".
Description
This dataset contains instruction-response pairs for training models to generate high-quality video diffusion prompts with:
Cinematic language and film terminology
De Forum aesthetic (noir, minimalist, art film style)
Technical parameters (aspect ratio, guidance scale, seeds)
Camera… See the full description on the dataset page: https://huggingface.co/datasets/Limbicnation/deforum-prompt-lora-dataset.korean-upper-limbenglish_limbumdeforum-prompt-lora-dataset-v2Human-Decomposition-Limbs-Image-Embeddings
