datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
shitspotter
Dataset Card for ShitSpotter ("ScatSpotter")
ShitSpotter (or "ScatSpotter" in formal settings) is an open dataset of images containing dog feces.
This dataset contains full-resolution smartphone images of dog feces ("poop") collected in urban outdoor
environments taken using a "before/after/negative" protocol.
It includes thousands of polygon annotations of feces in varied lighting, seasonal, and terrain conditions.
The dataset is designed for training and evaluating object… See the full description on the dataset page: https://huggingface.co/datasets/erotemic/shitspotter.ConsistCompose3Millustrations_for_children
数据集 README
数据集概述
欢迎使用我们的数据集,该数据集主要包含网络收集的儿童插画(儿插)。这些插画旨在为教育和研究目的提供丰富的视觉素材。我们鼓励用户在遵守本README中规定的条款和条件的前提下,充分利用这些资源进行学习和研究。
许可协议
本数据集遵循Creative Commons Attribution-NonCommercial-ShareAlike 3.0 (CC BY-NC-SA 3.0)许可协议。这意味着您可以:
自由分享:复制和分发数据集中的材料。
自由改编:基于本数据集的材料进行修改和再创作。
但请注意以下限制:
非商业性:您不得将本数据集用于商业目的。
相同方式共享:如果您对数据集进行了修改或衍生,您必须以相同的许可协议分发您的作品。
署名:您必须给出适当的署名,提供许可协议链接,并说明是否进行了更改。您可以以任何合理的方式进行署名,但不得以任何方式暗示许可人认可您或您的使用。
使用限制… See the full description on the dataset page: https://huggingface.co/datasets/shiertier/illustrations_for_children.animetimm-Danbooru-VLMDiff-training-testpixiv_tmpshitspotter
Dataset Card for ShitSpotter ("ScatSpotter")
ShitSpotter (or "ScatSpotter" in formal settings) is an open dataset of images containing dog feces.
This dataset contains full-resolution smartphone images of dog feces ("poop") collected in urban outdoor
environments taken using a "before/after/negative" protocol.
It includes thousands of polygon annotations of feces in varied lighting, seasonal, and terrain conditions.
The dataset is designed for training and evaluating object… See the full description on the dataset page: https://huggingface.co/datasets/rever77/shitspotter.illustrations_for_children_1024
数据集 README
数据集概述
欢迎使用我们的数据集,该数据集主要包含网络收集的儿童插画(儿插)。这些插画旨在为教育和研究目的提供丰富的视觉素材。我们鼓励用户在遵守本README中规定的条款和条件的前提下,充分利用这些资源进行学习和研究。
图像被处理为接近1024**2的分辨率大小(不放大)。
许可协议
本数据集遵循Creative Commons Attribution-NonCommercial-ShareAlike 3.0 (CC BY-NC-SA 3.0)许可协议。这意味着您可以:
自由分享:复制和分发数据集中的材料。
自由改编:基于本数据集的材料进行修改和再创作。
但请注意以下限制:
非商业性:您不得将本数据集用于商业目的。
相同方式共享:如果您对数据集进行了修改或衍生,您必须以相同的许可协议分发您的作品。
署名:您必须给出适当的署名,提供许可协议链接,并说明是否进行了更改。您可以以任何合理的方式进行署名,但不得以任何方式暗示许可人认可您或您的使用。
使用限制… See the full description on the dataset page: https://huggingface.co/datasets/shiertier/illustrations_for_children_1024.ShieldBreaker_Benchmark_Dataset
ShieldBreaker Benchmark Dataset
Overview
The ShieldBreaker Benchmark Dataset is a comprehensive collection of anti-CRISPR protein sequences and structures, designed for machine learning research in CRISPR-Cas system inhibition. This dataset contains both positive (anti-CRISPR) and negative (non-anti-CRISPR) samples with dual-modal data representations.
Dataset Structure
ShieldBreaker_Upload/
├── positive/
│ ├── fasta/
│ │ └──… See the full description on the dataset page: https://huggingface.co/datasets/Jumbol/ShieldBreaker_Benchmark_Dataset.danbooru_1024XiaoDing.Ziprlbench
RLBench Dataset
RLBench 18 tasks following the PerAct protocol, with image resolution 128×128.
100 demonstrations per task for training.
25 demonstrations per task for testing.
stylenetaesthetic-100kThis repo host first 100k of anime aesthetic dataset, intended for post training T2I models for anime image generation. As aesthetic is not well defined I decided to take liberty with it and included images that are considered "hard" to text-to-image generating system.
If you see no dataset files then I haven't done with it yet
fractal-dex12Twatermarktext-to-image-cleanedshigureui
