CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Goku-OpenLab /nano-banana-pro-prompts-datasets 🖼️ Nano Banana Pro Prompt Dataset 🖼️ The ultimate Nano Banana Pro prompt dataset (6GB+). 26,000+ image generation prompts with full metadata and preview images. Truly open source: No login, no ads, no redirection. Just pure data for AI image creators. This project is a massive collection of prompts used for Nano Banana Pro AI image model and the resulting generated images. The entire dataset exceeds 6GB and contains 26,000+ images, all structured into a comprehensive… See the full description on the dataset page: https://huggingface.co/datasets/Goku-OpenLab/nano-banana-pro-prompts-datasets.imagetext-to-image10K<n<100K1 likes19k downloads2mo agoHugging Face02JoyboyBrian /nano-omni-vlmimage1M<n<10M0 likes1.7k downloads2y agoHugging Face03nanonets /idp-leaderboard-resultsimage1K<n<10K0 likes913 downloads6mo agoHugging Face04emozilla /dolma-v1_7-305B-tokenized-llama2-nanosetimagen<1K0 likes852 downloads2y agoHugging Face05medarc /nanopathimage1K<n<10K0 likes793 downloads4mo agoHugging Face06bitmind /nano-banana Nano-Banana Generated Images 9,457 high-quality images generated using the Nano-Banana model (Google Gemini 2.5 Flash Image Preview). Dataset Overview Total Images: 9,457 images Generation Method: Nano-Banana (Google Gemini 2.5 Flash Image Preview) Storage Format: Optimized binary (Hugging Face Image type) File Organization: Normal large parquet files (not chunked) License: MIT Schema Column Type Description id int Unique identifier image Image… See the full description on the dataset page: https://huggingface.co/datasets/bitmind/nano-banana.imagetext-to-image1K<n<10K21 likes598 downloads1y agoHugging Face07mateoguaman /vamos_25pct_gpt5_nano vamos_25pct_gpt5_nano Description VLN Navigation dataset with 100% of tartandrive data, 50% of scand data, 25% of coda data, 100% of in-domain spot data, and 25% of annotated/augmented data using gpt5-nano. Whenever daatsets aren't 100%, they are ranked by curvature and output of length 5. Processing Parameters {} Dataset Configuration Train dataset: mixer: mateoguaman/vlmn_tartandrive100_scand50_coda25_spot100_sub5: 1.0… See the full description on the dataset page: https://huggingface.co/datasets/mateoguaman/vamos_25pct_gpt5_nano.image1M<n<10M0 likes479 downloads1y agoHugging Face08ArchaeonSeq /nanochat nanochat nanochat is the simplest experimental harness for training LLMs. It is designed to run on a single GPU node, the code is minimal/hackable, and it covers all major LLM stages including tokenization, pretraining, finetuning, evaluation, inference, and a chat UI. For example, you can train your own GPT-2 capability LLM (which cost $43,000 to train in 2019) for only $48 (2 hours of 8XH100 GPU node) and then talk to it in a familiar ChatGPT-like web UI. On a spot instance… See the full description on the dataset page: https://huggingface.co/datasets/ArchaeonSeq/nanochat.imagen<1K0 likes476 downloads23d agoHugging Face09Nilaksh404 /gpt-5-nanodocumentn<1K0 likes427 downloads10mo agoHugging Face10nanovdr /NanoVDR-Train NanoVDR-Train: Multilingual Visual Document Retrieval Training Data Training dataset for NanoVDR, comprising 1.49M query–image pairs across 6 languages for visual document retrieval. Paper: Our arxiv preprint is currently on hold. Details on training methodology, ablations, and full results will be available once the paper is published. Dataset Summary Statistic Value Total samples 1,489,252 (711K original + 778K augmented) Validation samples 14,518… See the full description on the dataset page: https://huggingface.co/datasets/nanovdr/NanoVDR-Train.imagefeature-extraction1M<n<10M0 likes365 downloads6mo agoHugging Face11Nanopocket-ai /FFHQ-2048 FFHQ-2048 — NanoPocket Enhanced (First 1,000) The first new high-quality public face dataset since 2019. 1,000 sharp, artifact-free 2048×2048 portraits, derived from FFHQ and enhanced with the NanoPocket Face Enhance model. Why this dataset exists FFHQ (NVIDIA, 2019) has been the gold-standard face dataset for the past five years — but the field has moved on. Modern generators (Flux, SD3 / SDXL, StyleGAN-T, portrait restoration nets) train at 1024²… See the full description on the dataset page: https://huggingface.co/datasets/Nanopocket-ai/FFHQ-2048.imageimage-to-image1K<n<10K0 likes342 downloads5mo agoHugging Face12nanochat-students /imagesimagen<1K1 likes257 downloads1y agoHugging Face13bitmind /Nano-banana-150kNano-consistent-150k. — the first dataset constructed using Nano-Banana that exceeds 150k high-quality samples, uniquely designed to preserve consistent human identity across diverse and complex editing scenarios image100K<n<1M9 likes217 downloads1y agoHugging Face14Abd0r /nanog-cancer-data NanoG - Cancer Foundation-Model Training Data Multimodal cancer corpus for NanoG1 (generative multimodal pretraining). Literature, structured biology, imaging, and grounded <simulate> traces. Hub: Abd0r/nanog-cancer-dataAuthor: Syed Abdur Rehman Ali (@Abd0r) · 17 · independent Train exclusion (hard): NCI-60 is out of training. Skip records whose source / path / text refer to NCI-60. Prefer NCI-ALMANAC, TCGA-sim, Polymathic, PMC/PubMed, TCGA omics, imaging. How… See the full description on the dataset page: https://huggingface.co/datasets/Abd0r/nanog-cancer-data.imagetext-generation10M<n<100M1 likes209 downloads2mo agoHugging Face15medarc /nanopath-evals NanoPath evaluation data This is the immutable data mirror used by NanoPath probe protocol v2. It contains only the exact development records consumed by medarc/nanopath: selected THUNDER training/validation images, prepared development-only slide caches, and the two PathoROB subsets. manifest.json records SHA-256 checksums and binds the snapshot to the checked-in benchmark manifests. No official THUNDER, HEST, or CPTAC classification test record is included. HEST is absent.… See the full description on the dataset page: https://huggingface.co/datasets/medarc/nanopath-evals.imageimage-classification10K<n<100K0 likes197 downloads20d agoHugging Face16qubvel-hf /ade20k-nanoimagen<1K0 likes186 downloads2y agoHugging Face17ash12321 /nano-banana-pro-generated-1k Nano Banana Pro (1K) Dataset 200 AI-generated images at 1K quality. License: MIT imagen<1K4 likes186 downloads9mo agoHugging Face18julienlucas /midjourney-dalle-sd-nanobananapro-dataset Dataset Card: Midjourney, DALL-E, Stable Diffusion & Nano Banana Pro vs Real Images Description Dataset de classification binaire pour détecter les images générées par IA (Midjourney, DALL-E, Stable Diffusion et Nano Banana Pro) vs images réelles. Dataset Structure Train set: 10,000 images Real: 5000 images Fake (AI-generated): 5000 images Test set: 2,000 images Real: 1000 images Fake (AI-generated): 1000 images Features { "image": Image… See the full description on the dataset page: https://huggingface.co/datasets/julienlucas/midjourney-dalle-sd-nanobananapro-dataset.imageimage-classification10K<n<100K6 likes183 downloads8mo agoHugging Face19Qwest /monuments_nanoimage1K<n<10K0 likes176 downloads2y agoHugging Face2034data /nano-receipts 🧾 Nano Receipts Dataset A diverse collection of 2428 hyper-realistic synthetic receipt images generated using state-of-the-art text-to-image AI models. 🚀 Quick Start from datasets import load_dataset # Load dataset (fast parquet format!) dataset = load_dataset("34data/nano-receipts") # Access images image = dataset["train"][0]["image"] # PIL Image filename = dataset["train"][0]["filename"] 📊 Dataset Details Total Images: 2428 receipts Format:… See the full description on the dataset page: https://huggingface.co/datasets/34data/nano-receipts.imageimage-to-text1K<n<10K0 likes170 downloads11mo agoHugging Face21exdysa /nano-banana-pro-generated-1k-clonename: nano-banana-pro-generated-1k-clone license: mit pipeline_tag: text-to-image tasks: - text-to-image - image-generation tags: - nano-banana language: en size_categories: - n<1K [!IMPORTANT] Original Model Link : https://huggingface.co/datasets/ash12321/nano-banana-pro-generated-1k-cloneimagen<1K2 likes125 downloads8mo agoHugging Face22FlameF0X /nano-banana-pro-gen-zh-enFlameF0X/nano-banana-pro-gen-zh-en is a translation of kaupane/nano-banana-pro-gen from Chinese to English imagetext-to-image1K<n<10K2 likes98 downloads4mo agoHugging Face23kaupane /nano-banana-pro-genimage1K<n<10K11 likes97 downloads7mo agoHugging Face24Deathspike /magical-girl-lyrical-nanoha-official-art-verimagen<1K0 likes83 downloads3y agoHugging Face25samarth010 /nano-receipts 🧾 Nano Receipts Dataset A diverse collection of 2428 hyper-realistic synthetic receipt images generated using state-of-the-art text-to-image AI models. 🚀 Quick Start from datasets import load_dataset # Load dataset (fast parquet format!) dataset = load_dataset("34data/nano-receipts") # Access images image = dataset["train"][0]["image"] # PIL Image filename = dataset["train"][0]["filename"] 📊 Dataset Details Total Images: 2428 receipts… See the full description on the dataset page: https://huggingface.co/datasets/samarth010/nano-receipts.imageimage-to-text1K<n<10K0 likes70 downloads29d agoHugging Face26lucasgblu /nanostep-datasetsimagen<1K0 likes63 downloads3y agoHugging Face27tomo2chin2 /nanoimage0 likes47 downloads11mo agoHugging Face28FlameF0X /NanoSRimage1K<n<10K1 likes45 downloads5mo agoHugging Face29Bofeee5675 /GUI-Net-NanoFor debugging purpose of training TongUI image1K<n<10K2 likes41 downloads1y agoHugging Face30thevatsalsaglani /coco_captions_nanoimage1K<n<10K0 likes39 downloads3y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.