banana
Datasets
All datasets matching “banana”nano-banana-pro-prompts-datasets
🖼️ Nano Banana Pro Prompt Dataset
🖼️ The ultimate Nano Banana Pro prompt dataset (6GB+). 26,000+ image generation prompts with full metadata and preview images. Truly open source: No login, no ads, no redirection. Just pure data for AI image creators.
This project is a massive collection of prompts used for Nano Banana Pro AI image model and the resulting generated images. The entire dataset exceeds 6GB and contains 26,000+ images, all structured into a comprehensive… See the full description on the dataset page: https://huggingface.co/datasets/Goku-OpenLab/nano-banana-pro-prompts-datasets.open-vision-banana-snvc-train-full
SNVC-50M v5_full — Multi-Task Vision Dataset
Description
This dataset is a curated subset of the SenseNova Vision Corpus 50M (SNVC-50M), containing 43,509 samples across 6 vision task families and 31 source datasets. Each sample follows a conversational format with interleaved <image> tokens, designed for training vision-language models (VLMs).
Coverage: 43,509 / 57,878 (75.2%) of the original sampling plan. 23 datasets at 100%, 8 partial, 12 unrecoverable… See the full description on the dataset page: https://huggingface.co/datasets/gatilin/open-vision-banana-snvc-train-full.banana-vidorev3-fullpipe
Banana ViDoRe v3 Fullpipe
Nano Banana Pro full-pipeline synthetic training data for ViDoRe v3 finance and industrial domains.
This repository contains 4670 training records and 40190 unique referenced images across
domain-separated ColFlor/ColQwen training splits. Images are included in the repository and paths in each JSONL are
relative to that domain directory.
Generated at: 2026-06-29T09:39:01.644860+00:00
Layout
finance/train.jsonl
finance/metadata.json… See the full description on the dataset page: https://huggingface.co/datasets/vkehfdl1/banana-vidorev3-fullpipe.banana-vidorev3-synthetic-arms
Banana ViDoRe v3 Synthetic Arms
Domain-separated ViDoRe v3 synthetic training arms for finance and industrial adaptation.
The Hub dataset uses finance and industrial as dataset configs/subsets. Within each config, splits separate
vlm_in_batch, vlm_ocr_bm25, banana_fullpipe, and hybrid_vlm_ocr_bm25_banana_fullpipe.
Generated at: 2026-06-29T11:49:45.670912+00:00
Total JSONL rows across configs/splits: 151691.
Images are stored once per subset under… See the full description on the dataset page: https://huggingface.co/datasets/vkehfdl1/banana-vidorev3-synthetic-arms.Med-Banana-80K
Med-Banana-50K: A Cross-modality Large-Scale Dataset for Text-guided Medical Image Editing
Paper Link | GitHub Repository
Summary
Med-Banana-50K is a comprehensive 50K-image dataset for instruction-based medical image editing spanning three modalities (Chest X-ray, Brain MRI, Fundus Photography) and 23 disease types. The dataset includes bidirectional edits (lesion addition and removal) generated from real medical images using Gemini-2.5-Flash-Image.
What… See the full description on the dataset page: https://huggingface.co/datasets/RichardChenZH/Med-Banana-80K.Cobot_Magic_cut_banana
Cobot_Magic_cut_banana
📋 Overview
This dataset uses an extended format based on LeRobot and is fully compatible with LeRobot.
Robot Type: agilex_cobot_decoupled_magic
| Codebase Version: v2.1
End-Effector Type: two_finger_gripper
🏠 Scene Types
This dataset covers the following scene types:
home
restaurant
🤖 Atomic Actions
This dataset includes the following atomic actions:
grasp
pick
cut
place
📊 Dataset Statistics… See the full description on the dataset page: https://huggingface.co/datasets/RoboCOIN/Cobot_Magic_cut_banana.
