CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01TIGER-Lab /OmniEdit-Filtered-1.2M OmniEdit In this paper, we present OMNI-EDIT, which is an omnipotent editor to handle seven different image editing tasks with any aspect ratio seamlessly. Our contribution is in four folds: (1) OMNI-EDIT is trained by utilizing the supervision from seven different specialist models to ensure task coverage. (2) we utilize importance sampling based on the scores provided by large multimodal models (like GPT-4o) instead of CLIP-score to improve the data quality. 📃Paper | 🌐Website |… See the full description on the dataset page: https://huggingface.co/datasets/TIGER-Lab/OmniEdit-Filtered-1.2M.image1M<n<10M132 likes63k downloads2y agoHugging Face02timbrooks /instructpix2pix-clip-filtered Dataset Card for InstructPix2Pix CLIP-filtered Dataset Summary The dataset can be used to train models to follow edit instructions. Edit instructions are available in the edit_prompt. original_image can be used with the edit_prompt and edited_image denotes the image after applying the edit_prompt on the original_image. Refer to the GitHub repository to know more about how this dataset can be used to train a model that can follow instructions. Supported Tasks… See the full description on the dataset page: https://huggingface.co/datasets/timbrooks/instructpix2pix-clip-filtered.image100K<n<1M48 likes7.6k downloads4y agoHugging Face03laion /filtered-wit Filtered WIT, an Image-Text Dataset. A reliable Dataset to run Image-Text models. You can find WIT, Wikipedia Image Text Dataset, here Data was taken from dalle-mini/wit Author Aarush Katta Data Structure The data is stored as tars, containing 10,000 samples per tar. The parquets contain the metadata of each tar, which was crated using this script Each tar contains a .jpg, .txt, and .json. The image is stored in .jpg, the caption in .txt. and the metadata in… See the full description on the dataset page: https://huggingface.co/datasets/laion/filtered-wit.image1M<n<10M11 likes6.6k downloads5y agoHugging Face04yxchng /laion_synthetic_filtered_large_part3image10M<n<100M0 likes3.2k downloads3y agoHugging Face05yxchng /laion_synthetic_filtered_large_part1image10M<n<100M2 likes3.2k downloads3y agoHugging Face06yxchng /laion_synthetic_filtered_large_part2image10M<n<100M0 likes3k downloads3y agoHugging Face07AlekseyKorshuk /product-photography-v1-tiny-prompts-tasks-collage-filteredimage1K<n<10K1 likes2.6k downloads3y agoHugging Face08Nikkozhang /puzzle-hle-filteredimagen<1K0 likes2.3k downloads1y agoHugging Face09yxchng /laion_synthetic_filtered_large_part4image10M<n<100M0 likes1.6k downloads3y agoHugging Face10williamium /ULVR-filtered ULVR-filtered Filtered subset of RuoliuYang/ULVR_v2_clean: the 101,951 training samples that Qwen2.5-VL-7B-Instruct answered incorrectly given only input_image, but correctly once the intermediate_image_* were also provided (judged by Qwen3-VL-32B-Instruct). Same schema / subsets / train-split structure as the source. subset rows scene_graph 3522 edge 1394 depth 537 segmentation 1328 bbox_highlight 15186 bbox_crop 15260 text_cot 27158 helper_interleaved… See the full description on the dataset page: https://huggingface.co/datasets/williamium/ULVR-filtered.imagevisual-question-answering100K<n<1M1 likes955 downloads3mo agoHugging Face11v1v1d /pdf_images_filtered Image Dataset with Parquet Format This dataset contains images with their IDs in parquet format for efficient loading. Dataset Structure Each configuration (language) contains: image: PIL Image object id: String identifier Languages Arabic (ar): 955 images Bengali (bn): 932 images German (de): 940 images English (en): 932 images Spanish (es): 951 images French (fr): 947 images Gujarati (gu): 949 images Hindi (hi): 891 images Italian (it): 1,005 images… See the full description on the dataset page: https://huggingface.co/datasets/v1v1d/pdf_images_filtered.image10K<n<100K0 likes776 downloads7mo agoHugging Face12yxchng /ccs_synthetic_filtered_largeimage10M<n<100M0 likes691 downloads3y agoHugging Face13ShinoharaHare /Danbooru-2024-Filtered-1Mimageimage-classification100K<n<1M7 likes616 downloads10mo agoHugging Face14darknoon /svg-stack-filtered Dataset Card for svg-stack-filtered This is an attempt to replicate the dataset used for SFT in the paper Rendering-Aware Reinforcement Learning for Vector Graphics Generation Processed: Optimized with svgo precision=2 Rasterized with cairosvg[^cairo] [^cairo] cairosvg doesn't implement all svg features, but matches how the original paper Filtered based on some heuristics: Removed any svg that couldn't be rendered with cairosvg (~30%) Removed solid-color images Removed some… See the full description on the dataset page: https://huggingface.co/datasets/darknoon/svg-stack-filtered.image1M<n<10M3 likes602 downloads1y agoHugging Face15diffusers /instructpix2pix-clip-filtered-upscaledimage10K<n<100K1 likes529 downloads3y agoHugging Face16Obscure-Entropy /CONCEPTUAL_CAPTIONS_HU_FILTEREDimage1M<n<10M0 likes520 downloads2y agoHugging Face17Jiwon-Kang /pixmo-points-filtered_0-20_imgContainedimage100K<n<1M0 likes510 downloads9mo agoHugging Face18guyue-wa /instructpix2pix-clip-filtered Dataset Card for InstructPix2Pix CLIP-filtered Dataset Summary The dataset can be used to train models to follow edit instructions. Edit instructions are available in the edit_prompt. original_image can be used with the edit_prompt and edited_image denotes the image after applying the edit_prompt on the original_image. Refer to the GitHub repository to know more about how this dataset can be used to train a model that can follow instructions. Supported Tasks… See the full description on the dataset page: https://huggingface.co/datasets/guyue-wa/instructpix2pix-clip-filtered.image100K<n<1M0 likes481 downloads8mo agoHugging Face19weikaih /ai2thor-random-views-20k-3obj-filteredimage1K<n<10K0 likes467 downloads1y agoHugging Face20theblackcat102 /amazon-all-beauty-filtered-limitedimage100K<n<1M0 likes394 downloads4mo agoHugging Face21continuallearning /real_0_put_bowl_filtered_raw_frames real_0_put_bowl_filtered Task: "put the bowl on the plate" Type: training (filtered) Robot: Franka FR3 Cameras: observation.images.primary, observation.images.wrist (image, 256x256) @ 15 FPS Statistics Metric Value Episodes 53 Total frames 16675 Avg frames/episode 314 FPS 15 Format LeRobot v3.0 Features Feature Type Shape observation.images.primary image [256, 256, 3] observation.images.wrist image [256, 256, 3]… See the full description on the dataset page: https://huggingface.co/datasets/continuallearning/real_0_put_bowl_filtered_raw_frames.imagerobotics10K<n<100K0 likes379 downloads6mo agoHugging Face22ohjoonhee /of_filtered_splitimage100K<n<1M0 likes369 downloads10mo agoHugging Face23SaiCharithaAkula21 /benchmark_coco_filteredCOCO Benchmark Dataset Description and Metadata imagen<1K1 likes329 downloads2y agoHugging Face24LuciusLan /MPM_train_filteredimage100K<n<1M0 likes329 downloads6mo agoHugging Face25andersonbcdefg /osatlas-fineweb-images-filtered-1image10K<n<100K0 likes320 downloads1y agoHugging Face26nielsr /datacomp-small-filtered Dataset Card for "datacomp-small-filtered" This is the DataComp-small dataset with CLIP-large-patch14 image embeddings added, as well as: captions filtered for English using a FastText model captions filtered to have at least complexity of 1 image1M<n<10M1 likes314 downloads3y agoHugging Face27raulsteleac /rl_expert_franka_dataset_for_pi0_1M_filtered_for_pickupsimage1M<n<10M0 likes312 downloads1y agoHugging Face28weikaih /ai2thor-perspective-qa-800-to-400-human-filtered-v2imagen<1K0 likes293 downloads9mo agoHugging Face29ShuhongZheng /UNO1m-filtered-splitimage100K<n<1M0 likes282 downloads1y agoHugging Face30ljnlonoljpiljm /datacap-recomp-1M-download-filteredimage100K<n<1M0 likes272 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.